review.tizen.org Git - platform/upstream/llvm.git/commit

projects / platform / upstream / llvm.git / commit

author	Tony <Tony.Tye@amd.com>
	Fri, 16 Oct 2020 07:09:38 +0000 (07:09 +0000)
committer	Tony <Tony.Tye@amd.com>
	Tue, 20 Oct 2020 22:55:12 +0000 (22:55 +0000)
commit	1bc7bfffdbabffcdb43cc2829c551c33aed57742
tree	9c721a505b7e063ccaccd60e79deeba106b2094f	tree \| snapshot
parent	1298252f80fec0cd77aabef5cb133e7b030852e4	commit \| diff

[AMDGPU] Optimize waitcnt insertion for flat memory operations

Change waitcnt insertion to check the memory operand tokens to see if
flat memory operations access VMEM in the same way it does to check if
accessing LDS. This avoids adding waitcnt for counters for address
spaces that are not accessed.

In addition, only generate the pessimistic waitcnt 0 if a flat memory
operation appears to access both VMEM and LDS.

This benefits flat memory operations that explicitly specify the
address space as GLOBAL or LOCAL.

Differential Revision: https://reviews.llvm.org/D89618

46 files changed:

llvm/lib/Target/AMDGPU/SIInsertWaitcnts.cpp		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/cvt_f32_ubyte.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/extractelement.i128.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/extractelement.i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/extractelement.i8.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/extractelement.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/fmed3.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/frem.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/insertelement.i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/insertelement.i8.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/lds-global-non-entry-func.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/llvm.amdgcn.atomic.dec.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/llvm.amdgcn.atomic.inc.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/llvm.amdgcn.div.fmas.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/llvm.amdgcn.div.scale.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/llvm.amdgcn.update.dpp.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/shl-ext-reduce.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/GlobalISel/zextload.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/bitreverse.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/copy-illegal-type.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/ctlz.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/cvt_f32_ubyte.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/fast-unaligned-load-store.global.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/fmax_legacy.f64.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/fmin_legacy.f64.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/frem.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/idot2.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/imm16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/insert_vector_elt.v2i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/llvm.amdgcn.cvt.pkrtz.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/llvm.amdgcn.image.sample.d16.dim.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/llvm.cos.f16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/llvm.sin.f16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/load-lo16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/lshr.v2i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/max.i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/saddo.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/shl.v2i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/shrink-add-sub-constant.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/sub.v2i16.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/trunc-combine.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/waitcnt-back-edge-loop.mir		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/waitcnt-looptest.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/waitcnt-vscnt.ll		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/waitcnt.mir		diff \| blob \| history
llvm/test/CodeGen/AMDGPU/widen-smrd-loads.ll		diff \| blob \| history

Domain: System / Toolchain;

RSS Atom