AMDGPU/SI: Enable lanemask tracking in misched Summary: This results in higher register usage, but should make it easier for the compiler to hide latency. This pass is a prerequisite for some more scheduler improvements, and I think the increase register usage with this patch is acceptable, because when combined with the scheduler improvements, the total register usage will decrease. shader-db stats: 2382 shaders in 478 tests Totals: SGPRS: 48672 -> 49088 (0.85 %) VGPRS: 34148 -> 34847 (2.05 %) Code Size: 1285816 -> 1289128 (0.26 %) bytes LDS: 28 -> 28 (0.00 %) blocks Scratch: 492544 -> 573440 (16.42 %) bytes per wave Max Waves: 6856 -> 6846 (-0.15 %) Wait states: 0 -> 0 (0.00 %) Depends on D18451 Reviewers: nhaehnle, arsenm Subscribers: arsenm, llvm-commits Differential Revision: http://reviews.llvm.org/D18452 llvm-svn: 264876

commit: 0bc954e3bc474383f87ab9e55ab1aa5ae996f9c0 [log] [tgz]
author: Tom Stellard <thomas.stellard@amd.com> Wed Mar 30 16:35:09 2016 +0000
committer: Tom Stellard <thomas.stellard@amd.com> Wed Mar 30 16:35:09 2016 +0000
tree: d428795eaee9170ce8fc563a9634e2568806147e
parent: f76123386a7867ff5fa63a55841668ac098e201e [diff] [blame]
diff --git a/llvm/test/CodeGen/AMDGPU/ds_read2st64.ll b/llvm/test/CodeGen/AMDGPU/ds_read2st64.ll
index c788561..3f3e40e 100644
--- a/llvm/test/CodeGen/AMDGPU/ds_read2st64.ll
+++ b/llvm/test/CodeGen/AMDGPU/ds_read2st64.ll

@@ -65,9 +65,9 @@
 
 ; SI-LABEL: @simple_read2st64_f32_over_max_offset
 ; SI-NOT: ds_read2st64_b32
-; SI: v_add_i32_e32 [[BIGADD:v[0-9]+]], vcc, 0x10000, {{v[0-9]+}}
-; SI: ds_read_b32 {{v[0-9]+}}, {{v[0-9]+}} offset:256
-; SI: ds_read_b32 {{v[0-9]+}}, [[BIGADD]]
+; SI-DAG: v_add_i32_e32 [[BIGADD:v[0-9]+]], vcc, 0x10000, {{v[0-9]+}}
+; SI-DAG: ds_read_b32 {{v[0-9]+}}, {{v[0-9]+}} offset:256
+; SI-DAG: ds_read_b32 {{v[0-9]+}}, [[BIGADD]]{{$}}
 ; SI: s_endpgm
 define void @simple_read2st64_f32_over_max_offset(float addrspace(1)* %out, float addrspace(3)* %lds) #0 {
   %x.i = tail call i32 @llvm.amdgcn.workitem.id.x() #1
commit	0bc954e3bc474383f87ab9e55ab1aa5ae996f9c0	[log] [tgz]
author	Tom Stellard <thomas.stellard@amd.com>	Wed Mar 30 16:35:09 2016 +0000
committer	Tom Stellard <thomas.stellard@amd.com>	Wed Mar 30 16:35:09 2016 +0000
tree	d428795eaee9170ce8fc563a9634e2568806147e
parent	f76123386a7867ff5fa63a55841668ac098e201e [diff] [blame]