[DAGCombiner] Fix miscompile bug in combineShiftOfShiftedLogic #89616

bjope · 2024-04-22T15:52:21Z

Ensure that the sum of the shift amounts does not overflow the
shift amount type when combining shifts in combineShiftOfShiftedLogic.

Solves a miscompile bug found when testing the C23 BitInt feature.

Targets like X86 that only use an i8 for shift amounts after
legalization seems to be extra susceptible for bugs like this as it
isn't legal to shift more than 255 steps.

llvmbot · 2024-04-22T15:52:56Z

@llvm/pr-subscribers-llvm-selectiondag

@llvm/pr-subscribers-backend-x86

Author: Björn Pettersson (bjope)

Changes

Ensure that the sum of the shift amounts does not overflow the
shift amount type when combining shifts in combineShiftOfShiftedLogic.

Solves a miscompile bug found when testing the C23 BitInt feature.

Targets like X86 that only use an i8 for shift amounts after
legalization seems to be extra susceptible for bugs like this as it
isn't legal to shift more than 255 steps.

Full diff: https://github.com/llvm/llvm-project/pull/89616.diff

2 Files Affected:

(modified) llvm/lib/CodeGen/SelectionDAG/DAGCombiner.cpp (+8-1)
(modified) llvm/test/CodeGen/X86/shift-combine.ll (+35)

diff --git a/llvm/lib/CodeGen/SelectionDAG/DAGCombiner.cpp b/llvm/lib/CodeGen/SelectionDAG/DAGCombiner.cpp
index e6e0a1fc7d8277..fd265b12d73ca4 100644
--- a/llvm/lib/CodeGen/SelectionDAG/DAGCombiner.cpp
+++ b/llvm/lib/CodeGen/SelectionDAG/DAGCombiner.cpp
@@ -9503,8 +9503,15 @@ static SDValue combineShiftOfShiftedLogic(SDNode *Shift, SelectionDAG &DAG) {
     if (ShiftAmtVal->getBitWidth() != C1Val.getBitWidth())
       return false;
 
+    // The fold is not valid if the sum of the shift values doesn't fit in the
+    // given shift amount type.
+    bool Overflow = false;
+    APInt NewShiftAmt = C1Val.uadd_ov(*ShiftAmtVal, Overflow);
+    if (Overflow)
+      return false;
+
     // The fold is not valid if the sum of the shift values exceeds bitwidth.
-    if ((*ShiftAmtVal + C1Val).uge(V.getScalarValueSizeInBits()))
+    if (NewShiftAmt.uge(V.getScalarValueSizeInBits()))
       return false;
 
     return true;
diff --git a/llvm/test/CodeGen/X86/shift-combine.ll b/llvm/test/CodeGen/X86/shift-combine.ll
index bb0fd9c68afbaf..ca3898d358e221 100644
--- a/llvm/test/CodeGen/X86/shift-combine.ll
+++ b/llvm/test/CodeGen/X86/shift-combine.ll
@@ -787,3 +787,38 @@ define <4 x i32> @or_tree_with_mismatching_shifts_vec_i32(<4 x i32> %a, <4 x i32
   %r = or <4 x i32> %or.ab, %or.cd
   ret <4 x i32> %r
 }
+
+; FIXME: Reproducer for a DAGCombiner::combineShiftOfShiftedLogic
+; bug. DAGCombiner need to check that the sum of the shift amounts fits in i8,
+; which is the legal type used to described X86 shift amounts. Verify that we
+; do not try to create a shift with 140+120 as shift amount, and verify that
+; the stored value do not depend on %a1.
+define void @combineShiftOfShiftedLogic(i128 %a1, i32 %a2, ptr %p) {
+; X86-LABEL: combineShiftOfShiftedLogic:
+; X86:       # %bb.0:
+; X86-NEXT:    movl {{[0-9]+}}(%esp), %eax
+; X86-NEXT:    movl {{[0-9]+}}(%esp), %ecx
+; X86-NEXT:    movl %eax, 20(%ecx)
+; X86-NEXT:    movl $0, 16(%ecx)
+; X86-NEXT:    movl $0, 12(%ecx)
+; X86-NEXT:    movl $0, 8(%ecx)
+; X86-NEXT:    movl $0, 4(%ecx)
+; X86-NEXT:    movl $0, (%ecx)
+; X86-NEXT:    retl
+;
+; X64-LABEL: combineShiftOfShiftedLogic:
+; X64:       # %bb.0:
+; X64-NEXT:    # kill: def $edx killed $edx def $rdx
+; X64-NEXT:    shlq $32, %rdx
+; X64-NEXT:    movq %rdx, 16(%rcx)
+; X64-NEXT:    movq $0, 8(%rcx)
+; X64-NEXT:    movq $0, (%rcx)
+; X64-NEXT:    retq
+  %zext1 = zext i128 %a1 to i192
+  %zext2 = zext i32 %a2 to i192
+  %shl = shl i192 %zext1, 130
+  %or = or i192 %shl, %zext2
+  %res = shl i192 %or, 160
+  store i192 %res, ptr %p, align 8
+  ret void
+}

nikic · 2024-04-23T06:21:42Z

llvm/test/CodeGen/X86/shift-combine.ll

@@ -787,3 +787,38 @@ define <4 x i32> @or_tree_with_mismatching_shifts_vec_i32(<4 x i32> %a, <4 x i32
  %r = or <4 x i32> %or.ab, %or.cd
  ret <4 x i32> %r
 }
+
+; FIXME: Reproducer for a DAGCombiner::combineShiftOfShiftedLogic


Leftover FIXME

Thanks! (I actually realized that I had forgotten to remove this when I woke up this morning, but was not quick enough to fix it before you found it.)

jayfoad · 2024-04-23T08:35:10Z

llvm/test/CodeGen/X86/shift-combine.ll

-; FIXME: Reproducer for a DAGCombiner::combineShiftOfShiftedLogic
-; bug. DAGCombiner need to check that the sum of the shift amounts fits in i8,
-; which is the legal type used to described X86 shift amounts. Verify that we
-; do not try to create a shift with 140+120 as shift amount, and verify that


Should be 130+160?

jayfoad

LGTM

Ensure that the sum of the shift amounts does not overflow the shift amount type when combining shifts in combineShiftOfShiftedLogic. Solves a miscompile bug found when testing the C23 BitInt feature. Targets like X86 that only use an i8 for shift amounts after legalization seems to be extra susceptible for bugs like this as it isn't legal to shift more than 255 steps.

AreaZR · 2024-04-23T13:31:44Z

/cherry-pick 5fd9bbd f9b419b

…89616) Ensure that the sum of the shift amounts does not overflow the shift amount type when combining shifts in combineShiftOfShiftedLogic. Solves a miscompile bug found when testing the C23 BitInt feature. Targets like X86 that only use an i8 for shift amounts after legalization seems to be extra susceptible for bugs like this as it isn't legal to shift more than 255 steps. (cherry picked from commit f9b419b)

llvmbot · 2024-04-23T13:36:39Z

/cherry-pick 5fd9bbd f9b419b

Error: Command failed due to missing milestone.

…89616) Ensure that the sum of the shift amounts does not overflow the shift amount type when combining shifts in combineShiftOfShiftedLogic. Solves a miscompile bug found when testing the C23 BitInt feature. Targets like X86 that only use an i8 for shift amounts after legalization seems to be extra susceptible for bugs like this as it isn't legal to shift more than 255 steps. (cherry picked from commit f9b419b)

bjope requested review from RKSimon and fzhinkin April 22, 2024 15:52

llvmbot added backend:X86 llvm:SelectionDAG SelectionDAGISel as well labels Apr 22, 2024

nikic reviewed Apr 23, 2024

View reviewed changes

bjope force-pushed the shiftcombine branch from d059635 to 06019aa Compare April 23, 2024 07:29

jayfoad reviewed Apr 23, 2024

View reviewed changes

jayfoad approved these changes Apr 23, 2024

View reviewed changes

bjope force-pushed the shiftcombine branch from 06019aa to 94c3f9b Compare April 23, 2024 10:33

bjope force-pushed the shiftcombine branch from 94c3f9b to 575a092 Compare April 23, 2024 12:10

bjope merged commit f9b419b into llvm:main Apr 23, 2024
3 of 4 checks passed

pointhex mentioned this pull request May 7, 2024

getStyleDiagHandler #91314

Closed

aemerson mentioned this pull request May 9, 2024

release/18.x: [AArc64][GlobalISel] Fix legalizer assert for G_INSERT_VECTOR_ELT - manual merge #91672

Merged

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

[DAGCombiner] Fix miscompile bug in combineShiftOfShiftedLogic #89616

[DAGCombiner] Fix miscompile bug in combineShiftOfShiftedLogic #89616

bjope commented Apr 22, 2024

llvmbot commented Apr 22, 2024 •

edited

Loading

nikic Apr 23, 2024

bjope Apr 23, 2024

jayfoad Apr 23, 2024

jayfoad left a comment

AreaZR commented Apr 23, 2024

llvmbot commented Apr 23, 2024

[DAGCombiner] Fix miscompile bug in combineShiftOfShiftedLogic #89616

[DAGCombiner] Fix miscompile bug in combineShiftOfShiftedLogic #89616

Conversation

bjope commented Apr 22, 2024

llvmbot commented Apr 22, 2024 • edited Loading

nikic Apr 23, 2024

Choose a reason for hiding this comment

bjope Apr 23, 2024

Choose a reason for hiding this comment

jayfoad Apr 23, 2024

Choose a reason for hiding this comment

jayfoad left a comment

Choose a reason for hiding this comment

AreaZR commented Apr 23, 2024

llvmbot commented Apr 23, 2024

llvmbot commented Apr 22, 2024 •

edited

Loading