JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Add copy buttons to all
 blocks\n(function() {\n function addCopyButtons() {\n document.querySelectorAll('pre code').forEach(function(codeBlock) {\n if (codeBlock.parentElement.hasAttribute('data-copy-added')) return;\n codeBlock.parentElement.setAttribute('data-copy-added', 'true');\n \n var btn = document.createElement('button');\n btn.textContent = 'Copy';\n btn.style.cssText = 'position:absolute;top:4px;right:4px;padding:2px 8px;font-size:11px;background:#4ecdc4;border:none;border-radius:4px;color:#1a1a2e;cursor:pointer;opacity:0.7;transition:opacity 0.2s;';\n btn.onmouseover = function() { this.style.opacity = '1'; };\n btn.onmouseout = function() { this.style.opacity = '0.7'; };\n btn.onclick = function() {\n navigator.clipboard.writeText(codeBlock.textContent).then(function() {\n btn.textContent = 'Copied!';\n setTimeout(function() { btn.textContent = 'Copy'; }, 1500);\n });\n };\n codeBlock.parentElement.style.position = 'relative';\n codeBlock.parentElement.appendChild(btn);\n });\n }\n \n addCopyButtons();\n \n // Re-run on dynamic content\n var observer = new MutationObserver(addCopyButtons);\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Add Copy Buttons to Code Blocks");
}
} catch(__e) { console.warn('[Userscript:Add Copy Buttons to Code Blocks]', __e); }
})();
(function(){
try {
var __m = "github.com";
var __re = new RegExp('^' + "github\\.com" + '
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Force GitHub README to respect dark mode\n(function() {\n var style = document.createElement('style');\n style.textContent = '\n .markdown-body {\n color-scheme: dark light;\n }\n .markdown-body pre { background: #161b22 !important; }\n .markdown-body code { background: rgba(110, 118, 129, 0.4) !important; }\n .markdown-body table th, .markdown-body table td { border-color: #30363d !important; }\n .markdown-body img { background: #0d1117; }\n .markdown-body blockquote { border-left-color: #8b949e; }\n .markdown-body hr { border-color: #30363d; }\n ';\n document.head.appendChild(style);\n})();", "GitHub Dark Mode README Fix"); } } catch(__e) { console.warn('[Userscript:GitHub Dark Mode README Fix]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Highlight search terms from Google/DuckDuckGo/Bing referrer\n(function() {\n var ref = document.referrer;\n var terms = [];\n \n if (ref.includes('google.com') || ref.includes('duckduckgo.com') || ref.includes('bing.com')) {\n var url = new URL(ref);\n var q = url.searchParams.get('q') || url.searchParams.get('p');\n if (q) {\n terms = q.split(/\\s+/).filter(function(t) { return t.length > 2; });\n }\n }\n \n if (terms.length === 0) return;\n \n var style = document.createElement('style');\n style.textContent = '.userscript-highlight { background: #fbbf24; color: #1a1a2e; padding: 1px 3px; border-radius: 2px; }';\n document.head.appendChild(style);\n \n function highlight(node) {\n if (node.nodeType === 3) { // text node\n var text = node.textContent;\n var found = false;\n terms.forEach(function(term) {\n var regex = new RegExp('(' + term.replace(/[.*+?^${}()|[\\]\\\\]/g, '\\\\') + ')', 'gi');\n if (regex.test(text)) {\n found = true;\n var frag = document.createDocumentFragment();\n var parts = text.split(regex);\n parts.forEach(function(part, i) {\n if (i % 2 === 0) {\n frag.appendChild(document.createTextNode(part));\n } else {\n var span = document.createElement('span');\n span.className = 'userscript-highlight';\n span.textContent = part;\n frag.appendChild(span);\n }\n });\n node.parentNode.replaceChild(frag, node);\n }\n });\n } else if (node.nodeType === 1 && node.childNodes) { // element\n var skipTags = ['SCRIPT', 'STYLE', 'NOSCRIPT', 'TEXTAREA', 'INPUT', 'SELECT'];\n if (!skipTags.includes(node.tagName)) {\n Array.from(node.childNodes).forEach(highlight);\n }\n }\n }\n \n highlight(document.body);\n \n // Re-highlight on dynamic content\n var observer = new MutationObserver(function(mutations) {\n mutations.forEach(function(m) {\n m.addedNodes.forEach(function(node) {\n if (node.nodeType === 1 || node.nodeType === 3) highlight(node);\n });\n });\n });\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Highlight Search Terms"); } } catch(__e) { console.warn('[Userscript:Highlight Search Terms]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Strip utm_, fbclid, gclid, etc. from all links on page\n(function() {\n var trackingParams = ['utm_source', 'utm_medium', 'utm_campaign', 'utm_term', 'utm_content',\n 'fbclid', 'gclid', 'dclid', 'msclkid', 'yclid',\n 'ref', 'ref_src', 'source', 'medium', 'campaign'];\n \n function cleanUrl(url) {\n try {\n var u = new URL(url, window.location.origin);\n var changed = false;\n trackingParams.forEach(function(p) {\n if (u.searchParams.has(p)) {\n u.searchParams.delete(p);\n changed = true;\n }\n });\n return changed ? u.toString() : url;\n } catch (e) {\n return url;\n }\n }\n \n function cleanLinks() {\n document.querySelectorAll('a[href]').forEach(function(a) {\n var clean = cleanUrl(a.href);\n if (clean !== a.href) a.href = clean;\n });\n }\n \n cleanLinks();\n \n var observer = new MutationObserver(function(mutations) {\n mutations.forEach(function(m) {\n m.addedNodes.forEach(function(node) {\n if (node.nodeType === 1) {\n if (node.tagName === 'A') cleanLinks();\n node.querySelectorAll('a[href]').forEach(function(a) {\n var clean = cleanUrl(a.href);\n if (clean !== a.href) a.href = clean;\n });\n }\n });\n });\n });\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "Remove Tracking Parameters from Links"); } } catch(__e) { console.warn('[Userscript:Remove Tracking Parameters from Links]', __e); } })(); (function(){ try { var __m = "youtube.com"; var __re = new RegExp('^' + "youtube\\.com" + '
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Auto-enable theater mode on YouTube\n(function() {\n function tryTheater() {\n var btn = document.querySelector('button[aria-label=\"Theater mode\"], ytd-player #player button[title=\"Theater mode\"]');\n if (btn && !btn.classList.contains('activated')) {\n btn.click();\n }\n }\n \n // Try immediately\n tryTheater();\n \n // Try after navigation (SPA)\n var lastUrl = location.href;\n setInterval(function() {\n if (location.href !== lastUrl) {\n lastUrl = location.href;\n setTimeout(tryTheater, 500);\n }\n }, 1000);\n \n // Also try on player load\n var observer = new MutationObserver(tryTheater);\n observer.observe(document.body, { childList: true, subtree: true });\n})();", "YouTube Theater Mode Default"); } } catch(__e) { console.warn('[Userscript:YouTube Theater Mode Default]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Remove or un-stick sticky/fixed headers that block content\n(function() {\n function unstick() {\n document.querySelectorAll('header, nav, [role=\"banner\"], .header, .navbar, .sticky, .fixed-top, [style*=\"position: fixed\"], [style*=\"position:sticky\"]').forEach(function(el) {\n if (el.style.position === 'fixed' || el.style.position === 'sticky' || \n getComputedStyle(el).position === 'fixed' || getComputedStyle(el).position === 'sticky') {\n el.style.position = 'static';\n el.style.top = 'auto';\n el.style.zIndex = 'auto';\n }\n });\n }\n \n unstick();\n \n var observer = new MutationObserver(unstick);\n observer.observe(document.body, { childList: true, subtree: true, attributes: true, attributeFilter: ['style', 'class'] });\n})();", "Kill Sticky Headers"); } } catch(__e) { console.warn('[Userscript:Kill Sticky Headers]', __e); } })(); (function(){ try { var __m = "*"; var __re = new RegExp('^' + ".*" + '
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding
, 'i'); if (__m === '*' || __re.test(location.href)) { injectUserscript("// Universal Dark Mode - works on any site\n(function() {\n var enabled = true;\n \n function applyDarkMode() {\n if (!enabled) return;\n \n // Create style element if it doesn't exist\n var style = document.getElementById('universal-dark-mode-style');\n if (!style) {\n style = document.createElement('style');\n style.id = 'universal-dark-mode-style';\n document.head.appendChild(style);\n }\n \n // Dark mode CSS - inverts colors but preserves images/video\n style.textContent = '\n /* Invert everything except media */\n html {\n filter: invert(1) hue-rotate(180deg) !important;\n background: #1a1a2e !important;\n }\n \n /* Restore images, videos, iframes, canvas */\n img, video, iframe, canvas, svg, picture, [style*=\"background-image\"] {\n filter: invert(1) hue-rotate(180deg) !important;\n }\n \n /* Preserve specific elements that should not be inverted */\n .no-dark-mode, .no-dark-mode *,\n [data-theme=\"light\"], [data-theme=\"light\"],\n .ace_editor, .ace_editor *,\n .CodeMirror, .CodeMirror *,\n .monaco-editor, .monaco-editor *,\n .markdown-body pre, .markdown-body pre *,\n .highlight, .highlight *,\n pre code, pre code * {\n filter: none !important;\n }\n \n /* Fix common UI elements */\n .modal, .popup, .dropdown-menu, .tooltip, .popover {\n filter: invert(1) hue-rotate(180deg) !important;\n background: #2d2d44 !important;\n border-color: #444 !important;\n }\n \n /* Scrollbars */\n ::-webkit-scrollbar { background: #1a1a2e !important; }\n ::-webkit-scrollbar-thumb { background: #444 !important; }\n ::-webkit-scrollbar-thumb:hover { background: #555 !important; }\n \n /* Selection */\n ::selection { background: #4ecdc4 !important; color: #1a1a2e !important; }\n ::-moz-selection { background: #4ecdc4 !important; color: #1a1a2e !important; }\n ';\n }\n \n function removeDarkMode() {\n var style = document.getElementById('universal-dark-mode-style');\n if (style) style.remove();\n }\n \n // Toggle with Alt+Shift+D\n document.addEventListener('keydown', function(e) {\n if (e.altKey && e.shiftKey && e.key === 'D') {\n e.preventDefault();\n enabled = !enabled;\n if (enabled) {\n applyDarkMode();\n console.log('[Universal Dark Mode] Enabled');\n } else {\n removeDarkMode();\n console.log('[Universal Dark Mode] Disabled');\n }\n }\n });\n \n // Apply on load\n applyDarkMode();\n \n // Re-apply on dynamic content\n var observer = new MutationObserver(function(mutations) {\n if (enabled && !document.getElementById('universal-dark-mode-style')) {\n applyDarkMode();\n }\n });\n observer.observe(document.head, { childList: true });\n \n console.log('[Universal Dark Mode] Loaded - Press Alt+Shift+D to toggle');\n})();", "Universal Dark Mode"); } } catch(__e) { console.warn('[Userscript:Universal Dark Mode]', __e); } })(); })();
Skip to content

JIT: Refactor morph cmp optimizations - #127548

Closed
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization
Closed

JIT: Refactor morph cmp optimizations#127548
BoyBaykiller wants to merge 8 commits into
dotnet:mainfrom
BoyBaykiller:EQ-NE-canonicalization

Conversation

@BoyBaykiller

@BoyBaykillerBoyBaykiller commented Apr 29, 2026

Copy link
Copy Markdown
Contributor

Unifying the compare opts to naturally handle canonicalization of constant to the right for GT_EQ, GT_NE. This now also calls fgOptimizeCmpWithCasts for GT_EQ, GT_NE which was missing too.

I also added 2 canonicalizations:

  1. '(A & pow2) == pow2' -> '(A & pow2) != 0'
  2. '(A & pow2) != pow2' -> '(A & pow2) == 0'

These will help optimizeBools and consequentenly #126852 and probably other stuff. I'd also like this for #125899 if I add e.g SELECT(x & 16 == 16, y | 16, y) -> y | (x & 16)

@github-actionsgithub-actionsBot added the area-CodeGen-coreclr CLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMI label Apr 29, 2026
@dotnet-policy-servicedotnet-policy-serviceBot added the community-contribution Indicates that the PR has been added by a community member label Apr 29, 2026
@dotnet-policy-service

Copy link
Copy Markdown
Contributor

Tagging subscribers to this area: @JulieLeeMSFT, @jakobbotsch
See info in area-owners.md if you want to be subscribed.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
case GT_EQ:
case GT_NE:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I see that there is already a logic that moves constants to the right, why do we repeat it here?

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I totally agreee. But the fact that its written in a switch makes it hard/ugly to freely share code. I noticed the exact thing when I was adding fgOptimizeDistributiveArithemtic from my other PR. Which personally motivates me to rewrite this whole thing to using classic if stmts without switch

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'm not sure I follow, for me it sounds like it repeats work that should be done elsewhere. It's already done in fgMorphSmpOp for all comparisons, we probably might want to have a some shared way to do that for all commutative opers

@BoyBaykillerBoyBaykillerApr 29, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I though you are talking about this right down below from where I stole it which only does it for some:

caseGT_LT:
caseGT_LE:
caseGT_GE:
caseGT_GT:
// Change "CNS relop op2" to "op2 relop* CNS"
if (op1->IsIntegralConst() && tree->OperIsCompare() && gtCanSwapOrder(op1, op2))
{
std::swap(tree->AsOp()->gtOp1, tree->AsOp()->gtOp2);
tree->gtOper = GenTree::SwapRelop(tree->OperGet());
oper = tree->OperGet();
op1 = tree->gtGetOp1();
op2 = tree->gtGetOp2();
}
if (op1->OperIs(GT_CAST) || op2->OperIs(GT_CAST))
{

Not sure what place you are refering to

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

we probably might want to have a some shared way to do that for all commutative opers

That'd be nice

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

The check just below it is also notably too lenient, it is valid for floating-point as well and we handle that for SIMD nodes.

You can likely just extract it out to a small helper function if you don't want to duplicate the logic between EQ/NE and the other compares, although notably EQ/NE are simpler since they are commutative.

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

If you happen to know where this code is that should already be handling comparisons, please let me know. So I can try to improve it. I will try to find it now.

Looks like it might indeed be missing, outside of costing.

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I'd recommend doing that one in it's own PR (it's much more broadly impacting and not isolated to the ispow2 work) and ideally fixing it to not only do it for integrals, we can do it for all types just fine and just check SwapRelop

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

and ideally fixing it to not only do it for integrals, we can do it for all types

are there optimizations on relops with floats? if yes where

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

In some cases, yes, and we do folding and other transforms based on them. There are more for vectors than there are for scalars right now, but ideally they are mirrored. In general we want to push constants right for other reasons as well, such as CSE, containment, VN, etc.

Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/morph.cpp Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
Comment threadsrc/coreclr/jit/gentree.h Outdated
if (IsIntegralConst())
{
return isPow2(AsIntConCommon()->IntegralValue());
if (IsCnsIntOrI())

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

nit: This is not correct as it checks for GT_CNS_INT where you must instead check for TypeIs(TYP_INT)

@BoyBaykillerBoyBaykillerApr 30, 2026

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

if (IsIntegralConst())
{
if (TypeIs(TYP_INT))
{
returnisPow2((uint32_t)AsIntCon()->IconValue());
}
returnisPow2((uint64_t)AsLngCon()->LngValue());
}

Input:

boolTest1(longA,longB){return8==(A&8);}

Gives me: Assertion failed 'OperIs(GT_CNS_LNG)'. I guess I should use AsIntConCommon()->IntegralValue().

@tannergoodingtannergoodingApr 30, 2026

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

What platform are you getting the assertion on? and for which line?

There shouldn't be any GT_CNS_LNG being produced on 64-bit, they all get routed to GT_CNS_INT. Then you shouldn't be producing a GT_CNS_LNG for TYP_INT on 32-bit either, correspondingly.

inlinessize_tGenTreeIntConCommon::IconValue() const
{
assert(OperIs(GT_CNS_INT)); // We should never see a GT_CNS_LNG for a 64-bit target!returnAsIntCon()->gtIconVal;
}

and

inlineINT64GenTreeIntConCommon::LngValue() const
{
#ifndef TARGET_64BIT
assert(OperIs(GT_CNS_LNG));
returnAsLngCon()->gtLconVal;
#elsereturnIconValue();
#endif
}

and

GenTreeLngCon(INT64 val)
: GenTreeIntConCommon(GT_CNS_NATIVELONG, TYP_LONG)
{
SetLngValue(val);
}

and

GenTreeIntCon(var_types type, ssize_t value DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(nullptr)
{
}
GenTreeIntCon(var_types type, ssize_t value, FieldSeq* fields DEBUGARG(bool largeNode = false))
: GenTreeIntConCommon(GT_CNS_INT, type DEBUGARG(largeNode))
, gtIconVal(value)
, gtCompileTimeHandle(0)
, gtFieldSeq(fields)
{
}

Copy link
Copy Markdown
ContributorAuthor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I will be able to properly answer you tomorrow but you should be able to reproduce on win-x64 main.

Comment threadsrc/coreclr/jit/gentree.h Outdated
@BoyBaykillerBoyBaykiller changed the title JIT: Add canonicalization to GT_EQ and GT_NEJIT: Refactor morph cmp optimizationsApr 30, 2026
@BoyBaykiller

Copy link
Copy Markdown
ContributorAuthor

Split into
Canon (cns to right): #127661 (more general, better version)
Canon (bit-test): #128533
Pow2: #127615
Unify: #127883

@github-actionsgithub-actionsBot locked and limited conversation to collaborators May 31, 2026
Sign up for freeto subscribe to this conversation on GitHub. Already have an account? Sign in.

Labels

area-CodeGen-coreclrCLR JIT compiler in src/coreclr/src/jit and related components such as SuperPMIcommunity-contributionIndicates that the PR has been added by a community member

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants

@BoyBaykiller@EgorBo@tannergooding