clang-p2996

Files

woruyu bbcebec3af [DAG] Refactor X86 combineVSelectWithAllOnesOrZeros fold into a generic DAG Combine (#145298 )

This PR resolves https://github.com/llvm/llvm-project/issues/144513

The modification include five pattern :
1.vselect Cond, 0, 0 → 0
2.vselect Cond, -1, 0 → bitcast Cond
3.vselect Cond, -1, x → or Cond, x
4.vselect Cond, x, 0 → and Cond, x
5.vselect Cond, 000..., X -> andn Cond, X

1-4 have been migrated to DAGCombine. 5 still in x86 code.

The reason is that you cannot use the andn instruction directly in
DAGCombine, you can only use and+xor, which will introduce optimization
order issues. For example, in the x86 backend, select Cond, 0, x →
(~Cond) & x, the backend will first check whether the cond node of
(~Cond) is a setcc node. If so, it will modify the comparison operator
of the condition.So the x86 backend cannot complete the optimization of
andn.In short, I think it is a better choice to keep the pattern of
vselect Cond, 000..., X instead of and+xor in combineDAG.

For commit, the first is code changes and x86 test(note 1), the second
is tests in other backend(node 2).

---------

Co-authored-by: Simon Pilgrim <llvm-dev@redking.me.uk>

2025-07-02 15:07:48 +01:00

AsmPrinter

[DwarfDebug] Slightly optimize computeKeyInstructions() (NFC) (#146357 )

2025-07-01 09:06:17 +02:00

GlobalISel

[GlobalISel] Allow expansion of urem by constant in prelegalizer (#145914 )

2025-07-02 13:46:36 +01:00

LiveDebugValues

[InstrRef] Skip clobbered EntryValue register recovery (#142478 )

2025-06-27 10:30:29 -07:00

MIRParser

[MIRParser] Report register class errors in a deterministic order (#142928 )

2025-06-06 10:03:34 +01:00

SelectionDAG

[DAG] Refactor X86 combineVSelectWithAllOnesOrZeros fold into a generic DAG Combine (#145298 )

2025-07-02 15:07:48 +01:00

AggressiveAntiDepBreaker.cpp

[LivePhysRegs] Use MCRegister instead of MCPhysReg in interface. NFC

2025-03-06 09:07:53 -08:00

AggressiveAntiDepBreaker.h

[AggressiveAntiDepBreaker] Use MCRegister. NFC

2025-03-02 15:17:23 -08:00

AllocationOrder.cpp

[CodeGen] Use Register or MCRegister. NFC

2025-03-06 09:08:21 -08:00

AllocationOrder.h

[NFC][regalloc] Fix typo in llvm/lib/CodeGen/AllocationOrder.h.

2025-05-02 10:11:34 +08:00

Analysis.cpp

…

AssignmentTrackingAnalysis.cpp

[RemoveDIs][NFC] Remove dbg intrinsic handling code from AssignmentTrackingAnalysis (#144674 )

2025-06-24 13:07:31 +01:00

AtomicExpandPass.cpp

[AtomicExpandPass] Match isIdempotentRMW with InstcombineRMW (#142277 )

2025-06-08 11:32:20 +01:00

BasicBlockPathCloning.cpp

[ARM,AArch64] Don't put BTI at asm goto branch targets (#141562 )

2025-06-03 08:44:13 +01:00

BasicBlockSections.cpp

…

BasicBlockSectionsProfileReader.cpp

…

BasicTargetTransformInfo.cpp

…

BranchFolding.cpp

[DLCov][NFC] Propagate annotated DebugLocs through transformations (#138047 )

2025-06-12 14:06:27 +01:00

BranchFolding.h

…

BranchRelaxation.cpp

[CodeGen] Remove unused local variables (NFC) (#138441 )

2025-05-04 00:26:37 -07:00

BreakFalseDeps.cpp

[X86][BreakFalseDeps] Using reverse order for undef register selection (#137569 )

2025-06-11 22:08:20 +08:00

CalcSpillWeights.cpp

[RegAlloc] Sort CopyHint by IsCSR (#131046 )

2025-04-14 09:58:46 -04:00

CallBrPrepare.cpp

…

CallingConvLower.cpp

[CallingConvLower] Use MCRegister. NFC

2025-03-02 23:46:18 -08:00

CFGuardLongjmp.cpp

…

CFIFixup.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

CFIInstrInserter.cpp

[CFIInstrInserter] Don't store Dwarf register number in Register. NFC

2025-03-02 11:33:09 -08:00

CMakeLists.txt

Add support for Windows Secure Hot-Patching (redo) (#145565 )

2025-06-24 14:56:55 -07:00

CodeGen.cpp

[CodeGen][NPM] Port VirtRegRewriter to NPM (#130564 )

2025-04-30 14:10:46 +05:30

CodeGenCommonISel.cpp

…

CodeGenPrepare.cpp

[CodeGenPrepare] Filter out unrecreatable addresses from memory optimization (#143566 )

2025-06-28 23:30:03 +02:00

CodeGenTargetMachineImpl.cpp

Triple: Forward declare Twine and remove include (#145685 )

2025-06-26 15:26:04 +09:00

CommandFlags.cpp

[TargetRegistry] Accept Triple in createTargetMachine() (NFC) (#130940 )

2025-03-12 17:35:09 +01:00

ComplexDeinterleavingPass.cpp

[LLVM][ComplexDeinterleaving] Update splat identification to include vector ConstantInt/FP. (#144516 )

2025-06-18 11:53:27 +01:00

CriticalAntiDepBreaker.cpp

[CriticalAntiDepBreaker] Fix another MSVC build error. NFC

2025-03-07 00:09:07 -08:00

CriticalAntiDepBreaker.h

[CriticalAntiDepBreaker] Use Register and MCRegister. NFC

2025-03-06 23:35:01 -08:00

DeadMachineInstructionElim.cpp

[CodeGen] NFC: Move isDead to MachineInstr (#123531 )

2025-01-23 12:54:29 -08:00

DetectDeadLanes.cpp

[CodeGen][NPM] Port DetectDeadLanes to NPM (#130567 )

2025-03-12 11:22:02 +05:30

DFAPacketizer.cpp

…

DroppedVariableStatsMIR.cpp

Add a pass to collect dropped var statistics for MIR (#126686 )

2025-02-12 14:08:18 -08:00

DwarfEHPrepare.cpp

…

EarlyIfConversion.cpp

[EarlyIfConverter] Fix reg killed twice after early-if-predicator and ifcvt (#133554 )

2025-04-01 12:06:30 +02:00

EdgeBundles.cpp

…

EHContGuardTargets.cpp

[win] NFC: Rename EHCatchret to EHCont to allow for EH Continuation targets that aren't catchret instructions (#129953 )

2025-03-06 09:28:44 -08:00

ExecutionDomainFix.cpp

[CodeGen] Use Register or MCRegister. NFC

2025-03-06 09:08:21 -08:00

ExpandFp.cpp

Rename ExpandLargeFpConvertPass to ExpandFpPass (#131128 )

2025-03-14 13:11:45 +01:00

ExpandLargeDivRem.cpp

…

ExpandMemCmp.cpp

…

ExpandPostRAPseudos.cpp

[CodeGen][NPM] Port ExpandPostRAPseudos to NPM (#129509 )

2025-03-04 11:49:09 +05:30

ExpandReductions.cpp

…

ExpandVectorPredication.cpp

[NFC][LLVM] Refactor IRBuilder::Create{VScale,ElementCount,TypeSize}. (#142803 )

2025-06-10 12:35:59 +01:00

FaultMaps.cpp

MC: Migrate away from operator<< MCExpr

2025-06-28 10:58:09 -07:00

FEntryInserter.cpp

[CodeGen][NPM] Port FEntryInserter to NPM (#129857 )

2025-03-17 10:35:53 +05:30

FinalizeISel.cpp

[AMDGPU] Fix undefined scc register in successor block of SI_KILL terminators (#134718 )

2025-04-30 09:02:45 -05:00

FixupStatepointCallerSaved.cpp

[llvm] Use *Set::insert_range (NFC) (#133353 )

2025-03-27 20:44:20 -07:00

FuncletLayout.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

GCEmptyBasicBlocks.cpp

…

GCMetadata.cpp

[GC] Use MapVector for GCStrategyMap (#132729 )

2025-05-14 17:51:18 +08:00

GCMetadataPrinter.cpp

…

GCRootLowering.cpp

…

GlobalMerge.cpp

[GlobalMerge][PPC] Don't merge globals in llvm.metadata section (#131801 )

2025-04-02 10:40:53 +02:00

GlobalMergeFunctions.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

HardwareLoops.cpp

[Remarks] Remove an upcast footgun. NFC (#142191 )

2025-05-31 11:07:54 -07:00

IfConversion.cpp

[IfConversion] Fix bug related to !HasFallThrough (#145471 )

2025-06-25 09:30:26 +02:00

ImplicitNullChecks.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

IndirectBrExpandPass.cpp

[llvm] Use std::none_of (NFC) (#143320 )

2025-06-08 16:18:38 -07:00

InitUndef.cpp

[CodeGen] Use non-static Register::virtRegIndex() instead of static Register::virtReg2Index. NFC (#125031 )

2025-01-30 00:14:08 -08:00

InlineSpiller.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

InterferenceCache.cpp

[CodeGen] Remove some implict conversions of MCRegister to unsigned by using(). NFC

2025-01-19 13:18:04 -08:00

InterferenceCache.h

…

InterleavedAccessPass.cpp

[IA] Remove recursive [de]interleaving support (#143875 )

2025-06-25 12:29:45 +01:00

InterleavedLoadCombinePass.cpp

…

IntrinsicLowering.cpp

…

JMCInstrumenter.cpp

…

KCFI.cpp

…

LatencyPriorityQueue.cpp

…

LazyMachineBlockFrequencyInfo.cpp

…

LexicalScopes.cpp

[llvm] Use *Set::insert_range (NFC) (#138237 )

2025-06-02 19:48:13 -07:00

LiveDebugVariables.cpp

[CodeGen] Remove unused local variables (NFC) (#138441 )

2025-05-04 00:26:37 -07:00

LiveInterval.cpp

[CodeGen] Use range-based for loops (NFC) (#145251 )

2025-06-22 19:09:38 -07:00

LiveIntervalCalc.cpp

…

LiveIntervals.cpp

[CodeGen] Introduce a VirtRegOrUnit class to hold virtual reg or physical reg unit. NFC (#123768 )

2025-01-24 18:30:28 -08:00

LiveIntervalUnion.cpp

[CodeGen] Use Register::id(). NFC

2025-03-06 09:08:21 -08:00

LivePhysRegs.cpp

[PowerPC] Deprecate uses of ISD::ADDC/ISD::ADDE/ISD::SUBC/ISD::SUBE (#133155 )

2025-04-03 13:22:49 -04:00

LiveRangeCalc.cpp

[CodeGen] Use Register or MCRegister. NFC

2025-03-06 09:08:21 -08:00

LiveRangeEdit.cpp

[CodeGen] Remove parameter from LiveRangeEdit::canRematerializeAt [NFC]

2025-03-14 09:12:07 -07:00

LiveRangeShrink.cpp

LiveRangeShrink: Early exit when encountering a code motion barrier.

2025-04-24 12:44:51 -07:00

LiveRangeUtils.h

…

LiveRegMatrix.cpp

[LiveRegMatrix] Use MCRegUnit instead of MCRegister for register unit. NFC

2025-01-20 10:57:34 -08:00

LiveRegUnits.cpp

…

LiveStacks.cpp

[CodeGen] Avoid repeated map lookups (NFC) (#127025 )

2025-02-13 09:11:17 -08:00

LiveVariables.cpp

CodeGen: Convert some assorted errors to use reportFatalUsageError (#142031 )

2025-05-30 08:06:53 +02:00

LocalStackSlotAllocation.cpp

[CodeGen] Use Register or MCRegister. NFC

2025-03-06 09:08:21 -08:00

LoopTraversal.cpp

…

LowerEmuTLS.cpp

IR: Make Module::getOrInsertGlobal() return a GlobalVariable.

2025-05-27 12:23:12 -07:00

LowLevelTypeUtils.cpp

…

MachineBasicBlock.cpp

[llvm] Remove unused includes (NFC) (#144293 )

2025-06-16 08:59:18 -07:00

MachineBlockFrequencyInfo.cpp

…

MachineBlockPlacement.cpp

[CodeGen][CodeLayout] Fix segfault on access to deleted block in MBP. (#142357 )

2025-06-23 23:04:22 +09:00

MachineBranchProbabilityInfo.cpp

…

MachineCFGPrinter.cpp

…

MachineCheckDebugify.cpp

…

MachineCombiner.cpp

[MachineCombiner][Targets] Use Register in TII genAlternativeCodeSequence interface. NFC (#131272 )

2025-03-13 23:27:56 -07:00

MachineConvergenceVerifier.cpp

…

MachineCopyPropagation.cpp

[MCP] Handle iterative simplification during forward copy prop (#140267 )

2025-06-02 11:21:41 -07:00

MachineCSE.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

MachineCycleAnalysis.cpp

[CodeGen][NewPM] Port MachineCycleInfo to NPM (#114745 )

2025-03-03 11:26:17 +05:30

MachineDebugify.cpp

[CodeGen] Avoid repeated hash lookups (NFC) (#131495 )

2025-03-16 11:02:57 -07:00

MachineDominanceFrontier.cpp

…

MachineDominators.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

MachineDomTreeUpdater.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

MachineFrameInfo.cpp

…

MachineFunction.cpp

[CodeGen][NFC] Fix quadratic c-t for large jump tables

2025-06-18 18:56:30 +02:00

MachineFunctionAnalysis.cpp

…

MachineFunctionPass.cpp

Add a pass to collect dropped var statistics for MIR (#126686 )

2025-02-12 14:08:18 -08:00

MachineFunctionPrinterPass.cpp

…

MachineFunctionSplitter.cpp

…

MachineInstr.cpp

[NFC][LLVM][CodeGen] Refactor MachineInstr operand accessors (#137261 )

2025-05-01 07:45:22 -07:00

MachineInstrBundle.cpp

[CodeGen] Simplify finalizeBundle. NFC. (#139234 )

2025-05-09 16:10:52 +01:00

MachineLateInstrsCleanup.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

MachineLICM.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

MachineLoopInfo.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

MachineLoopUtils.cpp

[CodeGen] Avoid repeated hash lookups (NFC) (#124078 )

2025-01-23 08:46:19 -08:00

MachineModuleInfo.cpp

…

MachineModuleInfoImpls.cpp

…

MachineModuleSlotTracker.cpp

…

MachineOperand.cpp

[CodeGen][LLVM] Fix MachineOperand::print crash when TII is nullptr. (#135170 )

2025-04-11 15:26:55 +08:00

MachineOptimizationRemarkEmitter.cpp

…

MachineOutliner.cpp

[MachineOutliner] Remove LOHs from outlined candidates (#143617 )

2025-06-30 14:29:06 -07:00

MachinePassManager.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

MachinePipeliner.cpp

[MachinePipeliner] Introduce a new class for loop-carried deps (#137663 )

2025-06-05 21:30:27 +09:00

MachinePostDominators.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

MachineRegionInfo.cpp

…

MachineRegisterInfo.cpp

Reapply "[AMDGPU][Scheduler] Refactor ArchVGPR rematerialization during scheduling (#125885 )" (#139548 )

2025-05-13 11:11:00 +02:00

MachineScheduler.cpp

[CodeGen] Use std::tie to implement a comparison functor (NFC) (#146252 )

2025-06-29 08:25:53 -07:00

MachineSink.cpp

[DLCov][NFC] Propagate annotated DebugLocs through transformations (#138047 )

2025-06-12 14:06:27 +01:00

MachineSizeOpts.cpp

…

MachineSSAContext.cpp

[CodeGen] Use Register or MCRegister. NFC

2025-03-06 09:08:21 -08:00

MachineSSAUpdater.cpp

…

MachineStableHash.cpp

[CodeGen] Use Register::id(). NFC

2025-03-06 09:08:21 -08:00

MachineStripDebug.cpp

…

MachineTraceMetrics.cpp

[MachineTraceMetrics] Use Register::id(). NFC

2025-03-06 23:29:17 -08:00

MachineUniformityAnalysis.cpp

[CodeGen] Port MachineUniformityAnalysis to new pass manager (#137578 )

2025-04-30 10:44:06 +08:00

MachineVerifier.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

MacroFusion.cpp

[CodeGen] Fix a warning

2025-06-05 01:18:34 -07:00

MBFIWrapper.cpp

…

MIRCanonicalizerPass.cpp

[MIRCanonicalizerPass] Use MCRegister. NFC

2025-03-02 23:46:16 -08:00

MIRFSDiscriminator.cpp

…

MIRNamerPass.cpp

…

MIRPrinter.cpp

[DebugInfo][RemoveDIs] Remove scoped-dbg-format-setter (#143450 )

2025-06-11 11:23:24 +01:00

MIRPrintingPass.cpp

[NFC][LLVM][CodeGen] Refactor MIR Printer (#137361 )

2025-05-01 10:00:54 -07:00

MIRSampleProfile.cpp

…

MIRVRegNamerUtils.cpp

[llvm] Call hash_combine_range with ranges (NFC) (#136511 )

2025-04-20 16:36:03 -07:00

MIRVRegNamerUtils.h

[MIRVRegNamerUtils] Use Register. NFC

2025-03-06 22:51:05 -08:00

MIRYamlMapping.cpp

…

MLRegAllocEvictAdvisor.cpp

Revert "[llvm][NFC] Use llvm::sort()" (#140668 )

2025-05-20 11:27:03 +08:00

MLRegAllocEvictAdvisor.h

…

MLRegAllocPriorityAdvisor.cpp

[CodeGen][NewPM] Port RegAllocPriorityAdvisor analysis to NPM (#118462 )

2025-02-20 09:35:49 +05:30

ModuloSchedule.cpp

[CodeGen] Remove unused local variables (NFC) (#138441 )

2025-05-04 00:26:37 -07:00

MultiHazardRecognizer.cpp

…

NonRelocatableStringpool.cpp

[llvm] Use *Map::try_emplace (NFC) (#141190 )

2025-05-22 23:50:58 -07:00

OptimizePHIs.cpp

[CodeGen] Use Register or MCRegister. NFC

2025-03-06 09:08:21 -08:00

PatchableFunction.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

PeepholeOptimizer.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

PHIElimination.cpp

[PHIElimination] Verify reappropriated COPY is of similar register class, update livevars. (#146337 )

2025-07-01 19:18:31 +03:00

PHIEliminationUtils.cpp

[CodeGen] Use MCRegister and Register. NFC

2025-03-02 22:33:26 -08:00

PHIEliminationUtils.h

[CodeGen] Use MCRegister and Register. NFC

2025-03-02 22:33:26 -08:00

PostRAHazardRecognizer.cpp

[CodeGen][NPM] Port PostRAHazardRecognizer to NPM (#130066 )

2025-04-09 16:36:22 +05:30

PostRASchedulerList.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

PreISelIntrinsicLowering.cpp

[PreISelIntrinsicLowering] Reuse previously generated GlobalVariable for memset_pattern16 when possible (#144677 )

2025-06-23 16:35:48 +01:00

ProcessImplicitDefs.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

PrologEpilogInserter.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

PseudoProbeInserter.cpp

Clean up external users of GlobalValue::getGUID(StringRef) (#129644 )

2025-04-28 11:09:43 +10:00

PseudoSourceValue.cpp

…

RDFGraph.cpp

[llvm] Use range constructors of *Set (NFC) (#133549 )

2025-03-28 19:55:18 -07:00

RDFLiveness.cpp

[CodeGen] Avoid repeated map lookups (NFC) (#140662 )

2025-05-19 21:36:31 -07:00

RDFRegisters.cpp

[CodeGen] Remove static member functions Register::stackSlot2Index/isStackSlot. NFC

2025-02-19 21:54:43 -08:00

ReachingDefAnalysis.cpp

[CodeGen] Use MCRegister and Register. NFC

2025-03-02 22:33:26 -08:00

README.txt

…

RegAllocBase.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegAllocBase.h

RegAlloc: Use new approach to handling failed allocations (#128469 )

2025-02-26 15:34:47 +07:00

RegAllocBasic.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegAllocEvictionAdvisor.cpp

[CodeGen][NewPM] Port RegAllocPriorityAdvisor analysis to NPM (#118462 )

2025-02-20 09:35:49 +05:30

RegAllocFast.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegAllocGreedy.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegAllocGreedy.h

[CodeGen][NewPM] Port RegAllocGreedy to NPM (#119540 )

2025-02-26 12:11:22 +05:30

RegAllocPBQP.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegAllocPriorityAdvisor.cpp

[CodeGen][NewPM] Port RegAllocPriorityAdvisor analysis to NPM (#118462 )

2025-02-20 09:35:49 +05:30

RegAllocScore.cpp

[llvm] annotate remaining CodeGen and CodeGenTypes library interfaces for DLL export (#145361 )

2025-06-25 13:00:59 -07:00

RegAllocScore.h

…

RegisterBank.cpp

…

RegisterBankInfo.cpp

[llvm] Call hash_combine_range with ranges (NFC) (#136511 )

2025-04-20 16:36:03 -07:00

RegisterClassInfo.cpp

[X86][BreakFalseDeps] Using reverse order for undef register selection (#137569 )

2025-06-11 22:08:20 +08:00

RegisterCoalescer.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegisterCoalescer.h

[CodeGen][NewPM] Port RegisterCoalescer to NPM (#124698 )

2025-02-03 13:41:51 +07:00

RegisterPressure.cpp

Revert "[CodeGen] Remove static member function Register::isVirtualRegister. NFC (#127968 )"

2025-02-20 22:06:21 +00:00

RegisterScavenging.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RegisterUsageInfo.cpp

[CodeGen] Construct SmallVector with iterator ranges (NFC) (#136258 )

2025-04-18 10:26:48 -07:00

RegUsageInfoCollector.cpp

[CodeGen] Remove some implict conversions of MCRegister to unsigned by using(). NFC

2025-01-19 13:18:04 -08:00

RegUsageInfoPropagate.cpp

…

RemoveLoadsIntoFakeUses.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

RemoveRedundantDebugValues.cpp

[CodeGen][NewPM] Port "RemoveRedundantDebugValues" to NPM (#129005 )

2025-03-03 19:57:50 +07:00

RenameIndependentSubregs.cpp

Revert "[CodeGen] Use range-based for loops (NFC) (#138434 )"

2025-05-04 17:36:52 -04:00

ReplaceWithVeclib.cpp

…

ResetMachineFunctionPass.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

SafeStack.cpp

[IRBuilder] Add new overload for CreateIntrinsic (#131942 )

2025-03-31 08:10:34 -07:00

SafeStackLayout.cpp

…

SafeStackLayout.h

…

SanitizerBinaryMetadata.cpp

[CodeGen][NPM] Port MachineSanitizerBinaryMetadata to NPM (#130069 )

2025-04-14 20:52:26 +05:30

ScheduleDAG.cpp

MachineScheduler: Improve instruction clustering (#137784 )

2025-06-05 15:28:04 +08:00

ScheduleDAGInstrs.cpp

[ScheduleDAG] Allow disabling the SchedModel / Itineraries during Scheduling (#138057 )

2025-05-05 14:07:23 -07:00

ScheduleDAGPrinter.cpp

…

ScoreboardHazardRecognizer.cpp

…

SelectOptimize.cpp

[CostModel] Remove optional from InstructionCost::getValue() (#135596 )

2025-04-23 07:46:27 +01:00

ShadowStackGCLowering.cpp

[GC] Use MapVector for GCStrategyMap (#132729 )

2025-05-14 17:51:18 +08:00

ShrinkWrap.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

SjLjEHPrepare.cpp

[CodeGen] Use *Set::insert_range (NFC) (#132651 )

2025-03-23 21:20:44 -07:00

SlotIndexes.cpp

[CodeGen] Combine two loops in SlotIndexes.cpp file (#127631 )

2025-03-07 08:17:23 +07:00

SpillPlacement.cpp

…

SplitKit.cpp

[llvm] Use llvm::SmallVector::pop_back_val (NFC) (#136533 )

2025-04-21 08:13:16 -07:00

SplitKit.h

SplitKit: Fix rematerialization undoing subclass based split (#122110 )

2025-03-04 10:04:14 +07:00

StackColoring.cpp

[CodeGen] Avoid repeated hash lookups (NFC) (#132329 )

2025-03-21 08:00:45 -07:00

StackFrameLayoutAnalysisPass.cpp

[CodeGen][NPM] Port StackFrameLayoutAnalysisPass to NPM (#130070 )

2025-04-15 12:37:19 +05:30

StackMapLivenessAnalysis.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

StackMaps.cpp

[CodeGen] Avoid repeated hash lookups (NFC) (#140838 )

2025-05-20 21:39:17 -07:00

StackProtector.cpp

Revert "[llvm][StackProtector] Add noreturn to __stack_chk_fail call" (#144452 )

2025-06-16 16:34:40 -07:00

StackSlotColoring.cpp

[StackSlotColoring] Fix issue where colors for a StackID are dropped (#138140 )

2025-05-02 10:23:39 +01:00

StaticDataAnnotator.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

StaticDataSplitter.cpp

[CodeGen] Remove unused includes (NFC) (#141320 )

2025-05-24 00:00:00 -07:00

SwiftErrorValueTracking.cpp

…

SwitchLoweringUtils.cpp

…

TailDuplication.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

TailDuplicator.cpp

Use early return/continue in TailDuplicator::duplicateInstruction [nfc]

2025-05-16 13:33:11 -07:00

TargetFrameLoweringImpl.cpp

Reland [AMDGPU] Support block load/store for CSR #130013 (#137169 )

2025-04-25 11:29:27 +02:00

TargetInstrInfo.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

TargetLoweringBase.cpp

RuntimeLibcalls: Pass in ABI name from MCOptions (#144894 )

2025-06-23 22:14:44 +09:00

TargetLoweringObjectFileImpl.cpp

[GOFF] Add writing of section symbols (#133799 )

2025-06-26 11:52:14 -04:00

TargetOptionsImpl.cpp

TargetOptions: Look up frame-pointer attribute once (#146639 )

2025-07-02 20:09:20 +09:00

TargetPassConfig.cpp

Add support for Windows Secure Hot-Patching (redo) (#145565 )

2025-06-24 14:56:55 -07:00

TargetRegisterInfo.cpp

[RegAlloc] Scale the spill weight by target factor (#113675 )

2025-03-13 12:27:59 +08:00

TargetSchedule.cpp

[ScheduleDAG] Allow disabling the SchedModel / Itineraries during Scheduling (#138057 )

2025-05-05 14:07:23 -07:00

TargetSubtargetInfo.cpp

…

TwoAddressInstructionPass.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

TypePromotion.cpp

[TypePromotion] Do not zero-extend getelementptr indexes since signed

2025-04-05 16:49:05 +02:00

UnreachableBlockElim.cpp

[llvm] Use range constructors of *Set (NFC) (#137552 )

2025-04-27 15:59:57 -07:00

ValueTypes.cpp

[AArch64] Generalize integer FPR lane stores for all types (#134117 )

2025-04-17 09:25:57 +01:00

VirtRegMap.cpp

[LLVM][CodeGen] Add convenience accessors for MachineFunctionProperties (#140002 )

2025-05-22 08:07:52 -07:00

VLIWMachineScheduler.cpp

…

WasmEHPrepare.cpp

IR: Make Module::getOrInsertGlobal() return a GlobalVariable.

2025-05-27 12:23:12 -07:00

WindowScheduler.cpp

[llvm] Use llvm::is_contained (NFC) (#135566 )

2025-04-13 16:35:29 -07:00

WindowsSecureHotPatching.cpp

Add support for Windows Secure Hot-Patching (redo) (#145565 )

2025-06-24 14:56:55 -07:00

WinEHPrepare.cpp

[WinEH] Track changes in WinEHPrepare pass (#134121 )

2025-05-30 15:29:32 +02:00

XRayInstrumentation.cpp

[XRay] Fix tail call sleds for AArch64 (#141403 )

2025-05-29 21:54:15 -07:00

README.txt

//===---------------------------------------------------------------------===//

Common register allocation / spilling problem:

        mul lr, r4, lr
        str lr, [sp, #+52]
        ldr lr, [r1, #+32]
        sxth r3, r3
        ldr r4, [sp, #+52]
        mla r4, r3, lr, r4

can be:

        mul lr, r4, lr
        mov r4, lr
        str lr, [sp, #+52]
        ldr lr, [r1, #+32]
        sxth r3, r3
        mla r4, r3, lr, r4

and then "merge" mul and mov:

        mul r4, r4, lr
        str r4, [sp, #+52]
        ldr lr, [r1, #+32]
        sxth r3, r3
        mla r4, r3, lr, r4

It also increase the likelihood the store may become dead.

//===---------------------------------------------------------------------===//

bb27 ...
        ...
        %reg1037 = ADDri %reg1039, 1
        %reg1038 = ADDrs %reg1032, %reg1039, %noreg, 10
    Successors according to CFG: 0x8b03bf0 (#5)

bb76 (0x8b03bf0, LLVM BB @0x8b032d0, ID#5):
    Predecessors according to CFG: 0x8b0c5f0 (#3) 0x8b0a7c0 (#4)
        %reg1039 = PHI %reg1070, mbb<bb76.outer,0x8b0c5f0>, %reg1037, mbb<bb27,0x8b0a7c0>

Note ADDri is not a two-address instruction. However, its result %reg1037 is an
operand of the PHI node in bb76 and its operand %reg1039 is the result of the
PHI node. We should treat it as a two-address code and make sure the ADDri is
scheduled after any node that reads %reg1039.

//===---------------------------------------------------------------------===//

Use local info (i.e. register scavenger) to assign it a free register to allow
reuse:
        ldr r3, [sp, #+4]
        add r3, r3, #3
        ldr r2, [sp, #+8]
        add r2, r2, #2
        ldr r1, [sp, #+4]  <==
        add r1, r1, #1
        ldr r0, [sp, #+4]
        add r0, r0, #2

//===---------------------------------------------------------------------===//

LLVM aggressively lift CSE out of loop. Sometimes this can be negative side-
effects:

R1 = X + 4
R2 = X + 7
R3 = X + 15

loop:
load [i + R1]
...
load [i + R2]
...
load [i + R3]

Suppose there is high register pressure, R1, R2, R3, can be spilled. We need
to implement proper re-materialization to handle this:

R1 = X + 4
R2 = X + 7
R3 = X + 15

loop:
R1 = X + 4  @ re-materialized
load [i + R1]
...
R2 = X + 7 @ re-materialized
load [i + R2]
...
R3 = X + 15 @ re-materialized
load [i + R3]

Furthermore, with re-association, we can enable sharing:

R1 = X + 4
R2 = X + 7
R3 = X + 15

loop:
T = i + X
load [T + 4]
...
load [T + 7]
...
load [T + 15]
//===---------------------------------------------------------------------===//

It's not always a good idea to choose rematerialization over spilling. If all
the load / store instructions would be folded then spilling is cheaper because
it won't require new live intervals / registers. See 2003-05-31-LongShifts for
an example.

//===---------------------------------------------------------------------===//

With a copying garbage collector, derived pointers must not be retained across
collector safe points; the collector could move the objects and invalidate the
derived pointer. This is bad enough in the first place, but safe points can
crop up unpredictably. Consider:

        %array = load { i32, [0 x %obj] }** %array_addr
        %nth_el = getelementptr { i32, [0 x %obj] }* %array, i32 0, i32 %n
        %old = load %obj** %nth_el
        %z = div i64 %x, %y
        store %obj* %new, %obj** %nth_el

If the i64 division is lowered to a libcall, then a safe point will (must)
appear for the call site. If a collection occurs, %array and %nth_el no longer
point into the correct object.

The fix for this is to copy address calculations so that dependent pointers
are never live across safe point boundaries. But the loads cannot be copied
like this if there was an intervening store, so may be hard to get right.

Only a concurrent mutator can trigger a collection at the libcall safe point.
So single-threaded programs do not have this requirement, even with a copying
collector. Still, LLVM optimizations would probably undo a front-end's careful
work.

//===---------------------------------------------------------------------===//

The ocaml frametable structure supports liveness information. It would be good
to support it.

//===---------------------------------------------------------------------===//

The FIXME in ComputeCommonTailLength in BranchFolding.cpp needs to be
revisited. The check is there to work around a misuse of directives in inline
assembly.

//===---------------------------------------------------------------------===//

It would be good to detect collector/target compatibility instead of silently
doing the wrong thing.

//===---------------------------------------------------------------------===//

It would be really nice to be able to write patterns in .td files for copies,
which would eliminate a bunch of explicit predicates on them (e.g. no side
effects).  Once this is in place, it would be even better to have tblgen
synthesize the various copy insertion/inspection methods in TargetInstrInfo.

//===---------------------------------------------------------------------===//

Stack coloring improvements:

1. Do proper LiveStacks analysis on all stack objects including those which are
   not spill slots.
2. Reorder objects to fill in gaps between objects.
   e.g. 4, 1, <gap>, 4, 1, 1, 1, <gap>, 4 => 4, 1, 1, 1, 1, 4, 4

//===---------------------------------------------------------------------===//

The scheduler should be able to sort nearby instructions by their address. For
example, in an expanded memset sequence it's not uncommon to see code like this:

  movl $0, 4(%rdi)
  movl $0, 8(%rdi)
  movl $0, 12(%rdi)
  movl $0, 0(%rdi)

Each of the stores is independent, and the scheduler is currently making an
arbitrary decision about the order.

//===---------------------------------------------------------------------===//

Another opportunitiy in this code is that the $0 could be moved to a register:

  movl $0, 4(%rdi)
  movl $0, 8(%rdi)
  movl $0, 12(%rdi)
  movl $0, 0(%rdi)

This would save substantial code size, especially for longer sequences like
this. It would be easy to have a rule telling isel to avoid matching MOV32mi
if the immediate has more than some fixed number of uses. It's more involved
to teach the register allocator how to do late folding to recover from
excessive register pressure.