llvm-project

Commit Graph

Author	SHA1	Message	Date
David Green	3651635eca	[ARM][DAG] BF16 constant handling. Much like f16 and f32, we shouldn't try to shrink bf16 to smaller fp constant. The code may not be optimal, but this allows us to legalize bf16 constants under Arm without errors.	2022-10-02 11:51:08 +01:00
Yeting Kuo	cefb7aab61	[VP][RISCV] Add vp.copysign and RISC-V support. Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134935	2022-10-01 10:19:10 +08:00
Eric Wang	5b26f4f042	Reland "[MLGO] ML Regalloc Priority Advisor" This relands commit `8f4f26ba5b`, which was reverted in `91c96a806c` because of Buildbot failures. The previous model test is not compatible with tflite. e.g. https://lab.llvm.org/buildbot/#/builders/6/builds/14041 Differential Revision: https://reviews.llvm.org/D133616	2022-09-30 16:27:26 -05:00
Guozhi Wei	feea3b990e	[LiveRangeEdit] Add a statistic variable for rematerialization Add a statistic variable for rematerialization. Differential Revision: https://reviews.llvm.org/D134907	2022-09-30 19:39:51 +00:00
Simon Pilgrim	61dc5014ac	[DAG] Update foldSelectWithIdentityConstant to use llvm::isNeutralConstant D133866 added the llvm::isNeutralConstant helper to track neutral/passthrough constants This patch updates foldSelectWithIdentityConstant to use the helper instead of maintaining its own opcode handling Differential Revision: https://reviews.llvm.org/D134966	2022-09-30 17:46:52 +01:00
Serge Pavlov	b3913a9cdf	[GlobalISel] Do not crash on widening vector result Function buildCopyToRegs did not handle properly the case when it should make wider vector result. It happened, for example, in a function that returns value of type <2 x f32>, which should be widen to <4 x f32> to fit XMM register. The function eventually calls MachineIRBuilder.buildUnmerge, which does not expect that only one destination register is specified. Now this case is treated specifically in buildCopyToRegs. Differential Revision: https://reviews.llvm.org/D128546	2022-09-30 21:30:55 +07:00
Pierre van Houtryve	7388520d1c	[GISel] Add more cases to isKnownNeverNaN Make it even with the DAG implementation as of D134854 Reviewed By: arsenm, foad Differential Revision: https://reviews.llvm.org/D134857	2022-09-30 14:10:56 +00:00
Pierre van Houtryve	653beae5a1	[AMDGPU][GISel] Add Identity BUILD_VECTOR Combines Folds-away BUILD_VECTOR-related noops in the post-legalizer combiner. Depends on D134433 Reviewed By: arsenm Differential Revision: https://reviews.llvm.org/D134953	2022-09-30 14:07:13 +00:00
Amaury Séchet	031a7ad575	[NFC] Fix erroneous indentation.	2022-09-30 12:30:27 +00:00
Yeting Kuo	1cc02b05b7	[SelectionDAG] Add helper function to check whether a SDValue is neutral element. NFC. Using this helper makes work about neutral elements more easier. Although I only find one case now, I think it will have more chance to be used since so many combine works are related to neutral elements. Reviewed By: RKSimon Differential Revision: https://reviews.llvm.org/D133866	2022-09-30 11:29:11 +08:00
Mircea Trofin	91c96a806c	Revert "[MLGO] ML Regalloc Priority Advisor" This reverts commit `8f4f26ba5b`. Buildbot failures, e.g. https://lab.llvm.org/buildbot/#/builders/6/builds/14041	2022-09-29 18:26:40 -07:00
Amaury Séchet	923909afbe	[DAG] Simplify the select of constant combine code. NFC	2022-09-30 01:03:14 +00:00
Amaury Séchet	d7600c7ccb	[DAG] select Cond, C, -1 --> or (sext (not Cond)), C when C is MVT::i1 In the spirit of D130765 . Get rid of cbranches and/or cmov. Usually shorter, but sometime not, becaus eit's hard to prededict when dependency breaking xor will be introduced. Reviewed By: RKSimon Differential Revision: https://reviews.llvm.org/D134736	2022-09-30 00:36:58 +00:00
Eric Wang	8f4f26ba5b	[MLGO] ML Regalloc Priority Advisor The bulk of the implementation is common between 'release' mode (==AOT-ed model) and 'development' mode (for training), the main difference is that in development mode, we may also log features (for training logs), inject scoring information and then produce the log file. Differential Revision: https://reviews.llvm.org/D133616	2022-09-29 16:55:15 -05:00
Thomas Symalla	a41dde2c62	[AMDGPU] Add use check in v_fma combine. In D132837, an existing v_fma combine was extended to regard nested fma instructions. Originally, the inner FMA was checked for being used only once. In its current state, this check is missing, which causes some regressions. In this patch, this check was added. Reviewed By: foad Differential Revision: https://reviews.llvm.org/D134856	2022-09-29 12:25:03 +02:00
Jessica Paquette	1eb49bbab6	[GlobalISel][CallLowering] Use hasRetAttr for return flags on CallBases Given something like this: ``` declare signext i16 @signext_callee() define i32 @caller() { %res = call i16 @signext_callee() ... } ``` CallLowering would miss that signext_callee's return value is sign extended, because it isn't on the call. Use hasRetAttr on the CallBase to allow us to catch this. (This now inserts G_ASSERT_SEXT/G_ASSERT_ZEXT like in the original review.) Differential Revision: https://reviews.llvm.org/D86228	2022-09-28 19:38:24 -07:00
Jessica Paquette	704b2e162c	[GlobalISel] Add isConstFalseVal helper to Utils Add a utility function which returns true if the given value is a constant false value. This is necessary to port one of the compare simplifications in TargetLowering::SimplifySetCC. Differential Revision: https://reviews.llvm.org/D91754	2022-09-28 15:44:26 -07:00
serge-sans-paille	16544cbe64	[iwyu] Move <cmath> out of llvm/Support/MathExtras.h Interestingly, MathExtras.h doesn't use <cmath> declaration, so move it out of that header and include it when needed. No functional change intended, but there's no longer a transitive include fromMathExtras.h to cmath.	2022-09-28 20:49:01 +02:00
Aiden Grossman	8d77f8fde7	[MLGO] Add per-instruction MBB frequencies to regalloc dev features This commit adds in two new features to the ML regalloc eviction analysis that can be used in ML models, a vector of MBB frequencies and a vector of indicies mapping instructions to their corresponding basic blocks. This will allow for further experimentation with per-instruction features and give a lot more flexibility for future experimentation over how we're extracting MBB frequency data currently. Reviewed By: mtrofin, jacobhegna Differential Revision: https://reviews.llvm.org/D134166	2022-09-28 18:45:04 +00:00
Jay Foad	2c12a04bba	[ISel] Fix DAG divergence after new FMA combine D132837 introduced a new DAG combine that used MorphNodeTo to morph an FMUL into an FMA. It turns out that MorphNodeTo does not properly update the divergence bit for users of the morphed node, causing an assertion failure on the new test case: llc: SelectionDAG.cpp:10486: void llvm::SelectionDAG::VerifyDAGDivergence(): Assertion `calculateDivergence(N) == N->isDivergent() && "Divergence bit inconsistency detected"' failed. Fixing MorphNodeTo to propagate the divergence bit is tricky because of the way it is used to select machine instructions, so use getNode and ReplaceAllUsesOfValueWith instead. Differential Revision: https://reviews.llvm.org/D134810	2022-09-28 19:41:51 +01:00
Craig Topper	12357e88af	[RISCV][SelectionDAGBuilder] Fix crash when copying a v1f32 vector between basic blocks. On a rv64 without f32 or vector support, this will be passed across the basic block as an i64. We need use i32 as an intermediate type with bitcast and anyext/trunc. Fixes PR58025 Reviewed By: RKSimon Differential Revision: https://reviews.llvm.org/D134758	2022-09-28 10:13:35 -07:00
Matt Arsenault	a61c3455c0	AtomicExpand: Use llvm.ptrmask instead of ptrtoint This removes the ptrtoint from the load's pointer operand, although we can't entirely eliminate these to get the LSB shift. In a future patch, this will avoid ptrtoint in the case where the atomic is overaligned to the word size.	2022-09-28 12:51:30 -04:00
Matt Devereau	0a4771a7e8	[AArch64][SVE] Expand gather index to 32 bits instead of 64 bits For gathers which load in 8 and 16 bit data then use that data as an index, the index can be extended to 32 bits instead of 64 bits Differential Revision: https://reviews.llvm.org/D130692	2022-09-28 14:42:12 +00:00
jacquesguan	465ac0b96e	[LegalizeTypes] Use getVectorElementCount to avoid crash of scalable vector. Reviewed By: RKSimon Differential Revision: https://reviews.llvm.org/D134718	2022-09-28 17:31:29 +08:00
eopXD	9677d70eb2	[VP][RISCV] Add vp.floor, vp.round, vp.roundeven and their RISC-V support Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134759	2022-09-27 19:45:58 -07:00
eopXD	163cb33854	[VP][RISCV] Add vp.ceil and RISC-V support Previous commit `8b00b24f85` missed to add `int_ceil` anchor for the llvm.ceil.* section under LangRef.rst Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134586	2022-09-27 12:04:09 -07:00
eopXD	384b8b3da7	Revert "[VP][RISCV] Add vp.ceil and RISC-V support" This reverts commit `8b00b24f85`.	2022-09-27 11:12:57 -07:00
eopXD	8b00b24f85	[VP][RISCV] Add vp.ceil and RISC-V support Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134586	2022-09-27 11:08:27 -07:00
Craig Topper	a6383bb51c	[VP][RISCV] Add vp.fmuladd. Expanded in SelectionDAGBuilder similar to llvm.fmuladd. Reviewed By: frasercrmck, simoll Differential Revision: https://reviews.llvm.org/D134474	2022-09-27 10:02:37 -07:00
Amaury Séchet	d1baed7c9c	[DAG] select Cond, -1, C --> or (sext Cond), C if Cond is MVT::i1 This seems to be beneficial overall, except for midpoint-int.ll . The X86 backend seems to generate zeroing that are not necesary. Reviewed By: shchenz Differential Revision: https://reviews.llvm.org/D131260	2022-09-27 12:54:52 +00:00
Yeting Kuo	04e1301f3d	[VP][RISCV] Add vp.maxnum and vp.minnum intrinsics and RISC-V support. Add vp.maxnum and vp.minnum which are vector predicted intrinsics of llvm.maxnum and llvm.minnum. Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134639	2022-09-27 13:36:45 +08:00
Paul Scoropan	ce004fb4f2	[PowerPC] XCOFF exception section support on the direct assembler path This feature implements support for making entries in the exception section on XCOFF on the direct assembly path using the ".except" pseudo-op. It also provides functionality to lower entries (comprised of language and reason codes) into the exception section through the use of annotation metadata attached to llvm.ppc.trap/trapd/tw/tdw intrinsics. Integrated assembler support will be provided in another review. https://reviews.llvm.org/D133030 needs to merge first for LIT tests Reviewed By: shchenz, RKSimon Differential Revision: https://reviews.llvm.org/D132146	2022-09-26 22:24:20 -04:00
Mircea Trofin	8a24e0cb5a	[nfc][mlgo] Lazily compute the regalloc reward Differential Revision: https://reviews.llvm.org/D134664	2022-09-26 15:34:29 -07:00
Craig Topper	afdd600a49	[LegalizeTypes][RISCV] Support f16 in ExpandIntRes_LLROUND_LLRINT. Promote f16 to f32 and use the f32 libcall. I deleted rv64zfh-half-intrinsics-strict.ll because it only existed due to this issue breaking rv32. Differential Revision: https://reviews.llvm.org/D134579	2022-09-26 11:09:33 -07:00
Amaury Séchet	b30bbd181b	Small formating nit in DAGCombiner. NFC	2022-09-26 13:36:11 +00:00
Yeting Kuo	43c5fbdd3a	[VP][RISCV] Add vp.sqrt intrinsic and RISC-V support. The patch modeled vp.fabs patch D132793. Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D133690	2022-09-26 10:47:40 +08:00
Yi Kong	32994b7357	Make MLIR model URLs cache variables This allows us to directly use the models published on Github. Differential Revision: https://reviews.llvm.org/D134566	2022-09-23 15:21:53 -07:00
Afanasyev Ivan	d2e434c378	[RegisterCoalescer] fix dst subreg replacement during remat copy trick Instructions might use definition register as its "undef" operand. It happens on architectures with predicated executon: ``` %0:subreg = instruction op_1, ..., op_N, undef %0:subreg, op_N+2, ... ``` RegisterCoalescer should take into account all remat instruction operands during destination subregister fixup. ``` ; remat result before fix: %1 = instruction op_1, ..., op_N, undef %1:subreg, op_N+2, ... ; remat result after fix (correct): %1 = instruction op_1, ..., op_N, undef %1, op_N+2, ... ``` Differential Revision: https://reviews.llvm.org/D125657	2022-09-23 18:52:29 +00:00
Philip Reames	b9c4733079	[DAG] Move one-use add of splat to base of scatter/gather This extends the uniform base transform used with scatter/gather to support one-use vector adds-of-splats with a non-zero base. This has the effect of essentially reassociating an add from vector to scalar domain. The motivation is to improve the lowering of scatter/gather operations fed by complex geps. Differential Revision: https://reviews.llvm.org/D134472	2022-09-22 18:45:12 -07:00
Eric Wang	83c53d346f	[NFC][MLGO] Introduce logRewardIfNeeded method This patch introduces a logRewardIfNeeded method to reuse regallocscoring. Differential Revision: https://reviews.llvm.org/D134232	2022-09-22 19:22:32 -05:00
Philip Reames	60c91fd364	[RISCV] Disallow scale for scatter/gather RISCV doesn't actually support a scaled form of indexed load and store. We previously handled this by forming the scaled SDNode, and then doing custom legalization during lowering. This patch instead adds a callback via TLI to prevent formation entirely. This has two effects: * First, the GEP gets expanded (and used). Instead of the shift being created with an SDLoc of the memory operation, it has the SDLoc of the GEP instruction. This avoids the scheduler perturbing IR order when there's no reason to. * Second, we fix what appears to be a bug in index calculation with RV32. The rules for GEPs require index calculation be done in particular bitwidth, and it appears the custom legalization code got this wrong for the case where index type exceeds pointer width. (Or at least, I trust the generic GEP lowering to be correct a lot more.) The DAGCombiner change to handle VPScatter/VPGather is technically separate, but is required to prevent a regression on those intrinsics. Differential Revision: https://reviews.llvm.org/D134382	2022-09-22 15:31:26 -07:00
Craig Topper	9d236d4dab	[SelectionDAGBuilder] Simplify how VTs is created for constrained intrinsics. NFC All constrained intrinsics return a single value. We can directly convert it to an EVT instead of going through ComputeValueTypes.	2022-09-22 14:21:22 -07:00
Philip Reames	46525fee81	[DAGCombine] Check both forms of a commutative transform The transform to fold an add into the base of a scatter/gather was only checking to see if the LHS was a splat. Included test change indicates that splats are not canonicalized to LHS, and that we need to check both sides.	2022-09-22 12:21:47 -07:00
Jonathan Camilleri	08288052ae	[DebugInfo] Emit access specifiers for typedefs The accessibility level of a typedef or using declaration in a struct or class was being lost when producing debug information. Reviewed By: dblaikie Differential Revision: https://reviews.llvm.org/D134339	2022-09-22 17:08:41 +00:00
Amara Emerson	885a87033c	[GlobalISel] Enforce G_ASSERT_ALIGN to have a valid alignment > 0.	2022-09-22 16:05:07 +01:00
Matt Arsenault	94ebd7d9ff	MachineVerifier: Verify REG_SEQUENCE Somehow there was no verification of this, other than an ad-hoc assertion in TwoAddressInstructions.	2022-09-22 09:51:15 -04:00
Christudasan Devadasan	32a8260ccc	-dot-machine-cfg for printing MachineFunction to a dot file This pass allows a user to dump a MIR function to a dot file and view it as a graph. It is targeted to provide a similar functionality as -dot-cfg pass on LLVM-IR. As of now the pass also support below flags: -dot-mcfg-only [optional][won't print instructions in the graph just block name] -mcfg-dot-filename-prefix [optional][prefix to add to output dot file] -mcfg-func-name [optional] [specify function name or it's substring, handy if mir file contains multiple functions and you need to see graph of just one] More flags and details can be introduced as per the requirements in future. This pass is inspired from -dot-cfg IR pass and APIs are written in almost identical format. Patch by Yashwant Singh <Yashwant.Singh@amd.com> (yassingh) Reviewed By: arsenm Differential Revision: https://reviews.llvm.org/D133709	2022-09-22 12:48:33 +05:30
Fangrui Song	2d975f1efe	[GlobalISel] Fix std::max after D134380	2022-09-21 14:09:04 -07:00
Amara Emerson	85cd376f70	[GlobalISel] Fix known bits for G_ASSERT_ALIGN. I don't know what was going on originally with these tests. It seems reasonable to have the immediate be the same byte alignment unit as the IR, in which case we need to take the log2 in order to set the right number of low bits. This fixes a miscompile in chromium. Differential Revision: https://reviews.llvm.org/D134380	2022-09-21 21:34:05 +01:00
Guozhi Wei	c39311eb40	[RegisterCoalescer] Use LiveRangeEdit to handle rematerialization This patch uses the API provided by LiveRangeEdit to handle rematerialization. It will make future maintenance and improvement more easier. No functional change. Differential Revision: https://reviews.llvm.org/D133610	2022-09-21 17:51:07 +00:00

1 2 3 4 5 ...

33017 Commits