clang-p2996

Author	SHA1	Message	Date
Matthias Springer	6040044f2f	[mlir][vector] VectorToSCF: Omit redundant out-of-bounds check There was a bug in `TransferWriteNonPermutationLowering`, a pattern that extends the permutation map of a TransferWriteOp with leading transfer dimensions of size ones. These newly added transfer dimensions are always in-bounds, because the starting point of any dimension is in-bounds. VectorToSCF inserts out-of-bounds checks based on the "in_bounds" attribute and dims that are marked as out-of-bounds but that are actually always in-bounds lead to unnecessary "scf.if" ops. Differential Revision: https://reviews.llvm.org/D155196	2023-07-14 09:50:37 +02:00
Andrzej Warzynski	60c9d2993b	[mlir][vector] Refine diagnostic messages Clarify a few diagnostics so that they are more consistent with the corresponding condition. For example: ``` if (positionAttr.size() > static_cast<unsigned>(getSourceVectorType().getRank())) ``` should lead to ("no greater than"): ``` return emitOpError( "expected position attribute of rank no greater than vector rank"); ``` as opposed to ("smaller"): ``` return emitOpError( "expected position attribute of rank smaller than vector rank"); ``` Differential Revision: https://reviews.llvm.org/D154998	2023-07-12 07:42:46 +01:00
Quinn Dawkins	5b6b2caf3c	[mlir][vector] Handle memory space conflicts in VectorTransferSplit patterns Currently the transfer splitting patterns will generate an invalid cast when the source memref for a transfer op has a non-default memory space. This is handled by first introducing a `memref.memory_space_cast` in such cases. Differential Revision: https://reviews.llvm.org/D154515	2023-07-11 22:58:23 -04:00
yzhang93	9a7677d8ee	[mlir] Narrow bitwidth emulation for vector.load This patch is a following for the previous patch https://reviews.llvm.org/D151519. With this patch, vector.load op with narrow bitwidth (e.g., i4) can be converted to supported wider bitwidth (e.g., i8). Reviewed By: hanchung Differential Revision: https://reviews.llvm.org/D154178	2023-07-11 13:38:15 -07:00
Diego Caballero	51ef80a7c2	[mlir][Vector] Add support for 0-D vectors to vector.insert/extract This is part of the process to remove vector.insertelement/extractelement from the Vector dialect. RFC: https://discourse.llvm.org/t/rfc-psa-remove-vector-extractelement-and-vector-insertelement-ops-in-favor-of-vector-extract-and-vector-insert-ops Differential Revision: https://reviews.llvm.org/D152644	2023-07-11 19:28:16 +00:00
Matthias Springer	867afe5e53	[mlir][vector] Remove duplicate tensor subset <-> vector transfer patterns Remove patterns that fold tensor subset ops into vector transfer ops from the vector dialect. These patterns already exist in the tensor dialect. Differential Revision: https://reviews.llvm.org/D154932	2023-07-11 11:12:29 +02:00
Matthias Springer	a7a5641bdc	[mlir][vector] Fix bug in `TransferWriteNonPermutationLowering` This pattern expands the rank of the vector. However, the rank of the mask was not expanded. Differential Revision: https://reviews.llvm.org/D154849	2023-07-10 17:21:03 +02:00
Matthias Springer	4106557a28	[mlir][transform] Improve transform.get_closest_isolated_parent * Rename op to `transform.get_parent_op` * Match parents by "is isolated from above" and/or op name, or just the direct parent. * Deduplication of result payload ops is optional. Differential Revision: https://reviews.llvm.org/D154071	2023-07-04 16:23:08 +02:00
Andrzej Warzynski	79c83e12c8	[mlir][VectorType] Allow arbitrary dimensions to be scalable At the moment, only the trailing dimensions in the vector type can be scalable, i.e. this is supported: vector<2x[4]xf32> and this is not allowed: vector<[2]x4xf32> This patch extends the vector type so that arbitrary dimensions can be scalable. To this end, an array of bool values is added to every vector type to denote whether the corresponding dimensions are scalable or not. For example, for this vector: vector<[2]x[3]x4xf32> the following array would be created: {true, true, false}. Additionally, the current syntax: vector<[2x3]x4xf32> is replaced with: vector<[2]x[3]x4xf32> This is primarily to simplify parsing (this way, the parser can easily process one dimension at a time rather than e.g. tracking whether "scalable block" has been entered/left). NOTE: The `isScalableDim` parameter of `VectorType` (introduced in this patch) makes `numScalableDims` redundant. For the time being, `numScalableDims` is preserved to facilitate the transition between the two parameters. `numScalableDims` will be removed in one of the subsequent patches. This change is a part of a larger effort to enable scalable vectorisation in Linalg. See this RFC for more context: * https://discourse.llvm.org/t/rfc-scalable-vectorisation-in-linalg/ Differential Revision: https://reviews.llvm.org/D153372	2023-06-27 19:21:59 +01:00
Matthias Springer	1826c728cf	[mlir][transform] SequenceOp: Top-level operations can be used as matchers As a convenience to the user, top-level sequence ops can optionally be used as matchers: the op type is specified by the type of the block argument. This is similar to how pass pipeline targets can be specified on the command line (`-pass-pipeline='builtin.module(func.func(...))`). Differential Revision: https://reviews.llvm.org/D153121	2023-06-19 09:06:18 +02:00
Matthias Springer	726d076784	[mlir][transform] ApplyPatternsOp: Add check to prevent modifying the transform IR Add an extra check to make sure that transform IR is not getting modified by this op while it is being interpreted. This generally dangerous and we may want to enforce this for all transform ops that modify the payload in the future. Users should generally try to apply patterns only to the piece of IR where it is needed (e.g., a matched function) and not the entire module (which may contain the transform IR). This revision is in response to a crash in a downstream compiler that was caused by a dead `transform.structured.match` op that was removed by the GreedyPatternRewriteDriver's DCE while the enclosing sequence was being interpreted. Differential Revision: https://reviews.llvm.org/D153113	2023-06-19 08:58:40 +02:00
Andrzej Warzynski	4d339ec91e	[mlir][Vector] Add pattern to reorder elementwise and broadcast ops The new pattern will replace elementwise(broadcast) with broadcast(elementwise) when safe. This change affects tests for vectorising nD-extract. In one case ("vectorize_nd_tensor_extract_with_tensor_extract") I just trimmed the test and only preserved the key parts (scalar and contiguous load from the original Op). We could do the same with some other tests if that helps maintainability. Differential Revision: https://reviews.llvm.org/D152812	2023-06-15 10:13:41 +01:00
Matthias Springer	80853a1673	[mlir][vector][bufferize] Better analysis for vector.transfer_write The destination operand does not bufferize to a memory read if it is completely overwritten. Differential Revision: https://reviews.llvm.org/D152823	2023-06-14 09:38:51 +02:00
Matthias Springer	faae4d5d81	[mlir][vector][transform] Expose tensor slice -> transfer folding patterns Add a new transform op to populate patterns: ApplyFoldTensorSliceIntoTransferPatternsOp. Differential Revision: https://reviews.llvm.org/D152531	2023-06-09 16:23:25 +02:00
Matthias Springer	8abbc96885	[mlir][vector][transform] Drop redundant "apply_" from op names This is for consistency with the other ops. Differential Revision: https://reviews.llvm.org/D152355	2023-06-08 09:00:12 +02:00
Quentin Colombet	1dd00d3903	[mlir][Vector] Fix a propagation bug with broadcast In the vector distribute patterns, we used to move `vector.broadcast`s out of `vector.warp_execute_on_lane0`s irrespectively of how they were defined. This could create broadcast operations with invalid semantic. E.g., ``` %r = warop ...[32] ... -> vector<1x2xf32> { %val = broadcast %in : vector<64xf32> to vetor<1x64xf32> vector.yield %val : vector<1x64xf32> } ``` => ``` %r = warop ...[32] ... -> vector<64xf32> { vector.yield %in : vector<64xf32> } // Broadcasting to a narrower type! broadcast %r : vector<64xf32> to vector<1x2xf32> ``` The root issue is we are trying to broadcast something that is not the same for each thread, so there is actually nothing to propagate here. The fix checks that the broadcast we want to create actually makes sense. Differential Revision: https://reviews.llvm.org/D152154	2023-06-06 16:40:15 +02:00
Matthias Springer	cc7f52432b	[mlir][transform] Use separate ops instead of PatternRegistry * Remove `transform::PatternRegistry`. * Add a new op for each currently registered pattern set. * Change names of vector dialect pattern selector ops, so that they are consistent with the remaining code base. * Remove redundant `transform.vector.extract_address_computations` op. Differential Revision: https://reviews.llvm.org/D152249	2023-06-06 11:53:03 +02:00
Matthias Springer	976d25ed9e	[mlir][vector] Use transform.apply_patterns in vector tests All vector transform ops are now `PatternDescriptorOpInterface` ops that merely select the patterns. The patterns are applied by the `apply_patterns` op. This is to ensure that ops are properly tracked. (TrackingListener is used in the implementation of `apply_patterns`.) Furthermore, handles are no longer invalidated when applying patterns in the vector tests. Differential Revision: https://reviews.llvm.org/D152174	2023-06-06 09:27:05 +02:00
Manish Gupta	9a795f0c59	[mlir][Vector] Adds a pattern to fold `arith.extf` into `vector.contract` Consider mixed precision data type, i.e., F16 input lhs, F16 input rhs, F32 accumulation, and F32 output. This is typically written as F32 <= F16F16 + F32. During vectorization from linalg to vector for mixed precision data type (F32 <= F16F16 + F32), linalg.matmul introduces arith.extf on input lhs and rhs operands. "linalg.matmul"(%lhs, %rhs, %acc) ({ ^bb0(%arg1: f16, %arg2: f16, %arg3: f32): %lhs_f32 = "arith.extf"(%arg1) : (f16) -> f32 %rhs_f32 = "arith.extf"(%arg2) : (f16) -> f32 %mul = "arith.mulf"(%lhs_f32, %rhs_f32) : (f32, f32) -> f32 %acc = "arith.addf"(%arg3, %mul) : (f32, f32) -> f32 "linalg.yield"(%acc) : (f32) -> () }) There are backend that natively supports mixed-precision data type and does not need the arith.extf. For example, NVIDIA A100 GPU has mma.sync.aligned.*.f32.f16.f16.f32 that can support mixed-precision data type. However, the presence of arith.extf in the IR, introduces the unnecessary casting targeting F32 Tensor Cores instead of F16 Tensor Cores for NVIDIA backend. This patch adds a folding pattern to fold arith.extf into vector.contract Differential Revision: https://reviews.llvm.org/D151918	2023-06-05 23:22:20 +00:00
Quentin Colombet	018d8ac974	[mlir][Vector] Fix a propagation bug with transfer_read In the vector distribute patterns, we used to move `vector.transfer_read`s out of `vector.warp_execute_on_lane0`s irrespectively of how they were defined. This could create transfer_read operations that would read values from within the warpOp's body from outside of the body. E.g., ``` warpop { %defined_in_body %read = transfer_read %defined_in_body vector.yield %read } ``` => ``` warpop { %defined_in_body vector.yield ... } // %defined_in_body is referenced outside of its scope. %read = transfer_read %defined_in_body ``` The fix consists in checking that all the values feeding the new `transfer_read` are defined outside of warpOp's body. Note: We could do this check before creating any operation, but that would mean knowing what `affine::makeComposedAffineApply` actually do. So the current fix is a trade off of coupling the implementations of this propagation and `makeComposedAffineApply` versus compile time. Differential Revision: https://reviews.llvm.org/D152149	2023-06-05 15:52:26 +02:00
Diego Caballero	834fcfed24	Reland "[mlir][Vector] Extend xfer drop unit dim patterns" This reverts commit `76d71f3792`.	2023-06-01 22:22:16 +00:00
Diego Caballero	d3e1398bef	[mlir][Vector] Prevent vector-to-scalar xfer patterns from triggering on sub-vectors Patterns that convert extract(transfer_read) into a scalar load where incorrectly triggering for cases where a sub-vector instead of a scalar was extracted. Reviewed By: nicolasvasilache, hanchung, awarzynski Differential Revision: https://reviews.llvm.org/D151862	2023-06-01 22:22:16 +00:00
Diego Caballero	0935c0556b	[mlir][Vector] Add support for 0-D 'vector.shape_cast' lowering This PR adds support for shape casting from and to 0-D vectors. Reviewed By: nicolasvasilache, hanchung, awarzynski Differential Revision: https://reviews.llvm.org/D151851	2023-06-01 22:22:16 +00:00
Diego Caballero	45b25d24f0	[mlir][Vector] Disable 'vector.extract' folding for unsupported 0-D vectors The `vector.extract` folding patterns do not support 0-D vectors (actually, 0-D vector support couldn't even be implemented as a folding pattern as it would require replacing `vector.extract` with a `vector.extractelement` op). This patch is bailing out folding when 0-D vectors are found. Reviewed By: nicolasvasilache, hanchung Differential Revision: https://reviews.llvm.org/D151847	2023-06-01 22:22:15 +00:00
Diego Caballero	76d71f3792	Revert "[mlir][Vector] Extend xfer drop unit dim patterns" This reverts commit `a53cd03dea`. This commit is exposing some implementation gaps in other patterns. Reverting for now.	2023-05-31 18:20:05 +00:00
Diego Caballero	a53cd03dea	[mlir][Vector] Extend xfer drop unit dim patterns This patch extends the transfer drop unit dim patterns to support cases where the vector shape should also be reduced (e.g., transfer_read(memref<1x4x1xf32>, vector<1x4x1xf32>) -> transfer_read(memref<4xf32>, vector<4xf32>). Reviewed By: hanchung, pzread Differential Revision: https://reviews.llvm.org/D151007	2023-05-23 20:58:51 +00:00
Diego Caballero	7ce2c3d71b	[mlir][Vector] Add 0-d vector support to 'vector.shape_cast` This patch adds support to shape cast a vector<1x1x1...1xElemenType> to a vector<ElementType> and the other way around. Differential Revision: https://reviews.llvm.org/D151169	2023-05-23 17:22:55 +00:00
Alex Zinenko	9813c184f7	[mlir] NFC: use !transform.any_op in relevant tests Update various tests using Transform dialect extensions to pervasively use `!transform.any_op` instead of `!pdl.operation`. Tests are sometimes used as source of knowledge for best practices and these were doing the opposite of what is considered best practices per https://discourse.llvm.org/t/rfc-type-system-for-the-transform-dialect/65702. Reviewed By: nicolasvasilache Differential Revision: https://reviews.llvm.org/D150785	2023-05-22 08:19:46 +00:00
Diego Caballero	14726cd691	[mlir][Vector] Extend xfer_read(extract)->scalar load to support multiple uses This patch extends the vector.extract(vector.transfer_read) -> scalar load patterns to support vector.transfer_read with multiple uses. For now, we check that all the uses are vector.extract operations. Supporting multiple uses is predicated under a flag. Reviewed By: hanchung Differential Revision: https://reviews.llvm.org/D150812	2023-05-19 21:03:18 +00:00
Diego Caballero	0c71a6e7c8	[mlir][Vector] Add folding for vector.mask with all-true masks This patch removes `vector.mask` operations with all-true masks (i.e., all lanes enabled). Reviewed By: hanchung Differential Revision: https://reviews.llvm.org/D150743	2023-05-18 19:07:07 +00:00
Lei Zhang	e000b62a34	[mlir][vector] Separate out vector transfer + tensor slice patterns These patterns touches the structure generated from tiling so it affects later steps like bufferization and vector hoisting. Instead of putting them in canonicalization, this commit creates separate entry points for them to be called explicitly. This is NFC regarding the functionality and tests of those patterns. It also addresses two TODO items in the codebase. Reviewed By: ThomasRaoux Differential Revision: https://reviews.llvm.org/D150702	2023-05-17 09:01:19 -07:00
Alex Zinenko	2fe4d90cac	[mlir] make structured transform ops use types Types have been introduced a while ago and provide for better readability and transform-time verification. Use them in the ops from the structured transform dialect extension. In most cases, the types are appended as trailing functional types or a derived format of the functional type that allows for an empty right hand size without the annoying `-> ()` syntax (similarly to `func.func` declaration that may omit the arrow). When handles are used inside mixed static/dynamic lists, such as tile sizes, types of those handles follow them immediately as in `sizes [%0 : !transform.any_value, 42]`. This allows for better readability than matching the trailing type. Update code to remove hardcoded PDL dependencies and expunge PDL from structured transform op code. Reviewed By: nicolasvasilache Differential Revision: https://reviews.llvm.org/D144515	2023-05-16 08:16:56 +00:00
Hanhan Wang	25cc5a71b3	[mlir][vector] Generalize vector.transpose lowering to n-D vectors The existing vector.transpose lowering patterns only triggers if the input vector is 2D. The revision extends the pattern to handle n-D vectors which are effectively 2-D vectors (e.g., vector<1x4x1x8x1). It refactors a common check about 2-D vectors from X86Vector lowering to VectorUtils.h so it can be reused by both sides. Reviewed By: dcaballe Differential Revision: https://reviews.llvm.org/D149908	2023-05-08 10:48:26 -07:00
Quinn Dawkins	650f04feda	[mlir][vector] Add pattern to break down vector.bitcast The pattern added here is intended as a last resort for targets like SPIR-V where there are vector size restrictions and we need to be able to break down large vector types. Vectorizing loads/stores for small bitwidths (e.g. i8) relies on bitcasting to a larger element type and patterns to bubble bitcast ops to where they can cancel. This fails for cases such as ``` %1 = arith.trunci %0 : vector<2x32xi32> to vector<2x32xi8> vector.transfer_write %1, %destination[%c0, %c0] {in_bounds = [true, true]} : vector<2x32xi8>, memref<2x32xi8> ``` where the `arith.trunci` op essentially does the job of one of the bitcasts, leading to a bitcast that need to be further broken down ``` vector.bitcast %0 : vector<16xi8> to vector<4xi32> ``` Differential Revision: https://reviews.llvm.org/D149065	2023-04-25 20:18:02 -04:00
Quinn Dawkins	435f7d4c2e	[mlir][vector] Add unroll pattern for vector.gather This pattern is useful for SPIR-V to unroll to a supported vector size before later lowerings. The unrolling pattern is closer to an elementwise op than the transfer ops because the index values from which to extract elements are captured by the index vector and thus there is no need to update the base offsets when unrolling gather. Differential Revision: https://reviews.llvm.org/D149066	2023-04-24 14:02:59 -04:00
Hanhan Wang	8d163e5045	[mlir][Vector] Add 16x16 strategy to vector.transpose lowering. It adds a `shuffle_16x16` strategy LowerVectorTranspose and renames `shuffle` to `shuffle_1d`. The idea is similar to 8x8 cases in x86Vector::avx2. The general algorithm is: ``` interleave 32-bit lanes using 8x _mm512_unpacklo_epi32 8x _mm512_unpackhi_epi32 interleave 64-bit lanes using 8x _mm512_unpacklo_epi64 8x _mm512_unpackhi_epi64 permute 128-bit lanes using 16x _mm512_shuffle_i32x4 permute 256-bit lanes using again 16x _mm512_shuffle_i32x4 ``` After the first stage, they got transposed to ``` 0 16 1 17 4 20 5 21 8 24 9 25 12 28 13 29 2 18 3 19 6 22 7 23 10 26 11 27 14 30 15 31 32 48 33 49 ... 34 50 35 51 ... 64 80 65 81 ... ... ``` After the second stage, they got transposed to ``` 0 16 32 48 ... 1 17 33 49 ... 2 18 34 49 ... 3 19 35 51 ... 64 80 96 112 ... 65 81 97 114 ... 66 82 98 113 ... 67 83 99 115 ... ... ``` After the thrid stage, they got transposed to ``` 0 16 32 48 8 24 40 56 64 80 96 112 ... 1 17 33 49 ... 2 18 34 50 ... 3 19 35 51 ... 4 20 36 52 ... 5 21 37 53 ... 6 22 38 54 ... 7 23 39 55 ... 128 144 160 176 ... 129 145 161 177 ... ... ``` After the last stage, they got transposed to ``` 0 16 32 48 64 80 96 112 ... 240 1 17 33 49 66 81 97 113 ... 241 2 18 34 50 67 82 98 114 ... 242 ... 15 31 47 63 79 96 111 127 ... 255 ``` Reviewed By: dcaballe Differential Revision: https://reviews.llvm.org/D148685	2023-04-23 11:05:41 -07:00
Benjamin Kramer	37a867a5a8	[vector] When trimming leading insertion dimensions, base the final result on the ranks This was incorrect when the number of dropped source dims was smaller than the number of dropped dst dims. We still need to insert zeros if there is anything dropped from the src. Differential Revision: https://reviews.llvm.org/D148636	2023-04-18 18:49:29 +02:00
Lei Zhang	5041fe8439	[mlir][vector] Fix integer promotion type mismatch We need to create a new type with transposed shape after transposing the operand in `CanonicalizeContractMatmulToMMT`. Reviewed By: kuhar, dcaballe Differential Revision: https://reviews.llvm.org/D148470	2023-04-17 11:29:23 -07:00
Nicolas Vasilache	e4e0bf63d0	[mlir][Vector] Split transform.vector.lower_mask in 2 ops. This gives us better control to lower masked operations independently of the create mask operations. It is often useful to maintain high-level mask information instead of lowering it too early to too fine-grained form. Differential Revision: https://reviews.llvm.org/D148162	2023-04-13 00:14:01 -07:00
Matthias Springer	3f7959ea3d	[mlir][bufferize] Simplify one_shot_bufferize transform op Restrict the op to functions and modules. Such ops are modified in-place. The transform now consumes the handle and produces a new handle. The `target_is_module` attribute is no longer needed because a result handle is produced in either case. Differential Revision: https://reviews.llvm.org/D147446	2023-04-06 12:59:35 +09:00
tyb0807	942b403ff1	[mlir] Fix casting of leading unit dims for vector.insert When dropping leading unit dims of vector.insert's operands and creating a new vector.insert, its new position rank should be computed explicitly in two steps: first based on the numbers of leading unit dims dropped from the vector.insert's destination, then based on the numbers of leading unit dims dropped from its source. Reviewed By: pifon2a Differential Revision: https://reviews.llvm.org/D147280	2023-03-31 12:12:35 +00:00
Diego Caballero	1cd434d007	[mlir][Vector] Add canonicalization pattern for vector.transpose(vector.constant_mask) We already had vector.transpose(vector.create_mask) -> vector.create_mask. This patch adds the constant mask version of it. Reviewed By: nicolasvasilache Differential Revision: https://reviews.llvm.org/D147099	2023-03-29 19:53:29 +00:00
Diego Caballero	7b70baa9ef	[mlir][Vector] Remove lhs and rhs masks from vector.contract This patch removes the historical lhs and rhs masks in vector.contract, now that vector.mask supports vector.contract and the lhs and rhs masks are barely supported by all the vector.contract lowerings and transformations. Reviewed By: nicolasvasilache Differential Revision: https://reviews.llvm.org/D144430	2023-03-29 19:53:29 +00:00
Nicolas Vasilache	8b51340740	[mlir][Vector][Transforms] Improve the control over individual vector lowerings and transforms This revision adds vector transform operations that allow us to better inspect the composition of various lowerings that were previously very opaque. This commit is NFC in that it does not change patterns beyond adding `rewriter.notifyFailure` messages and it does not change the tests beyond breaking them into pieces and using transforms instead of throwaway opaque test passes. Reviewed By: ftynse, springerm Co-authored-by: Alex Zinenko <zinenko@google.com> Differential Revision: https://reviews.llvm.org/D146755	2023-03-24 14:01:39 +00:00
Nicolas Vasilache	73bec2b2c3	[mlir][Vector] Retire one old filter-based test Differential Revision: https://reviews.llvm.org/D146742	2023-03-23 11:00:35 -07:00
Adam Paszke	61f33def13	[mlir][Vector] Make sure that vector.contract preserves extra attributes while parsing The old implementation parsed the optional attribute dict, only to replace its contents by using `assign`. Reviewed By: ftynse Differential Revision: https://reviews.llvm.org/D146707	2023-03-23 10:31:46 +00:00
Nicolas Vasilache	015cd84d3c	Revert "[mlir][Linalg][Transform] Avoid FunctionalStyleTransformOpTrait where unnecesseary to improve usability" This reverts commit `31aa8ea252`. This is currently not in a good state as we have some footguns due to missing listeners.	2023-03-20 07:07:27 -07:00
Nicolas Vasilache	31aa8ea252	[mlir][Linalg][Transform] Avoid FunctionalStyleTransformOpTrait where unnecesseary to improve usability Differential Revision: https://reviews.llvm.org/D146305	2023-03-20 03:17:44 -07:00
Diego Caballero	6d68ef4e38	[mlir][Vector] Canonicalize create_mask(transpose) When applying vector masking we may create a mask and then transpose it. Transpositions are extremely expensive so this patch introduces a new canonicalization pattern that remove the tranpose operation and create a new transposed mask. Differential Revision: https://reviews.llvm.org/D146193	2023-03-16 14:35:52 +00:00
Jakub Kuderski	a586c55100	[mlir][vector] Add mask fold test for gather lowering Check that `scf.if` checks are folded when the mask is all set / not set. This to address post-commit feedback for https://reviews.llvm.org/D145942. Reviewed By: dcaballe Differential Revision: https://reviews.llvm.org/D146144	2023-03-16 10:16:15 -04:00

1 2 3 4 5 ...

411 Commits