torch-mlir

Commit Graph

Author	SHA1	Message	Date
Rob Suderman	ec4cb8be44	Bump LLVM to llvm/llvm-project@0030fc4ac7 (#3079 ) Co-authored-by: Peiming Liu <peiming@google.com>	2024-04-01 16:34:59 -07:00
Thomas Dietert	3c33dbd987	[MLIR][Torch] Canonicalize torch.from_i1 and torch.to_i1 (#3067 ) When lowering `torch.aten.convolution`, it is expected that the 'transposed' argument is a torch.constant operation. In some cases, the argument was a `from_i1` operation converting an `arith.constant` operation into a torch.bool. This is not wrong semantically, but instead of generalizing the legality of the `torch.aten.convolution` op, we canonicalize `arith.constant` ops followed by `from_i1` ops to `torch.bool` ops. For example: ``` //===-------------------------------------------===// Legalizing operation : 'torch.aten.convolution'(0x124705b90) { %33 = "torch.aten.convolution"(%arg0, %20, %21, %31, %29, %30, %19, %32, %0) : (!torch.vtensor<[1,1,28,28],f32>, !torch.vtensor<[10,1,5,5],f32>, !torch.vtensor<[10],f32>, !torch.list<int>, !torch.list<int>, !torch.list<int>, !torch.bool, !torch.list<int>, !torch.int) -> !torch.vtensor<[1,10,24,24],f32> * Fold { } -> FAILURE : unable to fold * Pattern : 'torch.aten.convolution -> ()' { ** Failure : unimplemented: only constant transposed supported. <-- Resolved by this PR } -> FAILURE : pattern failed to match * Pattern : 'torch.aten.convolution -> ()' { ** Failure : not a supported Scalar to Tensor like op } -> FAILURE : pattern failed to match * Pattern : 'torch.aten.convolution -> ()' { ** Failure : not a supported elementwise op } -> FAILURE : pattern failed to match * Pattern : 'torch.aten.convolution -> ()' { ** Failure : not a supported reduce op } -> FAILURE : pattern failed to match } -> FAILURE : no matched legalization pattern //===-------------------------------------------===// <stdin>:21:11: error: failed to legalize operation 'torch.aten.convolution' that was explicitly marked illegal %17 = torch.operator "onnx.Conv"(%arg0, %0, %1) {torch.onnx.dilations = [1 : si64, 1 : si64], torch.onnx.group = 1 : si64, torch.onnx.kernel_shape = [5 : si64, 5 : si64], torch.onnx.pads = [0 : si64, 0 : si64, 0 : si64, 0 : si64], torch.onnx.strides = [1 : si64, 1 : si64]} : (!torch.vtensor<[1,1,28,28],f32>, !torch.vtensor<[10,1,5,5],f32>, !torch.vtensor<[10],f32>) -> !torch.vtensor<[1,10,24,24],f32> ^ <stdin>:21:11: note: see current operation: %33 = "torch.aten.convolution"(%arg0, %20, %21, %31, %29, %30, %19, %32, %0) : (!torch.vtensor<[1,1,28,28],f32>, !torch.vtensor<[10,1,5,5],f32>, !torch.vtensor<[10],f32>, !torch.list<int>, !torch.list<int>, !torch.list<int>, !torch.bool, !torch.list<int>, !torch.int) -> !torch.vtensor<[1,10,24,24],f32> ``` Additionally, we require the canonicalization of `to_i1` operating on a torch.constant bool to an `arith.constant ... : i1` for the e2e tests to pass successfully.	2024-04-01 14:25:51 -07:00
Stella Laurenzo	826786bdd0	[fx] Support ExportedProgram buffer mutation. (#3080 ) In the prior state when I supported mutation of user inputs by treating them as mutable-tensor SSA values, I had left the case of buffer mutation only vaguely implemented until a concrete use emerged. This patch reworks this buffer mutation support by assuming that buffers must be resolved via the hooks symbolically and treated with load/store semantics. This is implied in the structure since we have no SSA value that represents a buffer and we already assume that reading parameters happens via such a mechanism.	2024-04-01 14:18:12 -07:00
Xinan Jiang(姜曦楠)	1cdae6bc68	[MLIR][TORCH]Add support lowing aten.Int.bool to arith (#3083 ) Now there no lowing for `aten.Int.bool` in `convert-torch-to-arith` pass. this PR add this support. Below is the UT. ``` func.func @torch.aten.Int.bool(%arg0: !torch.bool) -> !torch.int { %0 = torch.aten.Int.bool %arg0 : !torch.bool -> !torch.int return %0 : !torch.int } ```	2024-04-01 10:05:08 -07:00
Stella Laurenzo	282e9b0e64	[fx] Fix type determination for multi-return ops and static `None` returns. (#3081 ) In practice, this was caught by the way that AOT autograd traces `convolution_backward`. For the unit test, we just repro it with a custom op.	2024-04-01 09:39:38 -07:00
Gaurav Shukla	129a79417a	[MLIR][ONNX] Fix onnx.gather_nd implementation (#3070 ) The indices should be expanded before the torch.gather operation. Signed-off-by: Gaurav Shukla <gaurav@amd.com>	2024-04-01 20:17:09 +05:30
Xida Ren (Cedar)	5f325749f9	add lowerings for AtenLtIntOp and AtenLeIntOp (#3061 ) Co-authored-by: Xida Ren <xida.ren.dev@gmail.com>	2024-03-27 10:06:43 -07:00
Yuanqiang Liu	0a581a97a7	[Torch Dialect] enhance aten.int.tensor's canonicalize (#3058 ) support fold with literal vtensor. change it to canonicalize because this pattern will create new op.	2024-03-27 09:51:58 +08:00
Stella Laurenzo	e2343cf4ce	[fx] Implement auto_functionalized higher order op. (#3063 ) * Also adds the basic scaffolding for handling more of these, which will be needed for cond, while, etc. * Refactors some of the support in the generic OpOverload emitter so it can be shared with these other special forms. This has been on my list for a while, but it just so happens that as part of upgrading to PyTorch 2.3 and a pure upstream flow in Turbine, we were using a feature that required integration with auto_functionalized. This is perhaps the "weirdest" of the higher-order ops and a poor place to start, but needs must. We have testing for this in Turbine. Full support in Turbine has an entire custom ops facility. I've reduced this down to a unit test in torch-mlir.	2024-03-26 17:06:05 -07:00
Rob Suderman	14b548f968	[torch] Improve shape inference for `torch-to-linalg` path for reshapes (#3055 ) Reshaping tensors depend on directly matching individual dimensions to their corresponding dim in the `torch.view` reshape dimensions. This involves decoupling dynamic dimensions from their static counterparts and support cleanup / canonicalization.	2024-03-26 12:41:40 -07:00
Vivek Khandelwal	9ae33e482e	[MLIR][TORCH] Add OnnxToTorch lowering for ops (#3049 ) This commit adds the OnnxToTorch lowering for the Mish, Softplus, HardSwish, Trilu, ThresholdedRelu op Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-03-25 20:29:07 +05:30
zjgarvey	6aa481c204	[ONNX] LogSoftmax to Torch (#3024 ) This PR adds support for onnx.LogSoftmax both for old versions (<13, with axis >=0), and new versions (13).	2024-03-22 11:01:39 -07:00
Gaurav Shukla	50635dd509	[ONNX][MLIR] Add support for onnx.gather_nd (#2988 ) Signed-off-by: Gaurav Shukla <gaurav@amd.com>	2024-03-22 21:38:39 +05:30
Stella Laurenzo	6ea857c644	[fx] Make the lift_fresh_copy -> clone special form use kwargs. (#3045 ) At some point, this op became kwarg-only instead of arg/kwarg. Discovered when upgrading to PyTorch 2.3. Also adds a test as this was untested in-tree (was caught out of tree).	2024-03-21 15:34:40 -07:00
penguin_wwy	7616d637fd	Add stateless fx graph import (#3036 )	2024-03-21 14:44:54 -07:00
Xida Ren (Cedar)	cb5cb506df	Fix SCF Forloop fails to convert to linalg when a tensor argument is supplied to the loop block (#3040 ) Co-authored-by: Rob Suderman <rob.suderman@gmail.com> Co-authored-by: Xida Ren <xida.ren.dev@gmail.com>	2024-03-20 11:04:02 -07:00
zjgarvey	6ff71b40c8	[ONNX] onnx.DynamicQuantizeLinear to Torch (#3009 ) This adds support for converting DynamicQuantizeLinear from torch-onnx to torch. I could not get an e2e test to pass, since there seems to be some issues with uint8 casting somewhere lower in the pipeline. For example compiling with IREE for llvm-cpu, I would get either the correct zero point (if zp < 128) or the correct zero-point minus 256 (if zp >= 128). The output tensor seems to always return a tensor of zeros, which also occurs when running uint8 examples through QuantizeLinear. Edit: the first problem can be resolved by casting the output back to uint8 on output, the second problem is resolved with PR #3018	2024-03-20 10:58:25 -07:00
jinchen	9cf6c45a39	Add OnnxToTorch support for Compress op (#3025 )	2024-03-20 17:12:08 +00:00
Pavani Chowdary	c51e2130f2	[onnx] support for lowering mod op from onnx to torch (#2859 ) nod-ai/Shark-Turbine#267 --------- Authored-by: boddu.pavani@research.iiit.ac.in Co-authored-by: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-03-18 17:54:37 +05:30
penguin_wwy	f34c187ac4	Normalize type hints to be compatible with multiple Python versions (#3028 ) Although we provide a wheel package for Python 3.8, it may actually throw the following exception: `TypeError: 'type' object is not subscriptable`	2024-03-15 08:29:48 -07:00
Sambhav Jain	0b2f9c89a2	Bring back `dynamic_shapes` constraints in fx importer API (#3026 ) https://github.com/llvm/torch-mlir/pull/2992 dropped `constraints` from the fx importer API, [breaking](https://github.com/cruise-automation/mlir-tcp/actions/runs/8284385380/job/22669774071) downstream AOT compile tests in `mlir-tcp` that use it. This knob has been soft-deprecated for a while now, replaced by `dynamic_shapes` - a more ergonomic interface. This PR brings back dynamic_shapes constraints in the new supported form. Also added a python lit test with dynamic shaped annotations.	2024-03-14 10:26:34 -07:00
aldesilv	6fa21bd8b1	OnnxToTorch lower celu op (#2920 )	2024-03-13 20:34:10 +05:30
Rob Suderman	8fb28661f9	[onnx] Fix onnx.ReduceMean lowering (#3002 ) Reduce mean lowerings did not succesfully lower to `linalg` via torched. There were two separate paths that could be consolidated to a single simpler pass. This resulted in a significant improvement in test coverage.	2024-03-11 11:32:53 -07:00
Rob Suderman	bd7f1baa42	[onnx] Fix expand operation for dynamic shape max (#3001 ) If the broadcast shape is length-1 at a dim while `?` in the input dim then we need to broadcast to the dynamic dim. This is equivalent to taking a max of two dimensions.	2024-03-08 16:23:07 -08:00
Rob Suderman	0723584936	[torch] Add folder for torch.aten.*.Scalar comparisons (#3000 ) This folds small version of the tensor-scalar comparison operators as they are commonly used for shape computations. This includes le, lt, ge, gt, eq, and ne.	2024-03-08 13:44:00 -08:00
Andreas Falkenberg	551a4e45f3	[onnx] Add support for `onnx.Gemm` with no bias (#2993 ) Previous gemm version required a bias vector. This provides an alternate path to `Torch::AtenMm` with no bias operation.	2024-03-07 15:58:38 -08:00
Rob Suderman	1964208d19	[onnx] Fix constant pad for dynamic shape (#2989 ) The current padding operation was not functional for dynamic shapes. Updated and enabled tests so that onnx.pad tests pass. Work TBD for reflection padding.	2024-03-07 13:29:50 -08:00
Scott Todd	7b18646def	[onnx] Handle optional arguments in Clip op pattern. (#2976 ) Spec: https://onnx.ai/onnx/operators/onnx__Clip.html	2024-03-07 17:25:14 +00:00
Vivek Khandelwal	6e84752c39	build: manually update PyTorch version (#2992 ) Set PyTorch and TorchVision version to nightly release 2024-03-07. This commit also removes the deprecated constraints API: `342e7929b8` Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-03-07 21:42:38 +05:30
Rob Suderman	c15f1a2bd2	[onnx] Adding lowering for `onnx.Size` operation (#2985 ) We can support `onnx.Size` by requesing the size of each dimensions and taking the product of the results, then packing it into a tensor. --------- Co-authored-by: Scott Todd <scott.todd0@gmail.com>	2024-03-06 17:01:05 -08:00
Rob Suderman	a78659742a	[onnx] Migrate `onnx.ReduceMax` to match `onnx.ReduceMin` (#2981 ) This mostly copy-pastes the reduce minimum implementation to reduce max to improve test coverage. We also improve the aten lowering for min/max dim for unsigned types.	2024-03-06 16:48:21 -08:00
Andreas Falkenberg	ea76dd12ba	[onnx][torch] Gridsampler E2E test and corrections of gridsampler (#2987 ) The addition of an e2e test is actually provided in the Shark-Testsuite. This adds 2 test cases for the gridsampler e2e test. Also as intended there were some items found which needed correction, so the Gridsampler op has also a change.	2024-03-06 10:56:58 -08:00
Rob Suderman	933db87a07	[onnx] Add support for constants of `i1`s (#2978 ) `getRawBuffer` expects a densely packed vector of `i1` values however `onnx` does not densely pack the values. Include code to handle the packing / unpacking.	2024-03-05 13:55:13 -08:00
Rob Suderman	a86e89ecb5	[torch] Additional folders for shape computations (#2972 ) A handful of operations are commonly used in shape calculations (slice, concat, broadcast). Added these additional folders to better propagate simple shape computations.	2024-03-04 11:46:49 -08:00
Chi_Liu	09875fabd1	[MLIR][ONNX] Add ONNX ReduceProd support (#2943 ) Alternatives to https://github.com/llvm/torch-mlir/pull/2908 Fix https://github.com/nod-ai/SHARK-Turbine/issues/353	2024-03-04 11:07:03 -08:00
Rob Suderman	d51e80b648	[onnx] Fix onnx.gather lowering for rank-0 indices (#2973 ) We assumed rank was atleast 1 however it can be rank-0, generating an illegal pair of flatten / unflatten operations. Corrected this.	2024-03-04 08:25:19 -08:00
Rob Suderman	61f0a5facf	[torch] Add an `aten.cat` length-0 canonicalization (#2966 ) If an input is length-0 along the dimension of canonicalization we can remove the tensor from the list	2024-03-01 21:41:12 -08:00
Vivek Khandelwal	579ac8b666	[MLIR][TORCH] Fix OnnxToLinalg lowering issue for sub and sum op (#2954 ) This commit adds the support for scalar conversion to byte. This commit also fixes the OnnxToLinalg lowering issue for Onnx.Sub and Onnx.Sum op. Fixes https://github.com/nod-ai/SHARK-Turbine/issues/466 Fixes https://github.com/nod-ai/SHARK-Turbine/issues/467 Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-02-29 21:48:46 +05:30
Aart Bik	f21b76b68a	[torch-mlir][sparse] fixed merge conflict (#2967 )	2024-02-28 17:14:00 -08:00
Peiming Liu	e85a2a87c5	[torch-mlir][sparse] support e2e sparse kernels with COO inputs. (#2939 )	2024-02-28 16:08:37 -08:00
Andreas Falkenberg	5437f32193	[onnx][torch] Lower `onnx.grid_sampler` to the `torch` equivalents (#2952 ) This is the lowering of gridsampler from onnx to torch using our prior implementation of AtenGridSamplerOp. Here are several checks for cornercases implemented. We may decide to have part of these checks in AtenGridSamplerOp instead of the onnx lowering portion.	2024-02-28 13:52:15 -08:00
Rob Suderman	e48fe45886	[onnx] Import `onnx` import to pass remaining tests (#2951 ) Finish supporting importing the vast majority of `onnx` operations. This includes: - region support - region value inherentance - `torch.string` support - `torch.list` support - `torch.optional` support	2024-02-28 12:18:02 -08:00
Rob Suderman	6f3d62ab04	[torch] Fix folders and `cat` and `view` torch lowerings (#2963 ) A bunch of small fixes are interlinked and trigger crashes if not addressed as a group. This includes: - aten view when expand from a rank-0 tensor - slice folder with negative indices - `aten._shape_as_tensor` folder on a rank-0 tensor - `aten.cat` of a tensor with a length-0 tensor	2024-02-28 12:04:52 -08:00
Rob Suderman	4a7a7d76f8	[onnx] Fix ReduceMean lowering to torch (#2956 ) Torch lowering only supported the most recent version. Refactored the lowering so more easily handle default values and optional operands / attributes.	2024-02-27 22:48:07 -08:00
Aart Bik	30212547a9	[torch-mlir][sparse] add JIT test for block sparse SpMV (#2955 ) This required adding a "decompose" pass to the torch lowering, since torch.mv was not directly handled by lowering to linalg	2024-02-27 11:49:32 -08:00
Rob Suderman	e30a083aff	[torch] Rework lowering to tm_tensor.scatter to stop serialization (#2940 ) We collapsed and broadcasted scatter indices to a single element version. We should instead upport `tm_tensor.scatter`s support for multiple indices and the implicitly broadcasted behavior. This avoids the serialization and materializing a needlessly large indices tensor.	2024-02-27 11:46:57 -08:00
Vivek Khandelwal	d81747eadb	[MLIR][TORCH] Extend support for OnnxToLinalg lowering for Dropout and Div op (#2938 ) Fixes https://github.com/nod-ai/SHARK-Turbine/issues/451, https://github.com/nod-ai/SHARK-Turbine/issues/452	2024-02-27 11:02:05 +05:30
Sambhav Jain	3cbe6c98ec	Expose `func_name` to the main fx import API (#2949 ) As titled.	2024-02-26 10:08:14 -08:00
Aart Bik	4147b280ce	[torch-mlir][sparse] add block sparsity to mlir lowering (#2942 ) Also note that we are in the process of proposing SparseTensorMetadata to PyTorch FX graph export (see https://github.com/pytorch/pytorch/pull/117907). This will hopefully eventually replace the current data structures in torch-mlir.	2024-02-23 11:57:20 -08:00
Andreas Falkenberg	55dc8deb92	[torch] GridSample TorchToLinalg lowering (#2883 ) Lowers `torch.grid_sample` to the equilvalent `linalg` representation.	2024-02-23 09:14:38 -08:00
Rob Suderman	53f6d06ab8	[onnx] Drop `ConstantOfShape` logic form importer, fix torch lowering (#2930 ) There is no reason to treat `ConstantOfShape` as a specialized import any as there exists a onnx-to-torch equivalent. Dropping the import coding and adding support for resource conversion substantially increases test coverage for dynamically shaped tests.	2024-02-21 21:34:43 -08:00
Srinath Avadhanula	0f80e75c2e	allow tosa.cast to convert from f32 to f16 (#2934 ) According to the [official TOSA spec](https://www.mlplatform.org/tosa/tosa_spec.html#_cast), `tosa.cast` allows a cast from `fp32` to `fp16`. We were not previously accounting for this in the `TorchToTosa` lowering. Also did a tiny bit of cleanup in the code to make it easier to spot which conversions are currently allowed. --------- Co-authored-by: Srinath Avadhanula <srinath.avadhanula@getcruise.com>	2024-02-20 14:22:38 -08:00
Rob Suderman	13553d49c9	[onnx] Update the importer to create a `none` for missing operands (#2931 ) Some operands are optional so we require a placeholder for missing operands. We invent an `onnx.None` operation as our placeholder.	2024-02-20 09:30:30 -08:00
Stella Laurenzo	4446fa00d8	Migrate passes in TorchConversion to use FunctionOpInterface. (#2935 ) This enables better re-use in downstreams which use different func implementations and should have no impact on those that don't except in opt pipelines if using the old form. With interfaces, explicit pipelines via `--pass-pipeline=` must be used.	2024-02-20 08:54:02 -08:00
Rob Suderman	135c81a416	[torch] Add folder for `prim.NumToTensor.Scalar` (#2921 ) Useful for `slice` lowerings that depend on tensors made form scalars.	2024-02-19 11:55:54 -08:00
Rob Suderman	e80054a3cc	[torch] Folders for `torch.aten.*.tensor` operators [add, sub, mul] (#2878 ) Simple folder for limited size aten tensor operations. This is primarily useful for shape computation folding as they unfortunately can use `aten` operators. Add, sub, mul are common examples of these folders.	2024-02-19 10:28:23 -08:00
Rob Suderman	cea51897a5	[onnx] Simplify onnx.slice lowering (#2919 ) Onnx slice lowering used arange needlessly instead of directly constructing the constant dimension values. This makes lowerings to linalg struggle as multiple folders are required to get what is a constant index value.	2024-02-19 10:26:29 -08:00
aldesilv	d29157b33f	OnnxToTorch support for onnx.InstanceNormalization op (#2710 ) https://github.com/nod-ai/SHARK-Turbine/issues/327	2024-02-19 19:53:48 +05:30
Rob Suderman	7a0d0e954b	[onnx] Fix onnx.gather lowering to use torch.aten.index_select (#2913 ) Onnx's gather maps directly to `torch.aten.index_select`. We should just use that path.	2024-02-16 16:05:44 -05:00
Aart Bik	c5d8c12469	[torch-mlir][sparse][NFC] fixed typo (#2917 ) grammar police	2024-02-16 13:02:00 -08:00
Stella Laurenzo	5253282c55	[fx] Support mutation in ExportedProgram. (#2916 ) As of https://github.com/pytorch/pytorch/pull/118969, `ExportedProgram` has the long awaited fixes to correctly categorize various things relating to parameters, buffers, mutated inputs and constants. With this additional modeling, we are finally able to implement (safely/soundly) the mutable semantics that were attempted on the TorchScript path. The difference is that on that path, we had to conservatively treat everything as mutable and run some dodgy heuristics (which have been the cause of many bugs relating to "MaximizeValueSemantics") to try to get back to an immutable state. The new model supports mutability at the graph edges, allowing both user inputs and buffers to be mutated (there is some more support than that, but that is all I fully tracked through to implementation). Therefore, when we receive programs like this, we now can selectively enable mutation at the edges. This happens to be the mutability model that IREE supports, which I expect to be a primary beneficiary. However, there is nothing stopping anyone else from handling the `!torch.tensor` types and the existing copy/overwrite ops that will be selectively added. Since this relies on API changes that will not release until 2.3, I'm being a bit cautious about not refactoring existing facilities.	2024-02-16 09:46:30 -08:00
Rob Suderman	074f112d6a	[onnx] Add testing using the `onnx` compilation using torch tests (#2795 ) We can route the torch tests via `onnx` using the `torch.onnx.export` tooling. We can then reimport, lower to torch, and compile to linalg to validate the onnx path is working correctly. The current implementation exposes some failures in the `onnx` path so we cannot enable the onnx test suite yet due to segmentation faults.	2024-02-15 10:17:13 -08:00
Vivek Khandelwal	d6d1a173dc	[MLIR][Torch] Add OnnxToTorch and TorchToLinalg support for trig ops (#2903 ) This commit adds the OnnxToTorch lowering for cosh, acosh, asin, asinh, and atanh op. This commit also adds the TorchToLinalg lowering for acosh, asin, asinh, and atanh op. Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-02-14 11:58:09 +05:30
Rob Suderman	e9cdd6cbc5	[torch] Fix tm_tensor.attention for end-to-end (#2907 ) Some operations include a backend matcher for specialized operations. We map these back to generics so they appropriately match to the high performance versions. This is done for the attention operation.	2024-02-13 21:18:01 -08:00
Scott Todd	d6e1d836ca	Drop torch attributes at the end of backend conversion. (#2876 ) Fixes https://github.com/llvm/torch-mlir/issues/2866 Some backends / downstream projects expect that a "fully converted" program has no remaining ops or attributes from the original dialect(s).	2024-02-13 14:32:02 -08:00
Aart Bik	24c2fc0b5f	[torch-mlir][sparse] add JIT test to expose pending issues (#2906 ) This test exposes issues that need fixing (1) propagate sparsity into the FX graph (over elt-wise) (2) batched dimensions need a new "dense(batch)" format	2024-02-13 13:42:56 -08:00
saienduri	9b967f6b5a	[MLIR][ONNX] Add OnnxToTorch support for Mean, IsInf, IsNaN, PRelu op (#2801 ) This commit adds the OnnxToTorch support for Mean, IsInf, IsNaN, and PRelu ops. All high priority ops were taken so went with these. The non trivial ones are Mean and IsInf which might require extra review --------- Co-authored-by: MaheshRavishankar <mravisha@amd.com>	2024-02-13 12:38:21 +05:30
Aart Bik	b6f4ca512e	[torch-mlir][sparse] sparsity metadata refinement (#2901 ) Various improvements on sparsity metadata: (1) define single data structure for all sparsity related metadata (2) handle batched dense dimensions, as well as dense subtensor dimensions (3) refine sparsity propagation for deeper networks	2024-02-12 16:10:57 -08:00
Aart Bik	be8375d350	[torch-mlir][sparse] implement first sparse_jit end-to-end path (#2894 ) This PR introduces a sparse_jit wrapper that can run simple models with sparse tensor inputs end-to-end. The implementation shows all required components on modifying sparse tensor types with a 1:N relation on the call sites. Two tests shows that the JIT runs end-to-end while computing the correct results. More details to follow (generalizing to COO and different ranks, as well as support for output sparse tensors), but the general concepts are all here now. _Update: Thanks to Rob, bump to proper LLVM/MLIR hash is done!_ _NOTE that all parameter passing changes are nicely done "downstream" in MLIR, so very little changes are required in torch-mlir code proper_ --------- Co-authored-by: Franz Haniel <77495327+frafranz@users.noreply.github.com> Co-authored-by: Franz Haniel <franz.haniel@amd.com>	2024-02-12 10:04:54 -08:00
Rob Suderman	c0f139be0f	[torch] Add `torch.aten.eq.Tensor` comparison folder (#2889 ) Added a folded for a equals operator. This allows an equivalent comparison folder, primarily for when shape computations occur small size tensor.	2024-02-09 15:02:20 -08:00
Rob Suderman	7d33ba69ac	[torch] Folder for torch.aten.select.int for splat cases (#2890 ) If the input or result is a splat value we can just constant fold the result. This is common for shape computations and can help with shape inference.	2024-02-09 14:02:54 -08:00
Ashay Rane	21f070e95f	onnx: fix checks in TorchOnnxToTorch pass to match the ONNX spec (#2848 ) This PR contains three commits to update the validation checks in the ONNX -> Torch conversion pass for the AveragePool, Pad, and Slice operators: > onnx: fix preconditions for lowering AveragePool ops > > The `pads` attribute of the AveragePool operator specifies the value to > pad at both the beginning as well as the end of the axis (see > https://onnx.ai/onnx/operators/onnx__AveragePool.html#attributes), so > the size of this attribute should be twice the rank of the input tensor. > However, our TorchOnnxToTorch bails out early since it incorrectly > compares the pads attribute with the rank (not twice the rank) of the > input tensor. > > This patch fixes the code to match the spec and adds a lit test. > onnx: allow optional constant value for Pad operator > > The `constant_value` input of the onnx.Pad operator is optional (see > https://onnx.ai/onnx/operators/onnx__Pad.html#inputs), but the existing > logic for lowering the operator into the Torch dialect assumes that it > is mandatory. > > This patch makes the attribute optional and constructs a default value > (a list of zeros the size of the input tensor) if the attribute was not > specified. > onnx: fix checks for axes and steps inputs of Slice operator > > The ONNX Spec for the Slice operator allows the `starts` and `ends` > inputs to have fewer indices that the dimensions of the `data` tensor > (see https://onnx.ai/onnx/operators/onnx__Slice.html), but our code > expects these inputs to be as many as the `data` tensor's dimensions. > > More precisely, the spec requires that the `starts` and `ends` inputs > are only as long as the `axes` input, but since the `axes` input is > optional, the default type for the `axes` input has to match the type > for the `starts` and `ends` inputs. Moreover, the number of indices in > the `steps` input also has to match those in the `axes` inputs (instad > of matching the dimensions of the `data` input). > > This patch fixes the checks in the TorchOnnxToTorch conversion so that > they match the ONNX spec.	2024-02-07 21:19:27 -08:00
Vivek Khandelwal	4df96616db	[MLIR][TORCH] Modify Onnx.Reshape lowering for static shape cases (#2852 ) This commit modifies the OnnxToTorch lowering of Onnx.Reshape op by creating the result shape list for the aten.reshape using the result shape values inferred from the op's result shape. Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-02-07 17:44:07 -08:00
Dave Liddell	23647ab2d1	[torhc] aten.index_select folder (#2871 ) Folds aten::index_select ops under the following conditions: 1. If the input and output are the same shape, the indexing operation is a NOP, so just return the input. 2. If the input has shape <1x1x...xNx...x1> (all 1's except for one dim), and the output shape is <1x1x...x1> (all 1's), then there is a single index, so extract the single element value and return a tensor with that value. --------- Co-authored-by: Dave Liddell <dliddell@xilinx.com>	2024-02-07 16:17:15 -08:00
Xida Ren (Cedar)	fc04bc7ee9	[torch] AtenSliceOp folder that produces splat results (#2869 ) Includes `slice` folder and lit tests --------- Co-authored-by: Xida Ren <xida.ren.dev@gmail.com>	2024-02-07 19:00:46 +00:00
saienduri	bfcf93ea21	Rename torch_mlir.compile APIs and introduce FX based analogs (#2842 ) Link to related RFC: https://discourse.llvm.org/t/rfc-rename-torch-mlir-compile-apis-and-introduce-fx-based-analogs/76646 This commit updates the documentation, tests, CMake files, and API for the proposed changes in the RFC. There is a new torch_mlir/fx.py for user level APIs related to importing modules and a corresponding test for this path can be found at test/python/fx_importer/basic_test.py. --------- Co-authored-by: MaheshRavishankar <mravisha@amd.com>	2024-02-06 19:07:59 -08:00
Xida Ren (Cedar)	cc06391630	AtenSortOp Folder (#2864 ) A chunk off https://github.com/llvm/torch-mlir/pull/2856 https://github.com/llvm/torch-mlir/pull/2860 --------- Co-authored-by: Xida Ren <xida.ren.dev@gmail.com> Co-authored-by: Rob Suderman <rob.suderman@gmail.com>	2024-02-06 21:12:12 +00:00
Daniel Garvey	faf7d4aaa5	[fx_importer] Add support for 0D tensors (#2870 ) Adds an escape hatch from creating a DenseResourceElementsAttr for single value tensors into DenseElementsAttr. For 0d or 1element, splats are better as DenseElementsAttr. Don't use DenseResourceElementsAttr for it	2024-02-06 00:19:31 -06:00
Dave Liddell	1cb14f6879	Rob's atenTensor folder (#2867 ) If a tensor is initialized by a list with a single constant integer, this folder turns it into a torch.vtensor.literal --------- Co-authored-by: Dave Liddell <dliddell@xilinx.com>	2024-02-05 17:10:42 -08:00
Rob Suderman	e3faef5224	[onnx] Convert `onnx.QLinearConv` to `torch` (#2851 ) Leaning on the QDQ functionality in torch we can support the QLinearConv operation by piggybacking through `torch.Convolution`. This includes some changes such as allowing the `onnx` rewriter to run recursively. Doing so allows `QLinearConv` to decopmose to `onnx.Convolution` which is then lowered to `torch`.	2024-02-05 16:09:41 -08:00
Rob Suderman	cb52c4b3cc	[onnx] Fix `onnx-to-torch` lowering for flatten shape (#2834 ) The existing `flatten` lowering did not define what the intermediate shape was. This could result in failures to lower further to linalg as the intermediate shape was unknown. Added a shape refinement section.	2024-02-05 14:23:46 -08:00
Gaurav Shukla	f4562a8eaa	[ONNX] Fix the lowering of onnx.expand op (#2861 ) Signed-off-by: Gaurav Shukla <gauravshukla789@gmail.com>	2024-02-05 23:46:58 +05:30
Xida Ren (Cedar)	24b8c8672a	[torch] Add folders for `torch.fill`, `torch.ones`, `torch.zeros` and `aten.getItem` (#2849 ) So that the CumSum Op in OPT can get the constant that it requires to be lowered to TMTensor --------- Co-authored-by: Rob Suderman <rob.suderman@gmail.com> Co-authored-by: Xida Ren <xida.ren.dev@gmail.com>	2024-02-02 10:46:33 -08:00
Rob Suderman	29baa813bd	[onnx] Fix `pool` lowering for non-symmetric padding (#2837 ) `torch` requires that padding be symmetric for pooling operations. To support non-symmetric pad we need to separately materialize out the padding operation. --------- Co-authored-by: James Newling <james.newling@gmail.com>	2024-02-01 14:35:21 -08:00
Dave Liddell	04be6ba773	Make the onnx importer more robust for internal/external and large models (#2794 ) Fix for https://github.com/llvm/torch-mlir/issues/2765 The onnx docs say that you can't do shape inference using the in-memory API for models > 2 GB. This fix replaces that API with the file-based API. Since the new API generates an intermediate file, also added a --keep switch to keep that file, which I delete by default. --------- Co-authored-by: Dave Liddell <dliddell@xilinx.com>	2024-01-31 21:58:43 -08:00
Sambhav Jain	8a17c98b74	Bump stablehlo to openxla/stablehlo@fd52182f76 (#2821 ) With the recent LLVM integrate and changes from https://github.com/llvm/llvm-project/pull/78260, we hit this build error in Stablehlo (which is quite old). ``` external/stablehlo/stablehlo/transforms/StablehloRefineShapes.cpp:1020:14: error: no member named 'startRootUpdate' in 'mlir::PatternRewriter' rewriter.startRootUpdate(op); ~~~~~~~~ ^ external/stablehlo/stablehlo/transforms/StablehloRefineShapes.cpp:1026:16: error: no member named 'finalizeRootUpdate' in 'mlir::PatternRewriter' rewriter.finalizeRootUpdate(op); ~~~~~~~~ ^ external/stablehlo/stablehlo/transforms/StablehloRefineShapes.cpp:1029:16: error: no member named 'cancelRootUpdate' in 'mlir::PatternRewriter' rewriter.cancelRootUpdate(op); ~~~~~~~~ ^ external/stablehlo/stablehlo/transforms/StablehloRefineShapes.cpp:1108:14: error: no member named 'updateRootInPlace' in 'mlir::PatternRewriter' rewriter.updateRootInPlace(op->getParentOp(), [&]() { return; }); ~~~~~~~~ ^ 4 errors generated. Target @torch-mlir//:torch-mlir-opt failed to build ``` I'm still puzzled as to how this didn't fail with the CMake merge gating CI (do we not test Stablehlo builds/tests?). In any case, bumping our submodule to https://github.com/openxla/stablehlo/pull/1918 fixes it. It exposes a new failing lit test in TorchToStablehlo though, that I have looped stablehlo developers into ([here](https://discord.com/channels/999073994483433573/999074539138990131/1201235845391331419)). ``` bazel run @torch-mlir//test/Conversion:TorchToStablehlo/scatter.mlir.test ...external/torch-mlir/test/Conversion/TorchToStablehlo/scatter.mlir within split at <stdin>:1 offset :33:8: error: unexpected error: Expects non-empty reduction block for type inference %0 = torch.aten.scatter.src %arg0, %int0, %arg1, %arg2 : !torch.vtensor<[?,?],si64>, !torch.int, !torch.vtensor<[?,?],si64>, !torch.vtensor<[?,?],si64> -> !torch.vtensor<[?,?],si64> ^ LLVM ERROR: Failed to infer result type(s). ``` Bazel CI: https://github.com/sjain-stanford/torch-mlir/actions/runs/7732673480/job/21083102228	2024-01-31 14:21:17 -08:00
Rob Suderman	3500523f75	[onnx] Convert resources to denseattr for `onnx.constant` to `torch` (#2830 ) `onnx` explicitly specifies that `raw_data` is stored in `little-endian` layout. While converting to `torch` we need to convert from a known endian format to an internal format of consistent layout. This means endianness must be correct during the import of `onnx.Constant`. --------- Co-authored-by: Xida Ren (Cedar) <cedar.ren@gmail.com>	2024-01-31 11:40:53 -08:00
Stella Laurenzo	943164d797	Fix some spurious `None` values in tests (broken at head). (#2840 )	2024-01-30 22:39:22 -08:00
Aart Bik	105aad6f57	[torch-mlir] provide FX traced graph importer for sparse tensors (#2817 ) Note that we are waiting for actual FX traced graph support for sparse tensors. For details see https://github.com/pytorch/pytorch/issues/117188 Until then, however, we provide this clever importer that builds the FX traced graph for for the dense case and then puts a sparse annotation back on the parameters. With import test.	2024-01-30 21:22:12 -08:00
Yuanqiang Liu	d778950f45	[Torch Dialect] add fold pattern for aten.clone (#2804 )	2024-01-31 09:43:21 +08:00
Rob Suderman	25a5a22cbd	[torch] Support `torch.convolution` quantized lowering to `linalg` (#2811 ) Linalg has quantized specific operations. We can lower to these operations when there is a known zeropoint and scale operations. This allows the `convolution` to occur with lower bitwidth's, improving the overall performance.	2024-01-30 13:46:47 -08:00
Aaron St George	4c557847bd	Don't fold `aten.detach` if result isn't same type as input. (#2824 ) We were seeing some assertion failures after some checks around folders were tightened up in LLVM: https://github.com/llvm/llvm-project/pull/75887 . This PR essentially moves the logic that used to be applied at the LLVM level into the folder, which seems to be the suggested fix. I'm not sure if the IR that caused issues for us _should_ be valid? ``` %1 = torch.aten.detach %arg0 : !torch.tensor<[1],f32> -> !torch.tensor ``` A better fix might be to create a verifier ensuring the result of `aten.detach` has the same type as its operand. --------- Co-authored-by: aaron-stgeorge <aaron.stgeorge@getcruise.com>	2024-01-30 09:45:51 -08:00
aldesilv	eff325abc3	OnnxToTorch ReduceMax lowering (#2768 ) Fixes https://github.com/nod-ai/SHARK-Turbine/issues/352	2024-01-30 11:44:48 +05:30
Rob Suderman	d3fd754b93	[onnx] `onnx.MatMulInteger` lowering to `torch.mm` and `quint*` types (#2761 ) Torch does not have an equivalent matmul operation for integers. Instead it sidechannels the information via its quantized types. For this lowering we setup these sidechannels then invoke `torch.mm`.	2024-01-29 09:40:21 -08:00
Aart Bik	46a25d7241	[torch-mlir][sparse] preserve sparsity during lowering torch to linalg (#2809 ) This preserves sparsity at the most obvious places of lowering TORCH tensors to MLIR RankedTensorType tensors. Other places are marked for audit. With some initial lowering tests.	2024-01-26 10:54:59 -08:00
Vivek Khandelwal	da7c6d2c16	[MLIR][TORCH] Add support for dynamic shape for Onnx.Transpose op (#2803 ) Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-01-26 09:46:54 -08:00
Phaneesh Barwaria	4964977e85	[ONNX][MLIR] support constantOfShape op (#2747 )	2024-01-26 09:36:39 -08:00
Aart Bik	0aed231e21	[torch-mlir][conversion-test] cleanup trailing whitespace in mlir files (#2807 )	2024-01-25 14:24:28 -08:00
Aart Bik	fe836ceebf	[torch-mlir][test] cleanup trailing whitespace in mlir files (#2806 )	2024-01-25 14:24:13 -08:00
Aart Bik	e824fbc65c	[torch-mlir][torch] add encoding field to torch type (#2799 ) This adds an encoding field to the torch type, using the interfaces for printing, parsing, and verification. Note that although this change prepares adding sparsity to the torch type (as illustrated by the round trip and invalid tests), nothing in this change depends on the actual contents of the encoding field!	2024-01-25 10:04:04 -08:00
Rob Suderman	f6f890520b	[torch][quant] Quantized `torch.mm` for linalg with end-to-end test (#2750 ) This includes custom op matching for decomposed operations and fusing dequantization into dense operations. As a validation we compare to the dequant+mm torch implementation.	2024-01-24 14:02:50 -08:00
Rob Suderman	60bf6c25af	[onnx] Lower `onnx.QLinearMatMul` lowering to `torch` operators (#2776 ) We can plumb the linear matmul into pytorch using its quantized types with side channel information. To handle the final int8 operation we dequantize and requantize.	2024-01-24 12:28:48 -08:00
Vivek Khandelwal	894805dd5e	[MLIR][TORCH] Support for `onnx.LayerNormalization` (#2789 ) Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-01-24 11:08:20 -08:00
Gaurav Shukla	12f123eff8	[ONNX][MLIR] Add support for pad op in the onnx pipeline (#2738 ) This commit adds mapping from `onnx.pad` op to `torch.pad` op. Currently it does not support `axes` parameter of `onnx.pad` op. Signed-off-by: Gaurav Shukla <gaurav.shukla@amd.com>	2024-01-25 00:33:37 +05:30
Phaneesh Barwaria	ac8975ea12	[MLIR] [ONNX] lowering for onnx tile op and sign op (#2725 )	2024-01-24 22:56:21 +05:30
Chi_Liu	77ae56337d	[ONNX][MLIR] Add support for onnx.Exp op (#2792 ) https://github.com/nod-ai/SHARK-Turbine/issues/312	2024-01-23 13:45:00 -08:00
James Newling	dc056e58e6	[MLIR][TORCH] Add onnx.cast cases used by OPT-1.25M (#2787 )	2024-01-23 21:06:25 +05:30
Gaurav Shukla	b7a0329676	[ONNX][MLIR] Fix padding size constraint for onnx.maxpool op (#2782 ) Signed-off-by: Gaurav Shukla <gaurav@nod-labs.com>	2024-01-23 19:23:01 +05:30
Chi_Liu	cad98e8113	[ONNX][TORCH-MLIR] Add TopK support (#2774 ) https://github.com/nod-ai/SHARK-Turbine/issues/331	2024-01-22 12:56:39 -08:00
Srinath Avadhanula	73b30604da	Do not try to legalize transposed convolution (#2721 ) Currently transposed convolution is not handled correctly by `TorchToTosa`. This PR allows transposed convolutions to pass through the conversion so that they can be handled by other conversion passes later in a pipeline. An example input which produces a compilation error is: ``` func.func @forward(%input: !torch.vtensor<[1,64,1,100],f32>) -> !torch.vtensor<[1,64,2,200],f32> { %true = torch.constant.bool true %int1 = torch.constant.int 1 %int2 = torch.constant.int 2 %weight = torch.vtensor.literal(dense<0.0> : tensor<64x64x3x3xf32>) : !torch.vtensor<[64,64,3,3],f32> %bias = torch.vtensor.literal(dense<0.0> : tensor<64xf32>) : !torch.vtensor<[64],f32> %stride = torch.prim.ListConstruct %int2, %int2 : (!torch.int, !torch.int) -> !torch.list<int> %int1x1 = torch.prim.ListConstruct %int1, %int1 : (!torch.int, !torch.int) -> !torch.list<int> %output = torch.aten.convolution %input, %weight, %bias, %stride, %int1x1, %int1x1, %true, %int1x1, %int1 : !torch.vtensor<[1,64,1,100],f32>, !torch.vtensor<[64,64,3,3],f32>, !torch.vtensor<[64],f32>, !torch.list<int>, !torch.list<int>, !torch.list<int>, !torch.bool, !torch.list<int>, !torch.int -> !torch.vtensor<[1,64,2,200],f32> return %output : !torch.vtensor<[1,64,2,200],f32> } ``` This MLIR produces an error about a cast operation with a size mismatch when passed through `torch-to-tosa`: ``` error: 'tensor.cast' op operand type 'tensor<1x64x1x50xf32>' and result type 'tensor<1x64x2x200xf32>' are cast incompatible ``` --------- Co-authored-by: Srinath Avadhanula <srinath.avadhanula@getcruise.com>	2024-01-22 10:57:56 -08:00
Dave Liddell	2f4924015d	[onnx] Added flatten (#2760 ) [https://github.com/nod-ai/SHARK-Turbine/issues/328](url) --------- Co-authored-by: Dave Liddell <dliddell@xilinx.com>	2024-01-19 16:18:16 -08:00
Gaurav Shukla	3b85c70748	[ONNX][MLIR] Add support for onnx.gather op (#2726 ) This commit adds support for gather op in the onnx pipeline. https://github.com/nod-ai/SHARK-Turbine/issues/242 Signed-off-by: Gaurav Shukla <gaurav.shukla@amd.com>	2024-01-19 21:58:29 +05:30
John Wu	704cfdaf08	Add aten.pool_max3d support to torch-to-linalg (#2735 ) Added verification logic to the abstract_interpreter_lib_gen.py Also made some unit tests Initially, I thought we can use `linalg::pooling_ndhwc_max` to help implement this problem. However, on a 5-dimensional matrix it does the pooling on dimensions (2, 3, 4) which is not what we want. We want pooling on dimensions (3, 4, 5). To achieve this, we would need to lower our code using the `linalg` dialect. Turns out the pooling code in `linalg` looks like this. ``` func @max_pooling_ncdhw(%I: memref<?x?x?x?x?xf32>, %K: memref<3xindex>, %O: memref<?x?x?x?x?xf32>, %strides: memref<3xindex>, %dilations: memref<3xindex>) { %c0 = arith.constant 0 : index %c1 = arith.constant 1 : index %N = memref.dim %I, %c0 : memref<?x?x?x?x?xf32> %C = memref.dim %I, %c1 : memref<?x?x?x?x?xf32> %D = memref.dim %I, 2 : memref<?x?x?x?x?xf32> %H = memref.dim %I, 3 : memref<?x?x?x?x?xf32> %W = memref.dim %I, 4 : memref<?x?x?x?x?xf32> %kernel_d = memref.load %K[%c0] : memref<3xindex> %kernel_h = memref.load %K[%c1] : memref<3xindex> %kernel_w = memref.load %K[2] : memref<3xindex> %stride_d = memref.load %strides[%c0] : memref<3xindex> %stride_h = memref.load %strides[%c1] : memref<3xindex> %stride_w = memref.load %strides[2] : memref<3xindex> %dilation_d = memref.load %dilations[%c0] : memref<3xindex> %dilation_h = memref.load %dilations[%c1] : memref<3xindex> %dilation_w = memref.load %dilations[2] : memref<3xindex> linalg.generic { indexing_maps = [ affine_map<(n, c, d, h, w, kd, kh, kw) -> (n, c, d * %stride_d + kd * %dilation_d, h * %stride_h + kh * %dilation_h, w * %stride_w + kw * %dilation_w)>, // Map for input tensor affine_map<(n, c, d, h, w, kd, kh, kw) -> (kd, kh, kw)>, // Map for kernel tensor affine_map<(n, c, d, h, w, kd, kh, kw) -> (n, c, d, h, w)> // Map for output tensor ], iterator_types = ["parallel", "parallel", "parallel", "parallel", "parallel", "reduction", "reduction", "reduction"], doc = "3D Max Pooling NCDHW with Strides, Dilations, and Kernel Size" } ins(%I, %K : memref<?x?x?x?x?xf32>, memref<3xindex>) outs(%O : memref<?x?x?x?x?xf32>) { ^bb0(%input_elem: f32, %kernel_elem: index, %output_elem: f32): %max_val = arith.maxf %input_elem, %output_elem : f32 linalg.yield %max_val : f32 } return } ``` This was implemented based on it's source code with the adjustments mentioned above: `4ca1b5e094/mlir/include/mlir/Dialect/Linalg/IR/LinalgNamedStructuredOps.yaml (L5647)` Issues related to this can be found here https://github.com/nod-ai/SHARK-Turbine/issues/324	2024-01-19 21:09:46 +05:30
Andreas Falkenberg	4de4d38b87	Initial commit of NonZero op (#2766 )	2024-01-18 15:23:13 -10:00
Rob Suderman	b5387c0f29	[onnx] Lowering `onnx.dequantize_linear` to `torch` (#2759 ) We can make the per-tensor version of the operation to the dequantize operation via marking with the make quantized tensor component. This introductions the `qint` and `quint` tensor type that can be lowered to teh appropriate dequantization behavior during the torch-to-linalg conversion.	2024-01-18 16:47:21 -08:00
Rob Suderman	bd11877f6f	[onnx] Support lowering quantize linear to `torch` (#2751 ) We can map the per_tensor case to the `torch.aten.quantize_per_linear` operation. In this case we extract the `scale` and `zeropoint` values and directly invoke the quantization, then return the integer representation value.	2024-01-18 16:33:10 -08:00
Ze Zhang	77a03f2069	torch-to-tosa lowering support for AtenLinalgVectorNormOp (#2734 ) This PR add torch-to-tosa lowering support for AtenLinalgVectorNormOp e2e test: python -m e2e_testing.main --config=tosa LIT tests: cmake --build build --target tools/torch-mlir/all --------- Co-authored-by: Ze Zhang <ze.zhang@getcruise.com>	2024-01-18 12:32:23 -08:00
Phaneesh Barwaria	eed144bfbc	[ONNX][MLIR] add Identity op support (#2754 )	2024-01-16 19:06:54 +05:30
kumardeepakamd	87389f0762	[ONNXToTorch] Add conversion for Onnx range (#2752 ) Implemented ONNX.Range. The spec says the data type for start, limit, delta are 0-D can be double, float, int16, int32, int64, All int types mapped to !torch.int and all float types mapped to !torch.float --------- Co-authored-by: Kumar Deepak <kumar@xilinx.com>	2024-01-15 14:26:46 -05:00
Rob Suderman	197b3b475c	[onnx] Convert `onnx.constant` to `torch` literal tensor (#2748 ) Handles the multiple cases of `onnx` constant values and converts them to `torch` literal tensors. This can include splats with a single integer or floating point value, a set of explicit integer values, or an elements array attr of values.	2024-01-15 09:31:22 -08:00
Han-Chung Wang	10acea71be	Bump LLVM to llvm/llvm-project@0cb024b (#2753 ) - Add fixes for `af78e5daf0` - Add fixes for `bb6d5c2200`	2024-01-15 07:12:12 -08:00
Chi_Liu	c7452af4fa	[MLIR][ONNX] Add OnnxToTorch support for Maxpool Op (#2695 ) Add Maxpool ONNX op support. Add Utils.h/cpp files to create a constant int list for ONNX.	2024-01-12 14:54:38 -08:00
Ze Zhang	670a99ae19	Handle torch.none type in tosa.clamp op (#2739 ) This PR updates the torch-to-tosa conversion with following changes: - Support torch.none as min/max input argument for tosa.clamp op - Support negative value as start index for tosa.slice op - Add tosa.logical_or lowering support e2e test: python -m e2e_testing.main --config=tosa LIT tests: cmake --build build --target tools/torch-mlir/all --------- Co-authored-by: Ze Zhang <ze.zhang@getcruise.com>	2024-01-11 10:36:48 -08:00
Andreas Falkenberg	5862854bc8	[ONNX][TORCH-MLIR] LayerNorm (#2716 ) Layer Normalization using the torch.aten.native_layer_norm https://github.com/nod-ai/SHARK-Turbine/issues/325	2024-01-11 14:27:04 +05:30
kumardeepakamd	29569713f3	support for onnx.expand operator (#2729 ) maps onnx.expand to torch aten broadcast_to, three tests added --------- Co-authored-by: Kumar Deepak <kumar@xilinx.com>	2024-01-10 13:05:37 -08:00
Vivek Khandelwal	208ae35583	[MLIR][ONNX] Add TorchToOnnx Support for DepthToSpace op Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-01-10 17:50:47 +05:30
Vivek Khandelwal	4707d3bdc6	[MLIR][ONNX] Add OnnxToTorch support for Bernoulli and CastLike op Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-01-10 16:24:06 +05:30
Vivek Khandelwal	35e8f86792	[MLIR][ONNX] Add OnnxToTorch support for Dropout and Elu op Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2024-01-10 16:23:55 +05:30
John Wu	4e5e34d215	[MLIR][ONNX] Add OnnxToTorch support for Slice Op (#2696 )	2024-01-03 19:41:10 -08:00
Xida Ren (Cedar)	1778314620	add basic cumsum. this doesn't support the exclusive and reverse attrs (#2717 ) fixes #2711	2024-01-03 09:52:59 -08:00
Xida Ren (Cedar)	9fc212ea9a	support Onnx opset 1-13 ReduceMean where axes is supplied as an attr (#2703 ) (instead of an input) Addresses part of #2689. fixes #2702	2023-12-28 09:31:41 -08:00
Xida Ren (Cedar)	d560698e3d	Lower `onnx.split` to `torch.aten` (#2686 )	2023-12-27 17:53:07 -08:00
aldesilv	2d796b7502	lower onnx max op to torch aten maximum op (#2618 ) lower onnx min op to torch aten minimum op	2023-12-27 11:07:35 -08:00
aldesilv	336cfb64b5	OnnxToTorch support for onnx.Mul op (#2699 )	2023-12-27 10:50:08 -08:00
Xida Ren (Cedar)	6847fc1fc6	Fix since-opset too high (#2701 ) Addresses two of the ops from https://github.com/llvm/torch-mlir/issues/2689 https://github.com/llvm/torch-mlir/issues/2700	2023-12-27 10:08:09 -08:00
aldesilv	abc6b0a25a	onnx to torch pow support (#2656 )	2023-12-27 09:34:48 -08:00
Vivek Khandelwal	4f252c88b4	[MLIR][ONNX] Add OnnxToTorch support for GlobalAveragePool op. (#2692 ) This commit adds the OnnxToTorch support for GlobalAveragePool op. Signed-Off By: vivekkhandelwal1424@gmail.com	2023-12-26 10:25:31 -08:00
saienduri	ee75e8d1ae	[MLIR][ONNX] Add OnnxToTorch support for Reshape Op (#2698 ) This commit adds the OnnxToTorch support for Reshape op.	2023-12-26 10:20:13 -08:00
Vivek Khandelwal	0849fd0a06	[MLIR][ONNX] Fix onnx.conv lowering to handle bias tensor Signed-Off By: Vivek Khandelwal <vivekkhandelwal1424@gmail.com>	2023-12-22 16:36:21 +05:30
Vivek Khandelwal	9a72c6584e	[MLIR][ONNX] Add OnnxToTorch support for BatchNormalization and Concat op. This commit adds the OnnxToTorch support for BatchNormalization and Concat op. Signed-Off By: vivekkhandelwal1424@gmail.com	2023-12-22 11:25:33 +05:30
Stella Laurenzo	ccd469ca0d	[fx] Upstream the turbine FxImporter to torch-mlir. (#2681 ) Changes made during upstreaming: * Removed comments attributing some copied code back to torch-mlir (since it is now repatriated). * Re-organized imports. * Inlined RefMapping/RefTracker and TypeSubclassMap from an external utility module. * Added FxImporter class comments. * Updated stack trace extraction to be fail safe. * Added an entry-point for `import_frozen_exported_program` which uses the shiny new upstream `torch.export.export()` API (versus the lower-level/older API that Turbine is presently using). This necessitated a small FX rewrite to line external state management up with current conventions. * Adapted one of Turbine's importer tests to go with this initial submission. Turbine unfortunately has a lot of more-integration-ey tests, and I would like to extract those as more of unit tests of the importer features and upstream them that way vs trying to copy directly. For now, one overall test with the initial submission gets us moving. I acknowledge that there are some code quality things that could be improved in this submission: this was authored over the course of many months (and often via some trial and error). I would like to keep it relatively converged with the downstream for the next few steps while getting the test suite upstreamed. And then it will be easier to take a hygienic pass through the code. Including co-authors for contributors in the git log of the original repository. Co-authored-by: Ean Garvey <87458719+monorimet@users.noreply.github.com> Co-authored-by: Avinash Sharma <aviator1994@gmail.com> Co-authored-by: Arham Khan <arhammkhan@gmail.com> Co-authored-by: brucekimrokcmu <kwangkyk@alumni.cmu.edu> Co-authored-by: saienduri <77521230+saienduri@users.noreply.github.com>	2023-12-21 08:40:10 -08:00
John Wu	46f2cb50dc	[onnx] Lower onnx.HardSigmoid to torch (#2682 ) The expression for HardSigmoid in Onnx (https://onnx.ai/onnx/operators/onnx__HardSigmoid.html): max(0, min(1, alpha * x + beta)) is inherently different from HardSigmoid in Torch (https://pytorch.org/docs/stable/generated/torch.nn.Hardsigmoid.html) which is: if x < -3 -> 0 elif x > 3 -> 1 else x/6 + 1/2 That being said, it was just better to compute out the entire expression when translating the Onnx expression to Torch mlir, which is done in this PR. Some of the logic is shared from the files in `DecomposeComplexOps`. Therefore, refactored some shared logic between `DecomposeComplexOps` and `DefaultDomainGToP` and put it in a `Utils` file.	2023-12-21 07:29:22 -08:00
Vivek Khandelwal	3226241521	[MLIR][ONNX] Add OnnxToTorch support for Conv and ConvTranspose op. This commit adds the OnnxToTorch support for Conv and ConvTranspose op. Signed-Off By: vivekkhandelwal1424@gmail.com	2023-12-21 11:12:14 +05:30
Rik Huijzer	8328998172	Allow printing all IR in `torch_mlir.compile` (#2669 ) This PR adds the `enable_ir_printing` option to `torch_mlir.compile`, which can be used to print the IR for all intermediate passes. When running the added test file via: ```shell $ python test/python/compile.py 2> tiny.stderr ``` the file `tiny.stderr` is about 700 KB.	2023-12-20 15:08:21 -06:00
Rob Suderman	11cc92d4ab	[onnx] Lowerings from `onnx.tan` (#2642 ) Started work on the `tan` lowerings for ONNX to Torch. Uses `sin` and `cos` to represent a `tan`.	2023-12-20 10:09:39 -08:00
Rob Suderman	a24aadbfab	[aten] Make `torch.aten.matmul` to `linalg` work for non-broadcasting case (#2659 ) Broadcasting for `torch.aten.matmul` is optional so a MxN with NxK matmul should be legalized to a `linalg.matmul`.	2023-12-20 10:09:10 -08:00
Andreas Falkenberg	ebaab4200f	[ONNX] ONNX -> TORCH for Erf (#2673 ) TorchOnnxToTorch For Erf function	2023-12-19 08:07:27 -08:00
Vivek Khandelwal	8649b84e3f	[MLIR][ONNX] Add OnnxToTorch support for AveragePool op. (#2672 ) This commit adds the OnnxToTorch support for AveragePool op. Signed-Off By: vivekkhandelwal1424@gmail.com	2023-12-18 18:17:11 -06:00
saienduri	698ff3a736	[MLIR][ONNX] Add OnnxToTorch support for Reduction Ops (#2657 ) This commit adds the OnnxToTorch support for ReduceSum, ReduceMean, and ReduceMin ops.	2023-12-18 12:37:31 -08:00
John Wu	deacb8ef38	[MLIR][ONNX] Add OnnxToTorch support for Gelu (#2647 ) This commit adds the OnnxToTorch support for Gelu op. --------- Co-authored-by: Rob Suderman <suderman@google.com>	2023-12-18 10:57:08 -08:00
Rob Suderman	791c666479	[torch] Lower `torch.aten.sinh` to `linalg` (#2662 )	2023-12-18 09:15:12 -08:00
Rob Suderman	ae1a6e4a5a	[onnx] Lower `onnx.Gemm` to `torch` (#2663 ) General lowering for `onnx.Gemm` to `torch`	2023-12-16 10:47:58 -08:00
Andreas Falkenberg	cee8563060	[onnx] Support of onnx.Greater, onnx.Less, onnx.GreaterOrEqual to Torch (#2649 ) The three remaining compare operations onnx.Greater onnx.Less onnx.GreaterOrEqual Are also added with this push request. This concludes a set of basic tensor compare functions.	2023-12-16 12:42:11 -05:00
Rob Suderman	61888690bb	[onnx] Add support for `onnx.sinh` (#2643 ) Adds a lowering from `onnx.sinh` to `aten.sinh`. This includes adding the `aten.sinh` operator.	2023-12-15 21:23:51 -08:00
Rob Suderman	705ea958ae	[onnx] Lowerings from `onnx.transpose` (#2641 ) Lowerings for `transpose` from ONNX to `aten`. Implementation depends on making multiple `aten.transpose` operations swapping pairs of dimensions. As `onnx.transpose` can swap around any dimensions it may require constructing multiple `aten.transpose`.	2023-12-15 15:30:05 -08:00
Quinn Dawkins	030b0140d4	[TorchToLinalg] Lower aten.cat to tensor.concat (#2650 ) This replaces the lowering of aten.cat with tensor.concat, allowing more efficient handling of concatenations in downstream flows. The refbackend populates concat decomposition patterns that can be used to recover the previous lowering.	2023-12-15 15:45:32 -05:00
Rob Suderman	061af696ce	[onnx] Lowering for `onnx.shape` to `torch` and `tensor` (#2648 ) Includes the lowering from the `aten` equivalent to `tensor` operations.	2023-12-15 11:37:49 -08:00
Gaurav Shukla	eb9249e601	[ONNX][MLIR] Add support for LeakyRelu and GatherElements op (#2655 ) This commit adds support for `LeakyRelu and GatherElements` op in the onnx pipeline. Signed-off-by: Gaurav Shukla <gaurav@nod-labs.com>	2023-12-15 11:18:28 -08:00
saienduri	f59c01fd2f	[MLIR][ONNX] Add OnnxToTorch support for q-z ops (specific ops in description) (#2601 ) This commit adds the OnnxToTorch support for Reciprocal, Round, ScatterElements, Sigmoid, Sin, Tanh, Sqrt, Sub, Sum, Where, Xor, Squeeze, Unsqueeze ops. For reviewers, the ops that weren't trivial and probably require extra review are Sum, Squeeze, and Unsqueeze.	2023-12-15 09:36:18 -08:00
Andreas Falkenberg	4ec8b9fc02	[onnx] add support for onnx.LessOrEqual (#2639 ) Added the less or equal operation to OnnxToTorch. onnx.LessOrEqual --------- Co-authored-by: root <andreas.falkenberg@amd.com>	2023-12-14 22:23:23 -05:00
Rob Suderman	4857606ffe	[onnx] Lowerings from `onnx.selu` (#2634 ) Lowerings for `selu` lowerings for ONNX to the corresponding torch implementations. Torch's `selu` implementation has fewer features so we use the a generalized `elu` with the input scale set to `1.0`.	2023-12-14 08:53:47 -08:00
John Wu	42392bc845	[MLIR][ONNX] Add OnnxToTorch support for matmul ops (#2629 ) This commit adds the OnnxToTorch support for Matmul.	2023-12-13 09:35:32 -08:00
Stella Laurenzo	ed4df38e8d	[onnx] Add torch-mlir-import-onnx tool. (#2637 ) Simple Python console script to import an ONNX protobuf to the torch dialect for additional processing. For installed wheels, this can be used with something like: ``` torch-mlir-import-onnx test/python/onnx_importer/LeakyReLU.onnx ``` Or from a dev setup: ``` python -m torch_mlir.tools.import_onnx ... ```	2023-12-12 22:01:30 -08:00
Stella Laurenzo	74f7a0c9d6	Upstream the ONNX importer. (#2636 ) This is part 1 of 2, which will also include upstreaming the FX importer. I started with ONNX because it forces some project layout updates and is more self contained/easier as a first step. Deviating somewhat from the RFCs on project layout, I made the following decisions: * Locating the `onnx_importer.py` into `torch_mlir.extras` as Maks already has opened up that namespace and it seemed to fit. Better to have fewer things at that level. * Setup the build so that the root project only contains MLIR Python and pure Python deps (like the importers), but this can be augmented with the `projects/` adding more depending on which features are enabled. * The default build continues to build everything whereas in `TORCH_MLIR_ENABLE_ONLY_MLIR_PYTHON_BINDINGS=1` mode, it builds a `torch-mlir-core` wheel with the pure contents only. `onnx_importer.py` and `importer_smoke_test.py` are almost verbatim copies from SHARK-Turbine. I made some minor local alterations to adapt to paths and generalize the way they interact with the outer project. I expect I can copy these back to Turbine verbatim from here. I also updated the license boilerplate (they have the same license but slightly different project norms for the headers) but retained the correct copyright. Other updates: * Added the ONNX importer unit test (which also can generate test data) in lit, conditioned on the availability of the Python `onnx` package. In a followup once I know everything is stable, I'll add another env var that the CI can set to always enable this so we know conclusively if tests pass. * Moved the ONNX conversion readme to `docs/`. * Renamed CMake option `TORCH_MLIR_ENABLE_ONLY_MLIR_PYTHON_BINDINGS` -> `TORCH_MLIR_ENABLE_PYTORCH_EXTENSIONS` and inverted the sense. Made the JitIR importer and LTC options `cmake_dependent_options` for robustness.	2023-12-12 19:02:51 -08:00
Frederik Harwath	b656c674ee	Implement e2e support for aten.acos op This depends on a change in the LLVM core repository which adds acos support to the MLIR Math dialect.	2023-12-12 10:52:02 +01:00
Vivek Khandelwal	0b4422a253	[MLIR][ONNX] Add OnnxToTorch support for bitwise and math ops This commit adds the OnnxToTorch support for BitwiseXor, BitwiseOr, Div, Equal, Cast, Ceil, Floor, Cos, and Clip op. This commit also adds the TorchToLinalg support for aten.clamp.Tensor and aten.clamp_min.Tensor op. Signed-Off By: vivekkhandelwal1424@gmail.com	2023-12-11 19:36:01 +05:30
Quinn Dawkins	141202bc01	[TorchToLinalg] Fix integer type handling for aten.mm (#2615 ) Despite aten.mm requiring the input and output types match, we still opt to maintain signedness semantics in case later passes try to do any sort of integer type narrowing.	2023-12-07 00:13:53 -05:00
Sambhav Jain	44f6942796	Bump LLVM and StableHLO (#2598 ) Bump LLVM to `5e5a22caf88ac1ccfa8dc5720295fdeba0ad9372` and StableHLO to `83f095e7217c897f1eccac5652600ceb944cb0e0`. Bazel GHA: https://github.com/sjain-stanford/torch-mlir/actions/runs/7027647674	2023-11-28 22:12:24 -08:00
Vivek Khandelwal	dc9ea08db5	[MLIR][ONNX] Add OnnxToTorch support for atan and bitwise ops This commit adds the OnnxToTorch support for Atan, Bitshift, BitwiseAnd, and BitwiseNot op. This commit also adds the TorchToLinalg support for AtenBitwiseLeftShiftTensorOp. Signed-Off By: vivekkhandelwal@nod-labs.com	2023-11-28 17:19:07 +05:30
Stella Laurenzo	e06efc5136	Initial TorchOnnxToTorch conversion pipeline. (#2585 ) Adds a pipeline to convert custom ops and metadata represented as `torch.operator` custom ops to corresponding `torch` ops where possible. This is part of a multi-part approach for building ONNX import in as a regular feature of torch-mlir. It is focused on the conversions vs the infra. We will end up maintaining a [pure-python importer](https://github.com/nod-ai/SHARK-Turbine/blob/main/python/shark_turbine/importers/onnx_importer.py) to go with this in torch-mlir, and we will also maintain test case generation utilities derived from it. I have left substantial documentation in the README of the conversion directory, including the recommended approach that we will take to keep building this out. (note that this organizes the code to coincide with the refactoring in #2442 versus the current flat arrangement)	2023-11-21 21:02:55 -08:00
Zhekun(Josh) Zhang	d67afa9e95	[Torch] Add fold rule for AtenMaskedFillTensorOp to AtenMaskedFillScalarOp (#2543 )	2023-11-21 13:26:17 +08:00
James Newling	647f2f5076	Additional tests for view lowering (#2584 ) The logic for lowering the aten view op to linalg is fairly complex. In this PR I have tried to follow all non-failing paths through the lowering and add unit tests where they're missing. There is 1 logical change to the lowering: redundant tensor.cast ops (same source and destination type) are folded.	2023-11-20 17:35:25 -08:00
Stella Laurenzo	5eae0adff1	Breakup python pytorch deps (#2582 ) This lifts the core of the jit_ir_importer and ltc out of the pt1 project, making them peers to it. As a side-effect of this layering, now the "MLIR bits" (dialects, etc) are not commingled with the various parts of the pt1 project, allowing pt1 and ltc to overlay cleanly onto a more fundamental "just MLIR" Python core. Prior to this, the Python namespace was polluted to the point that this could not happen. That "just MLIR" Python core will be introduced in a followup, which will create the space to upstream the FX and ONNX pure Python importers. This primary non-NFC change to the API is: * `torch_mlir.dialects.torch.importer.jit_ir` -> `torch_mlir.jit_ir_importer`. The rest is source code layering so that we can make the pt1 project optional without losing the other features. Progress on #2546.	2023-11-19 12:10:19 -08:00
James Newling	dad1f012f6	Add verification for torch permute op (#2551 ) - adds support for an optional verifier to the generated torch op tablegen (GeneratedTorchOps.td) - uses the above to add a verifier for the torch permute op. Motivation: I hit an unclear error from linalg while developing a decomposition pass for pixel_shuffle. The error would have been clearer if the problem had been detected earlier in the invalid aten.permute op. Testing: new tests added. To run added tests, from the base directory run ``` ./build/bin/llvm-lit test/Dialect/Torch/invalid.mlir ```	2023-11-15 11:47:54 -08:00
Yuanqiang Liu	3ab790c50a	[Torch Dialect] add canonicalize for aten.numel (#2562 )	2023-11-11 12:16:53 +08:00
Stella Laurenzo	6961f0a247	Re-organize project structure to separate PyTorch dependencies from core project. (#2542 ) This is a first step towards the structure we discussed here: https://gist.github.com/stellaraccident/931b068aaf7fa56f34069426740ebf20 There are two primary goals: 1. Separate the core project (C++ dialects and conversions) from the hard PyTorch dependencies. We move all such things into projects/pt1 as a starting point since they are presently entangled with PT1-era APIs. Additional work can be done to disentangle components from that (specifically LTC is identified as likely ultimately living in a `projects/ltc`). 2. Create space for native PyTorch2 Dynamo-based infra to be upstreamed without needing to co-exist with the original TorchScript path. Very little changes in this path with respect to build layering or options. These can be updated in a followup without commingling directory structure changes. This also takes steps toward a couple of other layering enhancements: * Removes the llvm-external-projects/torch-mlir-dialects sub-project, collapsing it into the main tree. * Audits and fixes up the core C++ build to account for issues found while moving things. This is just an opportunistic pass through but roughly ~halves the number of build actions for the project from the high 4000's to the low 2000's. It deviates from the discussed plan by having a `projects/` tree instead of `compat/`. As I was thinking about it, this will better accommodate the follow-on code movement. Once things are roughly in place and the CI passing, followups will focus on more in-situ fixes and cleanups.	2023-11-02 19:45:55 -07:00
Zhekun(Josh) Zhang	88d4c475d3	[Torch] Fix mixP case for non value semantic ops (#2540 ) NonValueSemantic Ops like Add_, div_, etc. expect result DType to be the same as the first input. However, current implementation would result in wrong result type for case like: ```python a = torch.randn(3, 3).half() # float16 b = torch.randn(3, 3) # float32 a += b # i.e. torch.ops.aten.add_(a, b) ``` torch expects `a` to be float16, but dtype refinement would infer float32 type, since it's replaced by `aten.add`.	2023-11-02 12:40:08 +08:00
Daniel Garvey	4901773f77	add uncovered cases in view lowering (#2524 ) removes unecessary checks from empty strided	2023-11-01 21:56:44 -05:00
Ze Zhang	4279b750da	update AtenClampOp in torch-to-tosa to handle fp inputs (#2516 ) As titled. --------- Co-authored-by: Ze Zhang <ze.zhang@getcruise.com>	2023-10-17 14:49:47 -07:00
Chi_Liu	14a4da923b	Update llvm-project to b44b3494f60296db6aca38a14cab061d9b747a0a (#2511 ) The main purpose is to bring in the new mesh dialect change. https://github.com/llvm/llvm-project/pull/68007	2023-10-16 19:29:48 -07:00
Ze Zhang	f2c53b8ca5	Add aten.isclose support and its torch-to-tosa lowering (#2512 ) Add aten.isclose op Add its torch-to-tosa lowering Update the TorchToTosa/basic.mlir tests To test e2e tosa lowering: `python -m e2e_testing.main -v -c=tosa` --------- Co-authored-by: Ze Zhang <ze.zhang@getcruise.com>	2023-10-16 09:44:53 -07:00
Ze Zhang	e649e06b7b	Add aten.unflatten.int support and its torch-to-tosa lowering (#2509 ) Add aten.unflatten.int op Add its torch-to-tosa lowering Update the TorchToTosa/basic.mlir tests To test e2e tosa lowering: `python -m e2e_testing.main -v -c=tosa` --------- Co-authored-by: Ze Zhang <ze.zhang@getcruise.com>	2023-10-13 18:39:41 -07:00
Quinn Dawkins	6f81ad7293	[TorchToLinalg] Improve broadcast lowerings in strict symbolic modes (#2505 ) With strict symbolic shapes, we can assume numpy-style dynamic broadcasts never occur. This improves the lowering in the presence of this assumption.	2023-10-05 15:15:26 -04:00
Quinn Dawkins	ae72eec224	Improve aten.broadcast_to folder when in strict symbol mode (#2504 ) Strict symbolic shapes allow us to assume numpy-style dynamic broadcasts never occur. This allows us to strengthen the folder for broadcasts to cases where the rank is the same and all shapes match (including dynamic sentinel values).	2023-10-05 09:02:10 -04:00
Stella Laurenzo	860be09a39	Elide dynamic broadcast checks when in strict symbolic shapes mode. (#2496 ) When importing dynamic shaped programs from Dynamo, via torch.compile or torch.export, we can assume that strict symbolic shape checks have been done prior to generating torch IR. Among other shape checking, this eliminates the case where an unknown dimension can be dynamically '1' in a way that signals a broadcast. Adds a `isAssumingStrictSymbolicShapes` utility which consults a `torch.assume_strict_symbolic_shapes` attribute on an enclosing scope and returns true if present. In the linalg pipeline, many runtime checks are elided when this returns true.	2023-09-29 16:45:48 -07:00
Stella Laurenzo	a00a0d4bfb	Integrate llvm-project and mlir-hlo. (#2454 ) Corresponding commits: * mlir-hlo: 16886a108eff5197f816ca0f1950cc5ff1b078d9 * stablehlo: 77a59815a82b34f7b08ed2d42a711d9920682d0e * llvm-project: 4acc3ffbb0af5631bc7916aeff3570f448899647 * Adapt to ByteCodeOpInterface changes. * Adapt to RegionBranchPoint changes: https://reviews.llvm.org/D159116 * Adapt inferReturnTypes to get the value from properties. * Adapt invalid.mlir to properties syntax * [TOSA] Align with custom assembly format change. * [TOSA] handle change of axis to int32 type * [TOSA] Restore improper convert to i32 Landing with Windows broken (it cannot be fixed because of the way the mlir-hlo dep is inserted). Will followup with an untangling. --------- Co-authored-by: TatWai Chong <tatwai.chong@arm.com> Co-authored-by: Eric Kunze <eric.kunze@arm.com>	2023-09-12 15:09:57 -07:00
Bruce Kim	cd1c7df8be	[MLIR][TORCH] Add E2E support for view_as_real op (#2419 ) * view_as_real test case, allow dtype in testutils.randn * abstract python upstream func implemented * fixed upstream dtype func, implemented view_as_real backend op * formatted AtenViewAsRealOp, removed change in e2etest/framework * removed test suit from reshape_like.py, because it's moved to basic.py * implemented C-API wrapper for mlirComplexF128 type * fixed torch.complex dtype width in MLIR and Torch MLIR, deleted float16 dtype dict * Changed IR input of aten fft_fft unit test * code refactored * code refactored and fixed ci test * refactored: removed white spaces, and rolled back to having both input/output affine expr * refactored: deleted output affine expr to reduce redundancy * xfail ltc backend * removed ComplexImag and ComplexReal from torchdynamo xfail set * copied and pasted from main branch as there's no change to be made in this file * refactored abstract_interp_lib_gen.py * refactored: torchtypes.td, formatted, removed commented out code	2023-09-01 21:12:01 -07:00
Quinn Dawkins	1fc4314b62	Add folder for aten.broadcast_to on unchanged static shapes (#2421 )	2023-09-01 14:50:34 -04:00
JianzheXiao	17d02811d5	[Torch Dialect] add folder for aten.any.bool (#2388 ) * update * update * update * update * update * update * update	2023-08-30 17:29:03 +08:00
jinchen62	1682b540bf	Prototype passes for lowering quantized group matmul (#2402 ) * Support brevitas custom op (#2320) * f16 change for brevitas * Adapt the change of brevitas quant custom op name * Add unit tests * Make brevitas conversions isolated * Address the comments --------- Co-authored-by: dan <danimal197@gmail.com>	2023-08-29 21:25:45 -07:00
Jiawei Wu	4c9d234b01	revert canonicalizer for PrimListConstructOp (#2408 )	2023-08-22 09:18:39 +08:00
Jiawei Wu	60bad54f27	[Torch Dialect] replace none-index in aten.Index.Tensor's param by manually generating it (#2344 ) * [Torch Dialect] replace none-index in aten.Index.Tensor's param by manually generating it Co-authored-by: Jiawei Wu <wujiawei.aml@bytedance.com> Co-authored-by: Jianzhe Xiao <jianzhe.xiao@bytedance.com> * minor typo fix * add new failed e2e tests for ltc * fix typo * Address comments * Add more e2e tests * add failed e2e tests for LTC * address comments * remove decomposition for AtenIndexTensorHackedTwinOp	2023-08-15 19:36:08 +08:00
Ramiro Leal-Cavazos	ff762100b8	Add handling of namespaces to library generator (#2391 ) When using custom ops, sometimes PyTorch will insert namespaces to the abstract interpretation function name in the format: `__torch__.{namespace_1}.{namespace_2}...{op_name}`. The extra namespaces are not part of the abstract interpretation function name, so it needs to be removed before generating the library of MLIR snippets of abstract interpretation functions. This commit adds support for removing the namespace information.	2023-08-11 09:56:19 -07:00
Jiawei Wu	4c12aceb81	[Torch-Dialect] add canonicalizer for prim::ListConstruct op (#2306 ) [Torch-Dialect] add canonicalizer for prim::ListConstruct op	2023-08-08 10:28:11 +08:00
Gleb Kazantaev	fb52a73cbe	LTC->MLIR Debug Info support (#1922 ) * LTC->MLIR Debug Info support * SW-95317 Propagate Lazy->Jit->MLIR scope name. * Enhance location information based on op names Currently, the location information attached to the ops just considers the filename, line number and column number. Attaching operation name would help identify the type of computation by just looking at the profile of execution. * Update locations logic; updated debug-info.py test * Use {scope}/{op_name} format to track names by default --------- Co-authored-by: Gleb Kazantaev <gleb.kazantaev@cerebras.net> Co-authored-by: Mark Browning <mark@cerebras.net> Co-authored-by: Vimal Patel <vimal@polymagelabs.com>	2023-08-02 10:29:11 -04:00
Matthias Gehre	0a67411719	test/CAPI/CMakeLists.txt: Depend on FileCheck (#2329 ) I saw test failing when FileCheck wasn't already build	2023-07-25 10:11:55 +02:00
Matthias Gehre	c56cb531d5	Ignore constants in the legality error (#2328 )	2023-07-25 10:11:40 +02:00
Jiawei Wu	026e8db2e4	[Stablehlo] add converter for aten.scatter.src op (#2295 )	2023-07-24 10:14:45 +08:00
Alexandre Rames	1e468e8294	Fix canonicalization of `torch.prim.TupleUnpack`.	2023-07-20 20:08:46 +02:00
Alexandre Rames	a20422ce65	Support `DerefineOp` in `RefinePublicReturn`.	2023-07-20 20:08:46 +02:00
Alexandre Rames	4847563bed	Clean up verification of calling conventions. The implementation at this place was a remnent of the times the pipeline was run only once. Rely instead on the backend verification, after optimizations have had an opportunity to resolve some uncertainties. (e.g. `!torch.optional`).	2023-07-20 20:08:46 +02:00
Matthias Gehre	64d7626a52	Fixes for split tensor and slice (#2314 ) * RecomposeComplexOps: Remove dead slice op * lib/Dialect/Torch/IR/TorchOps.cpp: Fold slice ops even when they are on non-value tensors * lib/Conversion/TorchToTosa/TorchToTosa.cpp: Fix slice start/end out of range/none * lib/Dialect/Torch/IR/TorchOps.cpp: AtenSliceTensorOp::fold: Fold slices that go from 0:int_max * More tests for aten.split.Tensor	2023-07-20 09:53:54 +02:00
Jiawei Wu	3f843c8fd9	[torch-dialect] fix aten.type_as op's folder (#2283 ) [torch-dialect] fix torch.type_as op's folder by decomposing it to prim.dtype + aten.to_dtype	2023-07-20 09:51:58 +08:00
Ramiro Leal-Cavazos	718f53ff8a	Fix handling of `!torch.number` in abstract interpretation library (#2309 ) In PyTorch, the `NumberType` is equal to `Union[int, float, complex]`. However, the abstract interpretation library was treating the `NumberType` as `Union[int, float]`, resulting in type mismatches when reifying certain dtype functions. This commit fixes the type inconsistency by having the abstract interpretation functions take as an input a `Union[int, float, complex]` for the ops that take `!torch.number` inputs.	2023-07-17 09:52:04 -07:00
Jiawei Wu	c7fa42b7d3	[Torch Dialect] Add canonicalizer for aten.to.other op (#2273 ) Canonicalize aten.to.other to prim.device + prim.dtype + aten.to.device Co-authored-by: wujiawei.aml <wujiawei.aml@bytedance.com>	2023-06-30 09:43:08 +08:00
Yuanqiang Liu	449cfb8375	[Torch Dialect] add more scalar op folders (#2265 )	2023-06-29 10:37:13 +08:00
Yuanqiang Liu	1ea2b57ab7	[Torch Dialect] add folder for aten.add (#2264 ) * [Torch Dialect] add folder for aten.add * update * update * update	2023-06-27 10:55:28 +08:00
Yuanqiang Liu	96b14e952e	[Torch Dialect] Support aten.device.with_index (#2254 )	2023-06-23 01:07:14 +08:00
Vivek Khandelwal	f6a6cfea4e	[MLIR][TORCH] Add support for negative index values for index.Tensor op (#2233 ) This commit adds the support for index.Tensor op when the index values are negative. This commit wraps around the index values by checking their values at run time. Signed-Off By: Vivek Khandelwal <vivek@nod-labs.com>	2023-06-16 14:21:04 -05:00
Yuanqiang Liu	7c6961bcbf	[Torch Dialect] Support aten.cuda and add canonicalizer for aten.cuda (#2231 )	2023-06-14 09:56:39 +08:00
Maksim Levental	0caaf8d32a	Bump LLVM (#2176 ) * Bump LLVM --------- Co-authored-by: Matthias Gehre <matthias.gehre@xilinx.com>	2023-06-13 16:17:23 +02:00
Yuanqiang Liu	ddea56a832	[Torch Dialect] fix torch.uint8's dtype infer (#2227 )	2023-06-13 10:38:20 +08:00
Christopher McGirr	b461daa06e	fix(TorchToTosa.cpp): adjust torch->tosa div conversion (#2200 ) check the return type of the division to figure out whether to use the floating point implementation of a division or to use the integer. the issue rose from the fact that the inputs are all integer but the result was casted to floating point. The conversion then chose to use the integer implementation of division which is not legal in tosa when all the inputs get casted to floating point. fix(TorchToLinalg): AtenDivScalarOp upcast self operand as well if applicable, the self operand must also be casted to float as it can be an integer.	2023-06-12 11:18:38 +02:00
Matthias Gehre	27a3d09917	Torch: Fold RuntimeAssertOp when condition is true (#2198 )	2023-06-09 19:06:25 +08:00
Yuanqiang Liu	5a7bf4e4cb	[Torch Dialect] Add canonicalize pattern for aten.is_floating_point (#2194 ) * [Torch Dialect] Add canonicalize pattern for aten.is_floating_point * implement as fold * add lit test	2023-06-07 17:05:31 +08:00
Vivek Khandelwal	da886280fe	[MLIR][TORCH] Add E2E support for aten.tril op (#2202 ) Signed-Off By: Vivek Khandelwal <vivek@nod-labs.com>	2023-06-05 16:17:01 -07:00
Ramiro Leal-Cavazos	dff3405d5a	Add alias analysis for cast-like ops to maximize-value-semantics (#2160 ) When `use_tracing=True` is used to import a model into Torch-MLIR, several casts get inserted in the IR to bridge the untyped inputs and outputs with the typed body of the computation. These casts create extra aliases of tensors that cause the current analysis in `maximize-value-semantics` to fail. In particular, the `maximize-value-semantics` analysis assumes that the only valid alias right after an overwrite is the overwritten alias. So, if there is a use of a casted version of the overwritten alias after the overwrite, the analysis fails. This commit improves the analysis by identifying all cast-like aliases of the overwritten alias and allowing such aliases to be used after an overwrite. Because this issue only arises when using tracing, it cannot be currently tested e2e, so only lit test is added.	2023-05-25 17:05:41 +00:00
Zhekun Zhang	f0b7b63be0	[Stablehlo] Add aten.uniform lowering (#2101 ) * add uniform stablehlo lowering * add unit test * new line * rm redundant file * Empty commit, trigger test * fix include * address comments --------- Co-authored-by: zhekun.zhang <zhekun.zhang@bytedance.com>	2023-05-25 10:32:55 +08:00
TatWai Chong	ed4ecb072f	[tosa] support lowering basic torch binary ops with mixed dtypes (#2122 ) Lowering torch operations that allow different compatible data types in its operands to tosa end up generating invalid tosa IR with mixed data types. In tosa spec, certain operations (generally element-wise operations) require all operands to have the same data type. Add wrapper functions for those element-wise tosa ops to perform op creation with type conversion if necessary.	2023-05-18 17:12:18 -07:00
Ramiro Leal-Cavazos	de02b56e17	Replace RefineTypes with dtype functions (#2105 ) This commit adds dtype functions for all the torch ops that did not previously have one and removes the pass `RefineTypes`, since the abstract interpretation library now takes care of all the dtype propagation. All dtype functions added are tested except for - `aten.embedding` - `aten._embedding_bag` - `aten.embedding_bag` These functions need a change to the testing framework to allow specifying the actual data inside the tensor used for testing. I will fix this in a follow up patch. Co-authored-by: Jiahao Li <liplus17@163.com>	2023-05-12 13:40:45 -07:00
Sean Silva	d7614c261d	Integrate LLVM LLVM: 26ee8947702d79ce2cab8e577f713685a5ca4a55 MHLO: 4805d8498dfb81566076f56f52273b426c1cc5bf Per: https://github.com/llvm/torch-mlir/issues/1178#issuecomment-1538492185	2023-05-09 10:14:27 -07:00
Chi_Liu	51e0a2c933	[Stablehlo] Add stablehlo support for aten.abs (#2068 ) Co-authored-by: AmosLewis <Amos_Lewsi@foxmail.com>	2023-05-08 22:13:00 -07:00
Yuanqiang Liu	ef6dae6ae2	[Linalg] fix lowering reduce max with -inf (#2097 )	2023-05-08 09:17:49 -07:00
Yuanqiang Liu	0096ceae2f	[Stablehlo] fix reduce max init_value with -inf (#2064 ) * [Stablehlo] fix reduce max init_value with -inf * update	2023-05-06 12:05:51 -07:00
Zhekun Zhang	0cf9ee340b	[Torch Dialect] Add to.dtype_layout canonicalize patterns (#2062 ) * add to.dtype_layout canonicalize patterns * update comment --------- Co-authored-by: zhekun.zhang <zhekun.zhang@bytedance.com>	2023-05-02 20:06:02 -07:00
Ramiro Leal-Cavazos	96d662647f	Fix import of constant bool tensor parameters (#2047 ) Bool tensors are represented in TorchScript as an array of `int8_t`s. However, when importing them into Torch-MLIR, the importer was assuming the array had `int32_t` elements, leading to the importer reading into memory that was out of bounds. This commit fixes the casting of the bool tensor.	2023-04-20 18:38:48 -07:00
Chi_Liu	f3d1eda09f	[TOSA] Add aten.abs support (#2032 )	2023-04-14 08:43:39 -07:00
Zhekun Zhang	1bd5747ca3	[StableHlo] Fix transposed convolution conversion (#2026 ) * fix conv bwd * fix * fix group case * clean up --------- Co-authored-by: zhekun.zhang <zhekun.zhang@bytedance.com>	2023-04-13 11:24:39 -07:00
Yuanqiang Liu	3e83a86354	[Torch Dialect] fix isValidSubtype with dynamic dim (#2018 )	2023-04-11 01:02:18 -07:00
Vivek Khandelwal	98747d09a8	[MLIR][TORCH] Add support for prims::view_of op This op does nothing and just returns the input operand as the result of the op. Signed-Off By: Vivek Khandelwal <vivek@nod-labs.com>	2023-04-11 07:58:10 +05:30
Vivek Khandelwal	e90ea3d7ab	[MLIR][TORCH] Extend implementation of aten._index_put_impl op. This commits adds the support for cases for index_put_op: 1.) where index is a 2-d tensor. 2.) where indices is a list of tensors and none, with exactly 2 non none tensors along the consecutive dimensions. This commit also adds a utility to compute the broadcast shape given the two input tensors. Signed-Off By: Vivek Khandelwal <vivek@nod-labs.com>	2023-04-05 14:04:30 +05:30
Alexandre Rames	d24fa71368	Minor fixes for `ConvertTorchConversionToMLProgram`. (#1991 ) * Only create the global seed variable if it does not exist already. * Make the pass a module pass. A func pass may not modify its parent op.	2023-04-04 09:09:58 -07:00
Yuanqiang Liu	c86f46bd70	[test] rename TorchToMhlo to TorchToStablehlo (#1995 )	2023-04-03 18:41:25 -07:00
Ramiro Leal-Cavazos	e0f301c890	Add `extra_library` kwarg to `torch_mlir.compile` (#1986 ) This commit adds the ability to specify extra abstract interpretation functions in `torch_mlir.compile` to use during type refinement. This allows users to easily add custom ops without having to interact with MLIR or C++ directly.	2023-03-30 09:20:19 -07:00
Ramiro Leal-Cavazos	d803ab4eeb	Cast `number` to `float` when shape function takes Scalar arg (#1978 ) To keep things simple in shape functions, `Scalar` inputs are considered `float`s. This means that when inserting the shape functions into the IR, we must cast any `!torch.number`s into `float`s so that the operand type matches the expected type in the shape function. This commit adds the cast from `Scalar` to `float`.	2023-03-28 09:30:31 -07:00
Maksim Levental	953ea39cb5	handles 2,3,4 from https://github.com/llvm/torch-mlir/issues/1963 (#1964 )	2023-03-24 21:50:01 -05:00
Michael Feliz	2389729fb9	Add support for aten_remainder in TorchToTosa (#1966 )	2023-03-23 17:55:58 -07:00
Ramiro Leal-Cavazos	eae3ff7f1c	Change dtype functions interface to take ints tuple for each tensor (#1965 ) The original design for the dtype functions outlined in https://github.com/llvm/torch-mlir/issues/1462 was unable to properly handle ops that take optional tensors as an input when the optional tensor has a value of None. By the time the op gets imported into torch-mlir, if an optional value is None, all information about the original type is lost from the op type signature, preventing torch-mlir from knowing if a value of None was from an optional tensor or not, which was crucial in the original design since each tensor argument must be turned into two separate arguments for the dtype function. This commit changes the interface to dtype functions such that each tensor turns into a tuple of two ints, the first representing the rank of the tensor and the second the dtype of the tensor. Since now there is a one-to-one correspondence between the operands of an op and the operands of its dtype function, there is no ambiguity about which operand of the op corresponds with which operand of the dtype function. To test the implementation, this commit defines dtype function for convolution op, which takes one optional tensor as an argument.	2023-03-23 11:05:39 -07:00
Sean Silva	c319a20828	Update to LLVM 029313cc979ae71877b65794b1063d4e51184cc8 - mergeBlockBefore -> inlineBlockBefore - move tosa-to-tensor pass ordering https://github.com/llvm/torch-mlir/issues/1178#issuecomment-1476217922	2023-03-21 04:16:20 -07:00
Matthias Gehre	aa5bcb3cf2	LowerToBackendContract: Explicitly error out on unimplemented operator (#1947 ) * LowerToBackendContract: Explicitly error out on unimplemented operator But only reject torch.operator when results are invalid. Otherwise it might be a custom op that the backend supports.	2023-03-20 16:27:08 +01:00
Ramiro Leal-Cavazos	d310bb12bd	Expand definition of tensor subtype to include shape/dtype info (#1929 ) Currently, the op `torch.tensor_static_info_cast` will not get canonicalized away if the result type has any shape or dtype information. This is because `isValidSubtype` only returns true when the tensor types being compared are exactly the same or the supertype has no shape and dtype information. Being unable to canonicalize away the `torch.tensor_static_info_cast` gets in the way of further optimizations, such as shape propagation. This commit improves `isValidSubtype` by adding logic that compares the shapes and dtypes of the two tensor types to determine of one type is indeed a valid subtype of the other. Fixes https://github.com/llvm/torch-mlir/issues/1926	2023-03-10 16:43:57 -08:00
Ziheng Jiang	dca2b8a40a	[TORCH] Improve type refinement for aten.cat. (#1908 ) * [TORCH] Fix type refinement for aten.cat. * Add test. * Address comments. * Update. * Update. * Update. * Update. * Update. --------- Co-authored-by: Ziheng Jiang <ziheng.jiang@bytedance.com>	2023-03-09 16:17:35 -08:00
Zhekun Zhang	1d3a7419c5	[Torch Dialect] add RSub, ScalarImplicit canonicalize (#1899 ) * add rsub, scalarimplit canonicalizer * reformat * address comments * fix bug * fix test * Update elementwise.py * resolve merge conflict * change to 3 * change to 3 * real fix * fix name * add torchdynamo fail test --------- Co-authored-by: zhekun.zhang <zhekun.zhang@bytedance.com>	2023-03-06 17:38:27 -08:00
Ramiro Leal-Cavazos	d30af8772b	Handle uninitialized lattice elements in RefineTypes (#1911 ) The data-flow analysis does not always propagate information to the entire graph. This results in some lattice elements being uninitialized. Currently the lattice elements are not checked to see if they are uninitialized before rewriting the graph, potentially resulting in invalid IR (see https://github.com/llvm/torch-mlir/issues/1896). This commit adds handling for uninitialized lattice elements.	2023-03-03 08:55:58 -08:00
Yuanqiang Liu	7a8304f935	[Torch Dialect] add folder for aten.sub.float (#1871 )	2023-03-02 09:07:33 -08:00
Yuanqiang Liu	fc1e091d6a	[Torch Dialect] add aten.pow.int_float op and it's folder (#1872 )	2023-02-28 09:36:05 -08:00
Maksim Levental	2eddb3fde7	WIP: No PyTorch dep (#1854 )	2023-02-13 14:21:06 -06:00
Yuanqiang Liu	6ab990e1e8	[Torch Dialect] add folder for aten.Int.float (#1863 )	2023-02-10 13:59:03 -08:00
Yuanqiang Liu	2f6fdb7f0b	[Torch Dialect] add folder for prim.min.int (#1864 )	2023-02-10 13:58:15 -08:00
Ashay Rane	711646d095	mhlo: migrate conversion to stablehlo (#1840 ) This patch replaces all MHLO operations with their StableHLO counterparts and adds a validation pass to ensure that no MHLO operations remain before translating all Stablehlo operations to the MHLO dialect for further lowering to the Linalg dialect. This patch also updates all lit tests so that they refer to the `convert-torch-to-stablehlo` pass and so that they check for StableHLO operations.	2023-02-02 07:29:47 -06:00
Chi_Liu	00fc14a6e1	[TOSA] Add to.dtype i1 to i64 (#1830 )	2023-01-27 09:21:06 -08:00
Gleb Kazantaev	3930588a7e	Enable VerifyBackendContract in LTC backend (#1798 ) * Enable VerifyBackendContract in LTC backend * Update VerifyBackendContract pass * Move convert_scalar_implicit to jit_utils * Rename VerifyBackendContract to VerifyBackendContractNoDecompositions * Update verify-backend-contract-error.mlir test	2023-01-24 22:14:17 -05:00
Ramiro Leal-Cavazos	6c86bec04f	build: update llvm tag to 9acc2f37 (#1828 ) This commit makes the following changes: - Update dialects to use fold API `kEmitFoldAdaptorFolder` and update signature of `fold` methods (see PSA https://discourse.llvm.org/t/psa-new-improved-fold-method-signature-has-landed-please-update-your-downstream-projects/67618) - Replace `makeArrayRef` with `ArrayRef` (see https://reviews.llvm.org/D140896) - Remove `TypeRange{}` arg from `b.create<scf::IfOp>` since builder no longer takes that argument - Make `func`s in `Torch/invalid.mlir` private, since symbol declarations cannot be public. (see https://discourse.llvm.org/t/rfc-symbol-definition-declaration-x-visibility-checks/2140)	2023-01-25 01:29:42 +00:00
Maksim Levental	8696752eb6	Expose metadata of torch-mlir types (plus verify DictType key and value types). (#1785 )	2023-01-16 10:25:02 -06:00
Ashay Rane	4e4a571104	[TOSA] Add LeakyReLU conversion pass (#1790 ) * feat(TorchToTOSA): LeakyReLU legalization * test(LeakyReLU): Add LIT test and enable e2e test Co-authored-by: Philipp Braun <philipp.braun@amd.com>	2023-01-10 21:42:07 -08:00
Ashay Rane	0faba6d2fc	build: update llvm tag to de3f0f7f (#1789 ) Credit to @vivekkhandelwal1 for finding the necessary changes. Summary of changes: - Switch Tosa_IntArrayAttr[N], Tosa_IntArrayAttrUpto[N] to DenseI64ArrayAttr. - Replace kNoIterationLimit with kNoLimit. (https://reviews.llvm.org/D140525) - Add dependency on MhloPasses when MHLO is enabled - Specify result type when using mhlo::DotOp	2023-01-10 17:07:19 -06:00
Raghavan Raman	0979df6589	Fix unsqueeze in Torch to Tosa conversion (#1780 )	2023-01-10 11:09:58 -08:00
Ramiro Leal-Cavazos	273664ded6	[custom op] Replace `tanh` dtype function with `expm1` (#1769 ) This commit replaces the `tanh` dtype function, which was being used to test the implementation of dtype functions in `a710237437`, with a dtype function for `expm1`. The dtype function for `expm1` is identical to the `tanh` one, so the same level of testing is maintained. Currently, there are ops getting dtype information from the `RefineTypes` pass and ops getting dtype information from the `TorchDtypeRefinementPipeline`. Since each pass can only propagete dtype information for the ops it knows how to handle, some models with many ops handled in both passes require the two dtype propagation passes to execute many times, reaching the iteration limit set in the `LowerToBackendContractPass`. To temporarily avoid this issue while the migration to `TorchDtypeRefinementPipeline` is finished, this commit switches `tanh` to `expm1`, since the latter is used a lot less in large models.	2023-01-03 14:18:26 -08:00
Ashay Rane	ac780529b4	Revert e2e support for aten logical or/and/xor/not ops (#1757 ) This reverts commit `eaab9be207`, since it is causing the post-merge CI tests to fail, causing subsequent PRs to be blocked. Specifically, the tests `ElementwiseAtenLogicalAndOpPromoteBroadcastModule_basic` and `ElementwiseAtenLogicalXorOpPromoteBroadcastModule_basic` fail because the oracle does not match the computed result. This patch reverts the commit to make the post-merge builds green again.	2022-12-29 21:01:06 -06:00
Jiahao Li	eaab9be207	Add e2e support for aten logical or/and/xor/not ops (#1752 )	2022-12-26 10:23:38 +08:00
Jiahao Li	49071f86e6	[MHLO] Evaluate RuntimeAssertOp at compile time (#1732 )	2022-12-22 17:12:52 +08:00
Tanyo Kwok	297fd3aa47	Revert "reimplement linear lowering torchToMhlo (#1524 )" (#1744 ) This reverts commit `50b524546f`.	2022-12-21 21:24:07 -08:00
zzp_miracle	50b524546f	reimplement linear lowering torchToMhlo (#1524 )	2022-12-22 10:15:16 +08:00
Jiahao Li	15b249777b	[Torch][MHLO] Decompose aten.copy op. Lower aten.rsqrt & sigmoid to mhlo. (#1734 )	2022-12-22 10:13:59 +08:00
Chi_Liu	9dc09ac8c5	[TOSA] Add aten.gather support for tosa (#1680 )	2022-12-21 11:04:07 -08:00
Chi_Liu	b2cefc0b64	[TOSA] Add aten.masked_fill.Tensor/Scalar support (#1735 )	2022-12-21 08:56:07 -08:00
ataheridezfouli-groq	17ee643aeb	[TORCH] Add Complex Number support (#1673 ) Add Complex number dtype support to torch tensors. Add aten.fft_fft op to test complex numbers.	2022-12-15 21:40:01 +00:00
Ramiro Leal-Cavazos	60db793feb	Pass op legality info to `verifyBackendContractPass` (#1705 ) In order to verify if a given IR satisfies the backend contract, the verifier needs to know if decompositions took place, and if so, which ops were decomposed and which were not. This commit adds two arguments to `verifyBackendContractPass` to specify if decompositions took place and which ops to consider backend legal, similar to the arguments of `LowerToBackendContractPass`.	2022-12-15 08:32:52 -08:00
Ahmed S. Taei	b1f6832849	Add aten.slice.Tensor & aten.cat folders (#1691 )	2022-12-13 13:02:47 -08:00
Ramiro Leal-Cavazos	a710237437	[custom op] Generalize shape library logic to work with dtypes (#1594 ) * [custom op] Generalize shape library logic to work with dtypes This commit generalizes the shape library logic, so that dtype rules for ops can also be expressed using the same mechanism. In other words, each op can now have a shape function and a dtype function specified in Python that is imported during lowering to calculate the shapes and dtypes throught a program. For more information about how to specify a dtype function, see the updated `docs/adding_a_shape_and_dtype_function.md`. For those not familiar with how the shape library works, the file `docs/calculations_lib.md` provides an overview.	2022-12-13 08:25:41 -08:00
Chi_Liu	163d19cce6	[TOSA] Add aten.add/sub.Scalar/Tensor si64 type support (#1604 )	2022-12-12 12:13:07 -08:00
Ramiro Leal-Cavazos	a54b334578	Allow running DecomposeComplexOps more than once (#1671 ) The current implementation of `DecomposeComplexOps` fails if an op expected to be decomposed does not get decomposed in the first iteration of the `createTorchSimplificationPipeline` in `LowerToBackendContractPass`. However, some graphs require multiple iterations of `createTorchSimplificationPipeline` to fully propagate all statically knowable information, such as dtypes and shapes, to the entire graph, sometimes resulting in the need to run `DecomposeComplexOps` more than once. This commit changes `DecomposeComplexOps` to use a greedy algorithm for pattern application and moves the legalization check of ops to the `LowerToBackendContractPass` to allow for the `DecomposeComplexOps` to run more than once.	2022-12-08 09:26:38 -08:00
Ramiro Leal-Cavazos	76190e8a3f	Remove unnecessary decompose-complex-ops tests (#1693 ) This commit removes lit tests from the `decompose-complex-ops` that are essentially testing a macro expansion, in accordance with https://github.com/llvm/torch-mlir/blob/main/docs/architecture.md#dos-and-donts-for-unit-vs-end-to-end-testing .	2022-12-08 08:22:08 -08:00
Ramiro Leal-Cavazos	dd35488da5	build: update llvm tag to 798fa4b4 (#1684 ) - Support for non-prefixed accessors has been removed. See: https://reviews.llvm.org/D136727 - Rename `operands` to `methodOperands` in `prim.CallMethod` since the name `operands` overlaps with a builtin method name. See: https://reviews.llvm.org/D136727 - Add passes in refbackend to lower memref.subview. See: https://reviews.llvm.org/D136377 - Replace `CopyToValueTensorOps` first in `RewriteViewLikeSubgraph` in maximize-value-semantics. The current implementation of the `RewriteViewLikeSubgraph` pass in maximize-value-semantics creates temporarily invalid IR. In particular, given a forward slice starting from a `CopyToNonValueTensorOp` and ending in `CopyToValueTensorOp`s, the pass first replaces all uses of the `CopyToNonValueTensorOp` with its operand, which results in all the `CopyToValueTensorOp` users having their operand have type `!torch.vtensor`, which is invalid. The correct way to do things is to first replace all the `CopyToValueTensorOp`s with their operand, and then replace all uses of the `CopyToNonValueTensorOp` with its operand. This only started failing now because the generated accessor `getOperand` for the `CopyToValueTensorOp` now returns a `TypedValue<NonValueTensorType>`, which has an assert checking that the value returned is of the expected type.	2022-12-07 12:20:41 -08:00
Vivek Khandelwal	f416953600	[MLIR][TORCH] Add TorchConversionToMLProgram and MLProgramBufferize pass This commit changes the `InsertRngGlobalsPass` to `TorchConversionToMLProgram` pass. This commit also adds the `MLProgramBufferize` pass for the bufferization of ml_program dialect ops to run on refbackend. Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-12-02 13:20:46 +05:30
Vivek Khandelwal	e7edcc62fd	build: update llvm tag to 147fe9de Summary of changes: - Replace call to `MemoryEffectOpInterface::hasNoEffect` with `isMemoryEffectFree`. - Make fix for the dynamic dims, since `kDynamicSize` value changed to `std::numeric_limits<int64_t>::min()` from `-1` in llvm - `makeShapeLLVMCompatible` and `makeShapeTorchCompatible` utilities convert shapes in order to remain consistent with the Torch and MLIR semantics. - Update tags llvm: 147fe9de29dc13c14835127b35280c4d95c8e8ba mhlo: 1944b5fa6062ec4c065d726c9c5d64f1487ee8c5 Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-12-01 13:36:50 +05:30
Ramiro Leal-Cavazos	0983a7f93a	Fix modulus calculation in LCG algorithm of refbackend (#1658 ) The current implementation sets the `nextSeed` value to `temp & 127`, which is wrong. The last step of the LCG algorithm for the multiplier and increment chosen should be `temp % 2^{64} = temp & (1 << 63)`. However, because we are dealing with i64 values, the modulus operation happens automatically, so it is not needed. See Donald Knuth's values for LCG here: https://en.wikipedia.org/wiki/Linear_congruential_generator	2022-11-30 08:46:52 -08:00
Tanyo Kwok	bbcdb38d99	Revert "Decompose torch.slice_scatter (#1622 )" (#1659 ) This reverts commit `f3f2f10030`.	2022-11-30 12:47:13 +08:00
Vivek Khandelwal	d9cbf01d1e	Revert "build: update llvm tag to 147fe9de" This reverts commit `e45ad313d4`.	2022-11-25 12:41:56 +05:30
Vivek Khandelwal	e45ad313d4	build: update llvm tag to 147fe9de Summary of changes: - Update call to `hasNoEffect` utility - `KDynamicSize` value changed to `std::numeric_limits<int64_t>::min()` from `-1` - Update tags llvm: 147fe9de29dc13c14835127b35280c4d95c8e8ba mhlo: 1944b5fa6062ec4c065d726c9c5d64f1487ee8c5 Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-11-24 12:44:43 +05:30
Tanyo Kwok	f3f2f10030	Decompose torch.slice_scatter (#1622 ) * Decompose torch.slice_scatter * fix compilation error * update file check * fix ci * fix i64 torch.tensor dtype	2022-11-23 18:14:12 +08:00
Vivek Khandelwal	da8fdc9f96	[MLIR][TORCH] Fix refine types crash This commit fixes https://github.com/llvm/torch-mlir/issues/1599. Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-11-23 15:17:37 +05:30
Tanyo Kwok	4aad5ccf39	fix #1626 return type mismatch (#1634 )	2022-11-23 15:02:41 +08:00
Vivek Khandelwal	55c7e66aa7	[MLIR][TORCH] Fix mean and mean.dim op for large-sized inputs This commit fixes the aten.mean and aten.mean.dim op decomposition for supporting large-sized inputs. This commit also fixes the formatting for the file stats.py Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-11-22 08:38:51 +05:30
Vivek Khandelwal	4cbd3927d7	[MLIR][TORCH] Add aten.sort.int op Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-11-20 19:00:41 +05:30
Chi_Liu	29c8f47723	[TOSA] Add aten.clamp op tosa support (#1609 ) Co-authored-by: AmosLewis <Amos_Lewsi@foxmail.com>	2022-11-18 13:32:13 -08:00
Sean Silva	39de4d6265	[cleanup] Make diagnostics better Also remove some unused imports.	2022-11-17 02:09:54 -08:00
Sambhav Jain	fc4c8d4ed9	Enable torch-mlir LIT tests in Bazel (#1585 ) Adds support to run `.mlir` LIT tests in bazel. ``` bazel test @torch-mlir//test/... ``` Follow-on PR will contain these updates: - Add tests to GHA CI workflow - Add `.py` LIT tests to bazel	2022-11-15 14:02:19 -08:00
Chi_Liu	dfe7513a45	[MLIR][TORCH] Fix aten.unsqueeze op (#1578 ) The range of the unsqueeze dim is: [-input.dim() - 1, input.dim() + 1), the bug forgets to add 1.	2022-11-14 09:09:15 -08:00
Daniel Ellis	a7ac0def45	Move single-tensor-tuple-return test to mlir unit test. Also, add multiple return test.	2022-11-10 09:23:53 -05:00
Xiafei Qiu	4f173c6e0f	update llvm tag to a2620e00. (#1567 ) - also update MHLO to 57ba12a2(branch greencommit/2022-11-07-a2620e00) - change -pass-pipeline format to make tests pass.	2022-11-10 18:39:28 +08:00
Tanyo Kwok	17bc7c89cc	build: update llvm tag to 74fb770d (#1539 ) * build: update llvm tag to 74fb770d This commit makes the following changes needed to update bump LLVM: + replace usages of `tensor::createPadScalarOp`, see https://reviews.llvm.org/D136493 + Update file checks	2022-11-01 15:27:09 +08:00
Ramiro Leal-Cavazos	b723186983	Remove all but one of valsem ops + move fill.Scalar to elementwise (#1531 ) This commit removes almost all of the valsem ops, since the value semantics version of the ops now exist in PyTorch. The only op missing is `aten.bernoulli_.float`. In addition, this commit also simplifies the implementation of `aten.fill.Scalar` by moving it to the pattern that converts elementwise ops.	2022-10-28 15:06:11 +00:00
Ashay Rane	a11ea93877	build: update llvm tag to f8b84268 (#1528 ) The only change required was to update a test to reflect the changes in https://reviews.llvm.org/D136541.	2022-10-26 15:33:53 -05:00
Vivek Khandelwal	ca87033d2f	[MLIR][TORCH] Add E2E support for aten.mse_loss op This commit adds decomposition for the `aten.mse_loss` op. Signed-Off By: Vivek Khandelwal <vivek@nod-labs.com>	2022-10-25 21:06:58 +05:30
Chi_Liu	ad6f5848cb	[MLIR][TORCH] Add TorchToTosa lowering for aten.where.self op (#1454 )	2022-10-18 09:39:39 -07:00
Ramiro Leal-Cavazos	82a3860e25	build: update llvm tag to 4546397e (#1502 ) This commit makes the following changes needed to update bump LLVM: - Replace `linalg.init_tensor` with `tensor.empty` (see: https://reviews.llvm.org/D135129) - Replace `NoSideEffect` with `Pure` (see https://reviews.llvm.org/D135505) - Replace `body` region accessor for `ReduceOp` and `ReduceWindowOp` with `getBody` - Fix incorrect use of `tosa::ReduceSumOp` in `AtenNativeLayerNormOp` conversion pattern. The result type of `tosa::ReduceSumOp` must have the same rank as the input type. (see: https://www.mlplatform.org/tosa/tosa_spec.html#_reduce_sum) Co-authored-by: Ashay Rane <ashay@users.noreply.github.com> Co-authored-by: Ashay Rane <ashay@users.noreply.github.com>	2022-10-18 04:22:53 +00:00
Gaurav Shukla	da90a25f90	[MLIR][TORCH] Add E2E support for `aten.[div.int\|bitwise_or.Tensor]` ops This commit adds lowering of `aten.div.int` and `aten.bitwise_or.Tensor` ops. Both these ops are required in order to support bloom_560m model. Signed-Off-by: Gaurav Shukla <gaurav@nod-labs.com>	2022-10-10 22:28:51 +05:30
Vivek Khandelwal	d3cc3f1aff	[tosa] Add lowering for aten.to.dtype and aten._to_copy op This commit adds the TorchToTosa lowering for `aten.to.dtype` and `aten._to_copy` op. Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-10-06 12:00:25 +05:30
Vivek Khandelwal	56f9a9b5de	[tosa] Add TorchToTosa lowering for torch.prim.NumToTensor.Scalar op Signed-Off By: Vivek Khandelwal<vivek@nod-labs.com>	2022-10-06 12:00:25 +05:30

... 4 5 6 7 8 ...

961 Commits (main)