Commits · 430d43e010bdd07d73c4d0d6536206d22d35a2cb · Lorenzo Albano / LLVM bpEVL

Jan 22, 2021

[mlir][Linalg] Disable fusion of tensor_reshape op by expansion when unit-dims are involved · 430d43e0

MaheshRavishankar authored Jan 22, 2021

Fusion of generic/indexed_generic operations with tensor_reshape by
expansion when the latter just adds/removes unit-dimensions is
disabled since it just adds unit-trip count loops.

Differential Revision: https://reviews.llvm.org/D94626

430d43e0

Add more explicit assert for failures · 73de3df1
Jacques Pienaar authored Jan 22, 2021
```
Differential Revision: https://reviews.llvm.org/D95201
```
73de3df1
[mlir][Linalg] Extend tile+fuse to work on Linalg operation on tensors. · 01defcc8
MaheshRavishankar authored Jan 22, 2021
```
Differential Revision: https://reviews.llvm.org/D93086
```
01defcc8

[mlir][Linalg] NFC: Refactor LinalgDependenceGraphElem to allow · bce318f5

MaheshRavishankar authored Jan 22, 2021

representing dependence from producer result to consumer.

With Linalg on tensors the dependence between operations can be from
the result of the producer to the consumer. This change just does a
NFC refactoring of the LinalgDependenceGraphElem to allow representing
both OpResult and OpOperand*.

Differential Revision: https://reviews.llvm.org/D95208

bce318f5

[mlir][spirv] Define spv.IsNan/spv.IsInf and add lowerings · e27197f3

Lei Zhang authored Jan 22, 2021

spv.Ordered/spv.Unordered are meant for OpenCL Kernel capability.
For Vulkan Shader capability, we should use spv.IsNan to check
whether a number is NaN.

Add a new pattern for converting `std.cmpf ord|uno` to spv.IsNan
and bumped the pattern converting to spv.Ordered/spv.Unordered
to a higher benefit. The SPIR-V target environment will properly
select between these two patterns.

Reviewed By: mravishankar

Differential Revision: https://reviews.llvm.org/D95237

e27197f3

[MLIR] Add support for extracting an integer sample point (if one exists) from... · 14056dfb

Arjun P authored Jan 22, 2021

[MLIR] Add support for extracting an integer sample point (if one exists) from an unbounded FlatAffineConstraints.

With this, we have complete support for finding integer sample points in FlatAffineConstraints.

Reviewed By: ftynse

Differential Revision: https://reviews.llvm.org/D95047

14056dfb

[mlir][StandardToSPIRV] Add support for lowering uitofp to SPIR-V · 2cb130f7

Hanhan Wang authored Jan 21, 2021

- Extend spirv::ConstantOp::getZero/One to handle float, vector of int, and vector of float.
- Refactor ZeroExtendI1Pattern to use getZero/One methods.
- Add one more test for lowering std.zexti which extends vector<4xi1> to vector<4xi64>.

Reviewed By: antiagainst

Differential Revision: https://reviews.llvm.org/D95120

2cb130f7

[mlir][Linalg] Introduce linalg.pad_tensor op. · 16d4bbef

Hanhan Wang authored Jan 21, 2021

`linalg.pad_tensor` is an operation that pads the `source` tensor
with given `low` and `high` padding config.

Example 1:

```mlir
  %pad_value = ... : f32
  %1 = linalg.pad_tensor %0 low[1, 2] high[2, 3] {
  ^bb0(%arg0 : index, %arg1 : index):
    linalg.yield %pad_value : f32
  } : tensor<?x?xf32> to tensor<?x?xf32>
```

Example 2:
```mlir
  %pad_value = ... : f32
  %1 = linalg.pad_tensor %arg0 low[2, %arg1, 3, 3] high[3, 3, %arg1, 2] {
  ^bb0(%arg2: index, %arg3: index, %arg4: index, %arg5: index):
    linalg.yield %pad_value : f32
  } : tensor<1x2x2x?xf32> to tensor<6x?x?x?xf32>
```

Reviewed By: nicolasvasilache

Differential Revision: https://reviews.llvm.org/D93704

16d4bbef

[mlir] Enable passing crash reproducer stream factory method · aee622fa

Jacques Pienaar authored Jan 21, 2021

Add factory to create streams for logging the reproducer. Allows for more general logging (beyond file) and logging the configuration/module separately (logged in order, configuration before module).

Also enable querying filename of ToolOutputFile.

Differential Revision: https://reviews.llvm.org/D94868

aee622fa

[mlir] Support FuncOpSignatureConversion for more FunctionLike ops. · 0a7a1ac7

mikeurbach authored Jan 18, 2021

This extracts the implementation of getType, setType, and getBody from
FunctionSupport.h into the mlir::impl namespace and defines them
generically in FunctionSupport.cpp. This allows them to be used
elsewhere for any FunctionLike ops that use FunctionType for their
type signature.

Using the new helpers, FuncOpSignatureConversion is generalized to
work with all such FunctionLike ops. Convenience helpers are added to
configure the pattern for a given concrete FunctionLike op type.

Reviewed By: rriddle

Differential Revision: https://reviews.llvm.org/D95021

0a7a1ac7

Jan 21, 2021

Add Python bindings for the builtin dialect · 922b26cd

Mehdi Amini authored Jan 20, 2021

This includes some minor customization for FuncOp and ModuleOp.

Differential Revision: https://reviews.llvm.org/D95022

922b26cd

Revert [mlir] Link mlir_runner_utils statically into cuda/rocm-runtime-wrappers (cf50f4f7) · bd3a387e
Christian Sigg authored Jan 21, 2021
```
There are cmake failures that I do not know how to fix.

Differential Revision: https://reviews.llvm.org/D95162
```
bd3a387e
Remove deprecated methods from OpState. · 8827e07a
Christian Sigg authored Jan 21, 2021
```
Reviewed By: rriddle

Differential Revision: https://reviews.llvm.org/D95123
```
8827e07a

[mlir]][SPIRV] Define OrderedOp and UnorderedOp and add lowerings from Standard. · 615167c9

MaheshRavishankar authored Jan 20, 2021

Define OrderedOp and UnorderedOp instructions in SPIR-V and convert
cmpf operations with `ord` and `uno` tag to these instructions
respectively.

Differential Revision: https://reviews.llvm.org/D95098

615167c9

[mlir][SPIRV] Rename OpSpecConstantOperation -> OpSpecConstantOp · 4234292e

MaheshRavishankar authored Jan 20, 2021

The SPIR-V spec uses OpSpecConstantOp. Using an inconsistent name
makes the dialect generation scripts fail. Update to use the right
operation name, and fix the auto generation scripts as well.

Differential Revision: https://reviews.llvm.org/D95097

4234292e

Add log1p lowering from standard to ROCDL intrinsics · 4ef38f9c
Frederik Gossen authored Jan 21, 2021
```
Differential Revision: https://reviews.llvm.org/D95129
```
4ef38f9c
Add log1p lowering from standard to NVVM intrinsics · 294e2544
Frederik Gossen authored Jan 21, 2021
```
Differential Revision: https://reviews.llvm.org/D95130
```
294e2544

[mlir] Remove complex ops from Standard dialect. · fc58bfd0

Alexander Belyaev authored Jan 20, 2021

`complex` dialect should be used instead.
https://llvm.discourse.group/t/rfc-split-the-complex-dialect-from-std/2496/2

Differential Revision: https://reviews.llvm.org/D95077

fc58bfd0

[mlir] Add a new builtin `unrealized_conversion_cast` operation · c78219f6

River Riddle authored Jan 20, 2021

An `unrealized_conversion_cast` operation represents an unrealized conversion
from one set of types to another, that is used to enable the inter-mixing of
different type systems. This operation should not be attributed any special
representational or execution semantics, and is generally only intended to be
used to satisfy the temporary intermixing of type systems during the conversion
of one type system to another.

This operation was discussed in the following RFC(and ODM):

https://llvm.discourse.group/t/open-meeting-1-14-dialect-conversion-and-type-conversion-the-question-of-cast-operations/

Differential Revision: https://reviews.llvm.org/D94832

c78219f6

[mlir] Add an interface for Cast-Like operations · 6ccf2d62

River Riddle authored Jan 20, 2021

A cast-like operation is one that converts from a set of input types to a set of output types. The arity of the inputs may be from 0-N, whereas the arity of the outputs may be anything from 1-N. Cast-like operations are removable in cases where they produce a "no-op", i.e when the input types and output types match 1-1.

Differential Revision: https://reviews.llvm.org/D94831

6ccf2d62

Jan 20, 2021

Revert "[mlir][Affine] Add support for multi-store producer fusion" · 735a07f0
Diego Caballero authored Jan 21, 2021
```
This reverts commit 7dd19885.

ASAN issue.
```
735a07f0

[mlir][sparse] add asserts on reading in tensor data · 5959c28f

Aart Bik authored Jan 20, 2021

Rationale:
Since I made the argument that metadata helps with extra
verification checks, I better actually do that ;-)

Reviewed By: penpornk

Differential Revision: https://reviews.llvm.org/D95072

5959c28f

[mlir:async] Fix data races in AsyncRuntime · a2223b09

Eugene Zhulenev authored Jan 20, 2021

Resumed coroutine potentially can deallocate the token/value/group and destroy the mutex before the std::unique_ptr destructor.

Reviewed By: mehdi_amini

Differential Revision: https://reviews.llvm.org/D95037

a2223b09

[mlir][Linalg] NFC - Fully compose map and operands when creating AffineMin in tiling. · 8dd58a50
Nicolas Vasilache authored Jan 20, 2021
```
This may simplify the composition of patterns but is otherwise NFC.
```
8dd58a50
[mlir] Add ComplexDialect to SCF->GPU pass. · b1e1bbae
Alexander Belyaev authored Jan 20, 2021

b1e1bbae

[mlir] Fix SubTensorInsertOp semantics · 866cb260

Nicolas Vasilache authored Jan 20, 2021

Like SubView, SubTensor/SubTensorInsertOp are allowed to have rank-reducing/expanding semantics. In the case of SubTensorInsertOp , the rank of offsets/sizes/strides should be the rank of the destination tensor.

Also, add a builder flavor for SubTensorOp to return a rank-reduced tensor.

Differential Revision: https://reviews.llvm.org/D95076

866cb260

[mlir][Linalg] NFC - Expose getSmallestBoundingIndex as an utility function · c0755726
Nicolas Vasilache authored Jan 20, 2021

c0755726
Avoid unused variable warning in opt mode · cad16e4a
Jacques Pienaar authored Jan 20, 2021

cad16e4a

[mlir][Affine] Add support for multi-store producer fusion · 7dd19885

Diego Caballero authored Jan 20, 2021

This patch adds support for producer-consumer fusion scenarios with
multiple producer stores to the AffineLoopFusion pass. The patch
introduces some changes to the producer-consumer algorithm, including:

* For a given consumer loop, producer-consumer fusion iterates over its
producer candidates until a fixed point is reached.

* Producer candidates are gathered beforehand for each iteration of the
consumer loop and visited in reverse program order (not strictly guaranteed)
to maximize the number of loops fused per iteration.

In general, these changes were needed to simplify the multi-store producer
support and remove some of the workarounds that were introduced in the past
to support more fusion cases under the single-store producer limitation.

This patch also preserves the existing functionality of AffineLoopFusion with
one minor change in behavior. Producer-consumer fusion didn't fuse scenarios
with escaping memrefs and multiple outgoing edges (from a single store).
Multi-store producer scenarios will usually (always?) have multiple outgoing
edges so we couldn't fuse any with escaping memrefs, which would greatly limit
the applicability of this new feature. Therefore, the patch enables fusion for
these scenarios. Please, see modified tests for specific details.

Reviewed By: andydavis1, bondhugula

Differential Revision: https://reviews.llvm.org/D92876

7dd19885

Added check if there are regions that do not implement the RegionBranchOpInterface. · 43f34f58

Julian Gross authored Jan 13, 2021

Add a check if regions do not implement the RegionBranchOpInterface. This is not
allowed in the current deallocation steps. Furthermore, we handle edge-cases,
where a single region is attached and the parent operation has no results.

This fixes: https://bugs.llvm.org/show_bug.cgi?id=48575

Differential Revision: https://reviews.llvm.org/D94586

43f34f58

[mlir] Link mlir_runner_utils statically into cuda/rocm-runtime-wrappers. · cf50f4f7

Christian Sigg authored Jan 11, 2021

The runtime-wrappers depend on LLVMSupport, pulling in static initialization code (e.g. command line arguments). Dynamically loading multiple such libraries results in ODR violoations.

So far this has not been an issue, but in D94421, I would like to load both the async-runtime and the cuda-runtime-wrappers as part of a cuda-runner integration test. When doing this, code that asserts that an option category is only registered once fails (note that I've only experienced this in Google's bazel where the async-runtime depends on LLVMSupport, but a similar issue would happen in cmake if more than one runtime-wrapper starts to depend on LLVMSupport).

The underlying issue is that we have a mix of static and dynamic linking. If all dependencies were loaded as shared objects (i.e. if LLVMSupport was linked dynamically to the runtime wrappers), each dependency would only get loaded once. However, linking dependencies dynamically would require special attention to paths (one could dynamically load the dependencies first given explicit paths). The simpler approach seems to be to link all dependencies statically into a single shared object.

This change basically applies the same logic that we have in the c_runner_utils: we have a shared object target that can be loaded dynamically, and we have a static library target that can be linked to other runtime-wrapper shared object targets.

Reviewed By: herhut

Differential Revision: https://reviews.llvm.org/D94399

cf50f4f7

[mlir][sparse] add narrower choices for pointers/indices · b5c542d6

Aart Bik authored Jan 19, 2021

Use cases with 16- or even 8-bit pointer/index structures have been identified.

Reviewed By: penpornk

Differential Revision: https://reviews.llvm.org/D95015

b5c542d6

[mlir][python] Swap shape and element_type order for MemRefType. · b62c7e04

Stella Laurenzo authored Jan 19, 2021

* Matches how all of the other shaped types are declared.
* No super principled reason fro this ordering beyond that it makes the one that was different be like the rest.
* Also matches ordering of things like ndarray, et al.

Reviewed By: ftynse, nicolasvasilache

Differential Revision: https://reviews.llvm.org/D94812

b62c7e04

Implement constant folding for DivFOp · 1bf2b166

Jackson Fellows authored Jan 19, 2021

Add a constant folder for DivFOp. Analogous to existing folders for
AddFOp, SubFOp, and MulFOp. Matches the behavior of existing LLVM
constant folding (https://github.com/llvm/llvm-project/blob/999f5da6b3088fa4c0bb9d05b358d015ca74c71f/llvm/lib/IR/ConstantFold.cpp#L1432).

Reviewed By: ftynse

Differential Revision: https://reviews.llvm.org/D94939

1bf2b166

Jan 19, 2021

[mlir][splitting std] move 2 more ops to `tensor` · be7352c0

Sean Silva authored Jan 14, 2021

- DynamicTensorFromElementsOp
- TensorFromElements

Differential Revision: https://reviews.llvm.org/D94994

be7352c0

[mlir][python] Add facility for extending generated python ODS. · 894d88a7

Stella Laurenzo authored Jan 19, 2021

* This isn't exclusive with other mechanisms for more ODS centric op definitions, but based on discussions, we feel that we will always benefit from a python escape hatch, and that is the most natural way to write things that don't fit the mold.
* I suspect this facility needs further tweaking, and once it settles, I'll document it and add more tests.
* Added extensions for linalg, since it is unusable without them and continued to evolve my e2e example.

Reviewed By: ftynse

Differential Revision: https://reviews.llvm.org/D94752

894d88a7

[mlir][python] Factor out standalone OpView._ods_build_default class method. · 71b6b010

Stella Laurenzo authored Jan 18, 2021

* This allows us to hoist trait level information for regions and sized-variadic to class level attributes (_ODS_REGIONS, _ODS_OPERAND_SEGMENTS, _ODS_RESULT_SEGMENTS).
* Eliminates some splicey python generated code in favor of a native helper for it.
* Makes it possible to implement custom, variadic and region based builders with one line of python, without needing to manually code access to the segment attributes.
* Needs follow-on work for region based callbacks and support for SingleBlockImplicitTerminator.
* A follow-up will actually add ODS support for generating custom Python builders that delegate to this new method.
* Also includes the start of an e2e sample for constructing linalg ops where this limitation was discovered (working progressively through this example and cleaning up as I go).

Differential Revision: https://reviews.llvm.org/D94738

71b6b010

[mlir][spirv] Define spv.GLSL.Fma and add lowerings · 3a56a966

Lei Zhang authored Jan 19, 2021

Also changes some rewriter.create + rewriter.replaceOp calls
into rewriter.replaceOpWithNewOp calls.

Reviewed By: hanchung

Differential Revision: https://reviews.llvm.org/D94965

3a56a966

[mlir][Affine] Revisit and simplify composeAffineMapAndOperands. · 93a873df

Nicolas Vasilache authored Jan 19, 2021

In prehistorical times, AffineApplyOp was allowed to produce multiple values.
This allowed the creation of intricate SSA use-def chains.
AffineApplyNormalizer was originally introduced as a means of reusing the AffineMap::compose method to write SSA use-def chains.
Unfortunately, symbols that were produced by an AffineApplyOp needed to be promoted to dims and reordered for the mathematical composition to be valid.

Since then, single result AffineApplyOp became the law of the land but the original assumptions were not revisited.

This revision revisits these assumptions and retires AffineApplyNormalizer.

Differential Revision: https://reviews.llvm.org/D94920

93a873df

[mlir] Add `complex.abs`, `complex.div` and `complex.mul` to ComplexOps. · 11f4c58c
Alexander Belyaev authored Jan 19, 2021
```
Differential Revision: https://reviews.llvm.org/D94911
```
11f4c58c