Commits · 3e067a8a1f2476487f71c8aefd34d2a0ed6d9197 · gaoqiong / MIGraphX

13 Oct, 2022 1 commit
- Fix trailing 1 shape mismatch with unsqueeze instead of outline · 3e067a8a
  Ted Themistokleous authored Oct 12, 2022
```
Fixes cases for trailing one testcases
```
  3e067a8a
12 Oct, 2022 2 commits
- Parse_if multi broadcast output for empty shape branches instead of adding outline · f24c65c3
  Ted Themistokleous authored Oct 12, 2022
```
This makes sure we're getting the correct output value of the branch if an
empty shape is given as a parameter to the IF operand.
```
  f24c65c3
- Remove debug test in parse_if operator · 9287e9c8
  Ted Themistokleous authored Oct 12, 2022
  
  9287e9c8
08 Oct, 2022 1 commit

Got model past if sequence but failing unit tests still · ea2d51bf

Ted Themistokleous authored Oct 07, 2022

- Gets past to the split section of the resnext model
- adding outline seems to solve if issues but verify calls broken
- Referencing wrong element now instead of output of correct if block?
- Need to determine proper output through verify tests.
- Modified protobuf to handle case of extra 1 to "vectorize" scalar
- Modified verify/tests to get things to "work", may need to be revised further.

ea2d51bf

07 Oct, 2022 1 commit

Get empty shapes working for parse_IF operator · abd3d63e

Ted Themistokleous authored Oct 07, 2022

- Update if_then/else_empty test protobuff and cases
- Need to update rand() vector used
- Make y empty instead of x for if_else_empty_test.onnx
- Regenerate protobufs with updates
- Add changes to handle empty/scalar input branch size to if operator.
- Add case where if both branches empty throw an error.
- Update verify tests with gold vectors and new shapes for empty input vec
  which we handle like a scalar before broadcasting

abd3d63e

05 Oct, 2022 2 commits

Add test files, protobufs and verification tests that capture errors with IF operator · 7c8c3bee

Ted Themistokleous authored Oct 05, 2022

- Verification tests that test each then/else branches for parsed IF operator
- Testing empty shape tensors for one branch -> output must be the other branch's shape
- Testing trailing 1 shape for one branch -> output must be union of both inputs

Current issue with IF operator is that we can't handle training vectors that match
in size correclty while also running into issues with empty inputs for one of the
branches for size/type checks.

7c8c3bee

Add additional test coverage for if_then case in verify · c1b0030b

Ted Themistokleous authored Oct 05, 2022

This seemed to be missing, just leveraging the existing protobuf made
to test parsing of if_then_test.onnx for this and using the tensor of all
ones to default to an ADD operation to ensure cond =1 is being handled and parsed
in correctly.

c1b0030b

04 Oct, 2022 13 commits
- Handle case to convert shapes for (x1,x2,x3,..xn) == (x1,x2,x3,..,xn) · 3e4991a6
  Ted Themistokleous authored Sep 14, 2022
```
Still busted work in prgoress. Keep running into

terminate called after throwing an instance of 'migraphx::version_1::exception'
  what():  /code/AMDMIGraphX/src/normalize_attributes.cpp:91: tune_attribute: TUNE_VECTOR: value out of range!
Aborted (core dumped)
```
  3e4991a6
- Fix unsqueeze on parse_if shape checks · 9199f074
  Ted Themistokleous authored Aug 22, 2022
  
  9199f074
- Only check types are equal, leave size checks till later. · 69977052
  Ted Themistokleous authored Aug 22, 2022
  
  69977052
- Attempt to fix scalar type and shape interpretation · 8b08a86f
  Ted Themistokleous authored Aug 04, 2022
```
Comming back to this once I've fixed parse_constant, looks like unallocated empty literals break this right now.
```
  8b08a86f
- Handle scalar conversion primarily empty scalars · 07e1755a
  Ted Themistokleous authored Aug 03, 2022
```
Need to have this to force the output to be a compatible to both the else/then cases

Still a work in progress
```
  07e1755a
- Handle static shapes only for conversion · 483c3e56
  Ted Themistokleous authored Aug 03, 2022
```
Need to handle static shapes explicitly since onnx says we should be able to just output a compatible output type.
```
  483c3e56
- fixup! First attempt at adding proper reshape for then/else modules in parse_if · 83215963
  Ted Themistokleous authored Aug 03, 2022
  
  83215963
- First attempt at adding proper reshape for then/else modules in parse_if · 7a141a27
  Ted Themistokleous authored Aug 03, 2022
  
  7a141a27
- Add more verbose loggign for if_op shape errors · dd6540bd
  Ted Themistokleous authored Aug 03, 2022
  
  dd6540bd
- Force scalars to be the same shape when comming out of an if then/else block · d799e44e
  Ted Themistokleous authored Aug 03, 2022
  
  d799e44e
- Fix parse_if cases for output result of if · db6e3a5c
  Ted Themistokleous authored Aug 02, 2022
```
The onnx spec mentions that the output shape of the resulting then/else branches must share the same type, but not the same shape. The only requirement is that the first dimension is compatible should one of the inputs have rank of one.

Without this we prematurely assert when an if is requred on the following case

int64, {1234, 1} (this is a 1 rank tensor)
int64, {1234}  (this result is scalar)
```
  db6e3a5c
- Stream sync Changset (#1358) · f7d987ba
  Ted Themistokleous authored Oct 04, 2022
```
Stream sync changes and associated API level changes
```
  f7d987ba
- Fast softmax (#1290) · a9a47402
  Paul Fultz II authored Oct 04, 2022
```
optimize the softmax operator
```
  a9a47402
03 Oct, 2022 1 commit

Add output_alias and runs_on_offload_target flags for the custom ops (#1309) · c9ffb38d

Umang Yadav authored Oct 03, 2022

Adds two methods for the custom_ops virtual class.

bool runs_on_offload_target(), if the custom op runs directly on the gpu then it should be set to true. in this case, custom op expects its parameters to reside in GPU memory and writes output to the GPU memory. If it is set to false then, custom op expects it's parameter to reside on the host and puts back the result into the host memory.

output_alias, if output of the custom op is aliasing the input buffer. i.e. interpreting the same input buffer with differnet shape and strides.

Update as_vector() in C++ API to handle non-standard shapes. It required exposing element_index to space_index conversion method for the shape class.

c9ffb38d

29 Sep, 2022 2 commits
- Use find_2.0 API for the convolution (#1346) · e19f78ae
  Umang Yadav authored Sep 29, 2022
```
Improvements/Additions to be made:

changes for the quant_convolution,
changes for the deconvolution,
Macros for MIOpen status checks
```
  e19f78ae
- Fix invalid program in debug mode from find_splits (#1390) · c2842c1e
  Paul Fultz II authored Sep 28, 2022
```
* Fix invalid program from find_splits
```
  c2842c1e
28 Sep, 2022 1 commit

Add compute_fp32 flag for quant_gemm tests (#1360) · 70e63960

Umang Yadav authored Sep 28, 2022

test_gpu_pack_int8_args fails on gfx908 machine, because it doesn't set compute_fp32 flag correctly. This PR fixes the test such that it checks for the device-name, and rocblas-versions and sets this flag accordingly.

70e63960

27 Sep, 2022 1 commit
- Add onnx mod operator gpu cpu (#1306) · 40118191
  Ted Themistokleous authored Sep 26, 2022
```
Implement operator for CPU and GPU implementations
```
  40118191
26 Sep, 2022 3 commits

Rewrite ONNX parse batch norm (#1362) · c00f8202

Charlie Lin authored Sep 26, 2022

Rewrites the BatchNormalization ONNX operator into other MIGX operators
- Added handling of 1D input tensor case (edge case in ONNX spec)
Removes the spatial and per_activation functionality (not in the ONNX spec)
- Did not remove the batch_norm_inference related code as the TensorFlow parser still uses it
- Can remove that code when the TF version is updated

c00f8202

Use larger vector size instead of preloading for broadcasted inputs (#1389) · 492c4a6c
Paul Fultz II authored Sep 26, 2022

492c4a6c
Upgrade cppcheck to 2.9 (#1400) · 66bbff1e
Paul Fultz II authored Sep 26, 2022
```
Upgrade cppcheck to 2.9 
```
66bbff1e

24 Sep, 2022 2 commits

check concurrency on PR level with one running and one pending performance tests (#1401) · 94bc41dc

Chris Austen authored Sep 24, 2022

Workflow has concurrency reintroduced with different set of rules. New expected behavior is to check concurrency on PR level with one running and one pending performance tests. In case of multiple commits in same PR, always the latest commit is queued after initiated performance test execution is completed. Any other PRs/commits are in pending/queued state

94bc41dc

update codecov version (#1402) · 1b575b5c
Chris Austen authored Sep 24, 2022
```
Codecov announced deprecating the bash uploader. Using updated uploader
```
1b575b5c

23 Sep, 2022 1 commit
- Remove unused device functions (#1394) · 8ea8473d
  Paul Fultz II authored Sep 23, 2022
```
* Remove device functions
* Update tests
```
  8ea8473d
21 Sep, 2022 2 commits

Parameterize epsilon for layernorm kernel (#1367) · d9578ba6

kahmed10 authored Sep 21, 2022

This PR allows for other values of epsilon to be matched when finding layernorm. Similarly, the calculation now uses the variable for epsilon.

d9578ba6

Multibroadcast find_mul_conv (#1384) · 9a70050b

Charlie Lin authored Sep 21, 2022

Change find_mul_conv to work with multibroadcast also. Checks the strides instead of the broadcast axis.

9a70050b

19 Sep, 2022 2 commits

Improve layernorm and reductions performance (#1348) · 97a1ed2d

Paul Fultz II authored Sep 19, 2022

Compute mean and variance in same reduction
Set block size to numbers divisible by 32 instead powers of 2
Global is also set exactly instead of being divisible by block size
More exact matching of global/local can help get rid of branching/loops
Reduce vectors first before doing dpp_reduce
Explicitly vectorize array operators since the compiler doesnt always vectorize them
Still uses old for loop when its computing at compile-time since the reinterpret_cast nor the all the vector types is supported

97a1ed2d

Disabled concurrency, queue added to perf-test.yml (#1386) · 34c08db7
Chris Austen authored Sep 19, 2022

34c08db7

16 Sep, 2022 2 commits
- Fix typo for add_sigmoid (#1385) · 10f37f49
  Umang Yadav authored Sep 16, 2022
```
* fix typo for add_sigmoid
```
  10f37f49
- Update deprecated Pybind constructor (#1382) · 255fb11a
  Umang Yadav authored Sep 16, 2022
```
* remove deprecated constructor
```
  255fb11a
15 Sep, 2022 1 commit

[mlir] Replaced `find_library` with `find_package` to locate MLIR static library (#1373) · e1e36cdc

Lixun Zhang authored Sep 15, 2022

* Replaced `find_library` with `find_package` to locate MLIR static library
* Unified the include dir for headers and remove backward compatibility
* Embedded the external/include dir into the exported library

e1e36cdc

14 Sep, 2022 2 commits
- Reduce problem size of unbatched_gemm tests (#1383) · 333860ce
  turneram authored Sep 14, 2022
```
The verify tests from pr #1354 were still causing some codecov timeouts after merge. This PR further reduces the problem sizes to avoid these failures.
```
  333860ce
- Fix split_reshape for slice len of 1 (#1379) · 4b76dd0d
  Umang Yadav authored Sep 14, 2022
```
* fix slice_dim1 for case
```
  4b76dd0d