Commits · 6d8e4c53c418697dcd7a68b5f95080382b716a01 · gaoqiong / MIGraphX

31 May, 2022 1 commit
- Merge branch 'jit-vector-softmax' into jit-layernorm · 6d8e4c53
  Paul authored May 31, 2022
  
  6d8e4c53
30 May, 2022 1 commit
- Merge branch 'develop' into jit-vector-reduce · e9885eb3
  Paul Fultz II authored May 29, 2022
  
  e9885eb3
27 May, 2022 1 commit
- renamed to main from master (#1226) · d436a723
  Chris Austen authored May 26, 2022
  
  d436a723
26 May, 2022 3 commits

Parallelize evaluations in propagate_constant (#1220) · bf603a76

shivadbhavsar authored May 26, 2022

Addressing issue #1166 - propagate_constant pass currently uses a recursive approach to find all instructions in a module that can be evaluated to a literal and performs the replacement in the same call.

New approach:

Perform single pass though instructions in the module to determine which instructions can be evaluated
Evaluate selected instructions in parallel
Replace the selected instructions with the corresponding literal

bf603a76

Merge branch 'develop' into jit-vector-reduce · 35579249
Paul Fultz II authored May 26, 2022

35579249
Upgrade to cppcheck 2.8 and fix new issues found (#1225) · a401e72a
Paul Fultz II authored May 26, 2022
```
* Upgrade to cppcheck 2.8
```
a401e72a

25 May, 2022 8 commits

Add missing header · 124ed38d
Paul authored May 25, 2022

124ed38d
Merge branch 'jit-vector-reduce' into jit-vector-softmax · 19ea0bf9
Paul authored May 25, 2022

19ea0bf9

Merge branch 'jit-vector-reduce' of... · 85897b5a

Paul authored May 25, 2022

Merge branch 'jit-vector-reduce' of github.com:ROCmSoftwarePlatform/AMDMIGraphX into jit-vector-reduce

85897b5a

Format · 49e1e618
Paul authored May 25, 2022

49e1e618
Set kernel name · 85f22ffd
Paul authored May 25, 2022

85f22ffd
Merge branch 'develop' into jit-vector-reduce · a0c504f2
Paul Fultz II authored May 25, 2022

a0c504f2
Used wrong path to download the bertsquad-10.onnx model (#1221) · bd746ccf
Chris Austen authored May 25, 2022
```
raw is the download for the file, blob is the url for the github page.
```
bd746ccf

Bump tensorflow from 2.5.3 to 2.6.4 in /examples/nlp/python_bert_squad (#1219) · 4e18f991

dependabot[bot] authored May 25, 2022

Bumps [tensorflow](https://github.com/tensorflow/tensorflow) from 2.5.3 to 2.6.4.
- [Release notes](https://github.com/tensorflow/tensorflow/releases)
- [Changelog](https://github.com/tensorflow/tensorflow/blob/master/RELEASE.md)
- [Commits](https://github.com/tensorflow/tensorflow/compare/v2.5.3...v2.6.4

)

---
updated-dependencies:
- dependency-name: tensorflow
  dependency-type: direct:production
...
Signed-off-by: dependabot[bot] <support@github.com>
Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>
Co-authored-by: Chris Austen <causten@users.noreply.github.com>

4e18f991

24 May, 2022 5 commits

Merge branch 'develop' into jit-vector-reduce · 5d941047
Paul Fultz II authored May 24, 2022

5d941047
Improve applicable batched gemms (#1214) · bf0a4713
Paul Fultz II authored May 24, 2022
```
* Improve applicable batched gemms for bert
```
bf0a4713

Remove std references in runtime compilation (#1186) · 150d6d20

Paul Fultz II authored May 24, 2022

Remove std references in runtime compilation since these are not available when using hiprtc and the headers may not be available on the system

150d6d20

Fuse gemm add with pointwise fusions (#1213) · a500620e
Paul Fultz II authored May 24, 2022
```
* Fuse gemm add with pointwise fusions
```
a500620e

Fix onnx mean parsing for integral inputs (#1209) · d895104a

shivadbhavsar authored May 23, 2022

As described in #1196, the ONNX mean parser does not work correctly for integral types. This update fixes the issue by handling integral types separately, where summation is performed before division. Additional test cases have also been added for handling integral types.

d895104a

23 May, 2022 4 commits
- Merge · c4cd8b0a
  Paul authored May 23, 2022
  
  c4cd8b0a
- Format · 92324d57
  Paul authored May 23, 2022
  
  92324d57
- Vectorize softmax · 4d66f031
  Paul authored May 23, 2022
  
  4d66f031
- Merge branch 'jit-vector-reduce' into jit-vector-softmax · 17f4ba28
  Paul authored May 23, 2022
  
  17f4ba28
20 May, 2022 2 commits
- Rename pointwise ops (#1145) · 4a312201
  kahmed10 authored May 20, 2022
```
For clarity on kernel names found when profiling. The new names are set to the order of the ops being compiled. For example: add + relu = add_relu_kernel.
```
  4a312201
- Improve matching with has_value when there are convert operators (#1212) · 27af0170
  Paul Fultz II authored May 19, 2022
  
  27af0170
19 May, 2022 1 commit
- Fix perf regression · c84154b8
  Paul authored May 19, 2022
  
  c84154b8
18 May, 2022 1 commit
- Fix tidy issue · 7133eee6
  Paul authored May 18, 2022
  
  7133eee6
17 May, 2022 4 commits
- Merge branch 'develop' into jit-vector-reduce · d9a5acbd
  Paul Fultz II authored May 17, 2022
  
  d9a5acbd
- Format · d0b7fc9a
  Paul authored May 17, 2022
  
  d0b7fc9a
- Fix wrong global size · 5515c9a5
  Paul authored May 17, 2022
  
  5515c9a5
- renamed variables for module from p to m (#1204) · a27dd28c
  shivadbhavsar authored May 17, 2022
```
Updated variable names according to #1193
```
  a27dd28c
13 May, 2022 1 commit

Update install_prereqs.sh for individual use (#1197) · 8c94ad07

Chris Austen authored May 13, 2022

Our documentation indicates a user with sudo can run the install_prereqs.sh file. Turns out that the file is not complete enough to run on Ubuntu 18.04/20.04 independently. I updated the file to resolve the failures.

resolves #1191

8c94ad07

12 May, 2022 3 commits
- Fix vec_reduce · b4c4234d
  Paul authored May 12, 2022
  
  b4c4234d
- Fix div by zero · 172f47f5
  Paul authored May 12, 2022
  
  172f47f5
- Fix tidy · 8344791c
  Paul authored May 12, 2022
  
  8344791c
11 May, 2022 5 commits
- Prefuse layernorm for gpu (#1190) · 671f24be
  Paul Fultz II authored May 11, 2022
```
Fuse layernorm and added triadd_layernorm fusion.  This is a prep performance booster
```
  671f24be
- Format · db2def39
  Paul authored May 10, 2022
  
  db2def39
- Fix vec issues · f1f60be1
  Paul authored May 10, 2022
  
  f1f60be1
- Format · c13780c2
  Paul authored May 10, 2022
  
  c13780c2
- Add vectorization to reduction · 15fd8205
  Paul authored May 10, 2022
  
  15fd8205