Commits · 701c20147e6f4c5cf28aa3c8f2a6e0eae4d3ba77 · gaoqiong / MIGraphX

11 Apr, 2022 2 commits

scatter operator refactoring to include reduction (#1124) · 701c2014

bpickrel authored Apr 11, 2022

Change the "scatter" struct and op to a base/child set of three: scatter_none, scatter_add, scatter_mul to mirror Onnx' ScatterElements op. and its three reduction options. (Onnx Scatter op is deprecated and is equivalent to scatter_none.)

Provides both a reference op. and update to Onnx parsing. Tests updated and new test case added.

701c2014

fix a bug in create tensor_view with vec data type (#1155) · 3c301efa

Shucai Xiao authored Apr 11, 2022

When create a tensor_view with vector date type, the last dimension of the shape should be divided by the vec_size.

3c301efa

08 Apr, 2022 1 commit
- Fix comparisons in migraphx::value class (#1146) · 1e0bbd78
  Paul Fultz II authored Apr 08, 2022
```
* Fix comparisons in migraphx::value class
```
  1e0bbd78
06 Apr, 2022 1 commit

Python Binding for the Manual Graph Buidling (#1143) · c4b6469a

Umang Yadav authored Apr 06, 2022

Adds following API binding and tests to python :

add_return
add_instruction
add_parameter
create_module.

c4b6469a

01 Apr, 2022 1 commit

Update developer overview, fix doc CMakeLists (#1140) · 0295965d

Charlie Lin authored Apr 01, 2022

* Fix and change doc CMakeLists
1. Fix include directory location with hange from #1088
2. Create a DoxygenWarningLog.txt file in <build_dir>/doc/doxygen
3. Move compiled html or pdf files to <build_dir>/doc/[pdf, html]

0295965d

31 Mar, 2022 1 commit
- Change the doc to mention only gpu or ref as targets (#1153) · c59f4079
  Umang Yadav authored Mar 31, 2022
```
Documentation update for valid targets
```
  c59f4079
29 Mar, 2022 3 commits

Python binding for shape : Fix constructor for the shape and enable tests (#1135) · b5c96d34
Umang Yadav authored Mar 29, 2022
```
Follow up to #1128
```
b5c96d34

Refactor runtime compiled kernels to use the same compile_ops pipeline (#1125) · 661046c6

Paul Fultz II authored Mar 29, 2022

This adds the infrastructure so we can compile everything in parallel, whereas before only pointwise kernels were compiled in parallel. This will also directly integrate with lowering and the gpu-driver. The kernels for pointwise and roialign are using this infrastructure. Scatternd is not since it does require standard shape.

This also makes it easier to add new runtime compiled kernels in the future.

661046c6

Remove Navi from CI temporarily (#1147) · 024b4abc
Chris Austen authored Mar 29, 2022
```
modify CI temporarily to stop using Navi hardware
```
024b4abc

28 Mar, 2022 2 commits

Use ifdef instead of comment for the auto-generated method declarations for... · 8e4d622f

Paul Fultz II authored Mar 28, 2022

Use ifdef instead of comment for the auto-generated method declarations for type erased classes (#1138)

It seems the formatting of comments are unreadable for larger methods, so instead just generate a struct with the methods in the interface and add a comment if its optional. It wraps this in #ifdef TYPE_ERASED_DECLARATION(assuming this would never be defined) instead of #if 0, so most editors can still provide syntax highlighting(although I think vscode with clangd will still gray it out unfortunately).

8e4d622f

Use ccache for runtime compilation (#1131) · ad056b1f
Paul Fultz II authored Mar 28, 2022
```
* Use ccache for runtime compilation
```
ad056b1f

25 Mar, 2022 1 commit
- Improve handling of string literals in value class (#1141) · c73c0dae
  Paul Fultz II authored Mar 25, 2022
```
* Handle string literal in construction
* Improve get_default with vector
```
  c73c0dae
24 Mar, 2022 1 commit
- Add initial experimental custom op (#1109) · 251cdd74
  Paul Fultz II authored Mar 24, 2022
```
This creates a custom op which has name() and compute_shape() methods. 
```
  251cdd74
22 Mar, 2022 1 commit
- Remove borrowed lifetime from operators that are no longer borrowing their lifetime (#1134) · cd165ebd
  Paul Fultz II authored Mar 22, 2022
```
Operators using arg.reshape() method the lifetime will be extended.
```
  cd165ebd
21 Mar, 2022 1 commit
- Lp normalization op (#1129) · 03225b57
  Charlie Lin authored Mar 21, 2022
```
* LpNormalization ONNX parser
```
  03225b57
18 Mar, 2022 2 commits

Complete GPU implementation of CumSum op (#1094) · 548783c8

turneram authored Mar 18, 2022

Add exclusive and reverse modes to gpu implementation of prefix_scan_sum, which completes support for ONNX op CumSum

548783c8

Make get_context experimental (#1137) · e521fa3f

Paul Fultz II authored Mar 18, 2022

The get_context may change in the future(when we support multi-targets) so make this experimental for now.

e521fa3f

15 Mar, 2022 2 commits

Expose APIs for the MIGraphX program (#1093) · 64e79a94

Umang Yadav authored Mar 15, 2022

API includes following
create_module,
get_main_module
add_instruction without module args
add_instruction with module args
add_parameter
add_return

64e79a94

Add iterators to kernels tensor_view and fix roialign to work with non-standard shape (#1126) · 31e63991

Paul Fultz II authored Mar 15, 2022

This adds iterators to tensor_view, which can allow kernels to work with non-standard shapes like for roialign.

To improve the performance of indexing when using the iterators, the shape class was updated to use integral_constants since the compiler doesn't always fold the const values. An integral_constant will at least enforce that in the AST.

Finally, since index calculations with single integers are improved, I also updated pointwise to use single index rather than multi index. There is about 4% improvement in some cases.

31e63991

14 Mar, 2022 3 commits
- Update gitignore to other comparable ROCm projects (#1117) · 2d1efd69
  Charlie Lin authored Mar 14, 2022
```
Have git ignore build directory and compiled python
Add ignores from other ROCm projects that look applicable
Ignore downloaded models in test/
Remove including visual studio settings
```
  2d1efd69
- Increase max groups in kernel (#1120) · d353641d
  Shucai Xiao authored Mar 14, 2022
```
change max number of groups in a kernel to 1B for greater performance
```
  d353641d
- Show the operator fields in the driver (#1103) · 9077db18
  Paul Fultz II authored Mar 14, 2022
```
* Show the operator fields in the driver
```
  9077db18
11 Mar, 2022 1 commit

Improve print ins (#1096) · b3b44f5d

Shucai Xiao authored Mar 11, 2022

The module::debug_print(ins) is very slow, which makes the trave_eval==1/2 very slow. The reason is printing an ins involves search the whole module to get the instruction, the print it.  This change is to fix that by calling module::print() to get names of all instructions of a program, then print the instruction by getting its name from a hash map.

b3b44f5d

09 Mar, 2022 3 commits
- Celu ONNX parser and tests (#1114) · 5b37c53c
  Charlie Lin authored Mar 09, 2022
```
Add Celu ONNX operator
```
  5b37c53c
- Add python API to construct shape class (#1128) · 4467c158
  Paul Fultz II authored Mar 09, 2022
```
Add python API to construct shape class
```
  4467c158
- Expose context in C++ API (#1118) · 0e6bd17c
  kahmed10 authored Mar 09, 2022
```
Add a callable C++ API to migraphx
```
  0e6bd17c
08 Mar, 2022 1 commit
- Size ONNX op (#1122) · d71a7b6a
  Charlie Lin authored Mar 08, 2022
```
* Implement size ONNX operator and tests
```
  d71a7b6a
07 Mar, 2022 1 commit
- Use `add_common_op` for handling types and broadcast in Clip Onnx parsing (#1121) · a0ae2f79
  Umang Yadav authored Mar 07, 2022
```
add_common_op for parse_clip
Should fix #1119
```
  a0ae2f79
04 Mar, 2022 2 commits

EyeLike Operator (#1087) · 8f184d4a
Charlie Lin authored Mar 04, 2022
```
Adds EyeLike ONNX parser and unit tests.
```
8f184d4a

Mode as enum for pooling and roi_align (#1091) · a2e90b5d

bpickrel authored Mar 04, 2022

Changed the pooling values for two structures from strings to specialized enum classes. Many test and operator parsing changes to support this. Introduces one new source file, op_enums.cpp.

a2e90b5d

03 Mar, 2022 3 commits
- Boost the max number of workgroups for pointwise ops (#1113) · d9d17a11
  Paul Fultz II authored Mar 03, 2022
```
Boost the max number of workgroups for pointwise ops by matching what we are doing in launch.hpp
```
  d9d17a11
- Use fp32 compute_type when calling rocBLAS API (#1085) · 36b01ba5
  kahmed10 authored Mar 03, 2022
```
better performance doing it this way
```
  36b01ba5
- Add ScatterND operator (#1074) · 832f28c6
  turneram authored Mar 02, 2022
```
Add onnx parser and ref and gpu implementations of ONNX op ScatterND
```
  832f28c6
02 Mar, 2022 3 commits
- isnan operator (#1100) · bfedcd45
  Charlie Lin authored Mar 02, 2022
```
Implements the IsNaN operator, ref, gpu, and onnx parser.
```
  bfedcd45
- Clang format ver10 (#1106) · 9852aaef
  bpickrel authored Mar 02, 2022
```
Update the base version of clang-format from 5.0 to 10.0
```
  9852aaef
- Free space on github jobs (#1107) · af0148ce
  Paul Fultz II authored Mar 01, 2022
```
* Free space on github jobs

* Delete more stuff

* Use sudo
```
  af0148ce
28 Feb, 2022 1 commit
- Bump version to 2.2 (#1105) · c1b56607
  Chris Austen authored Feb 28, 2022
```
Release branch created for ROCm 5.1 so moving develop branch to 2.2
```
  c1b56607
25 Feb, 2022 3 commits
- Add with_type to shape class (#1102) · 85b0563c
  Paul Fultz II authored Feb 25, 2022
```
Add with_type to shape class
```
  85b0563c
- Add reverse lookup of c++ class to c class (#1099) · 40c087bd
  Paul Fultz II authored Feb 25, 2022
```
Needed for custom_op so we can generically convert the C type back to the C++ type in the function pointer.
```
  40c087bd
- Add get_queue to context to get the current stream (#1097) · e5242676
  Paul Fultz II authored Feb 24, 2022
```
wrapped in a any_ptr class so the type can be checked at runtime for a mismatch.
```
  e5242676