- 10 Jun, 2021 1 commit
-
-
Chao Liu authored
* experimenting magic number division * overhauling fwd-v4r4 to clearly reflect transformation graph * added fwd-v4r5 * bug fix for make_dynamic_naive_tensor_descriptor_aligned_v2 * bug fix and added sanity-check in transform_dynamic_tensor_descriptor * added conv_driver_v2
-
- 24 Jun, 2020 1 commit
-
-
Chao Liu authored
* tuning para, * testing on v100 * add fp16 * remove deprecated tensor descriptor * sync with miopen * update build script Co-authored-by:Jing Zhang <jizhan@amd.com>
-
- 03 Dec, 2019 1 commit
-
-
Chao Liu authored
* enabled atomic add in tensor copy * added gridwise GEMM * added backward data conv using GEMM + atomic * added backward data conv using GEMM, no atomic
-
- 13 Jun, 2019 1 commit
-
-
Chao Liu authored
-
- 15 May, 2019 1 commit
-
-
Chao Liu authored
-
- 04 Nov, 2018 1 commit
-
-
Chao Liu authored
-
- 22 Oct, 2018 1 commit
-
-
Chao Liu authored
-
- 14 Oct, 2018 1 commit
-
-
Chao Liu authored
-
- 09 Oct, 2018 1 commit
-
-
Chao Liu authored
-