"include/ck/utility/reduction_operator.hpp" did not exist on "dfb80c4e39ec7b304c3ebc88bab2a204bc4906b9"
DL GEMM fp32/fp16/int8 (#41)
* add threadwise copy the copy a tensor in one copy, added kpack to DL GEMM * add kpack into fwd v4r5 nchw fp32
Showing
Please register or sign in to comment