"...resnet50_tensorflow.git" did not exist on "4a0551a85b3b6d99c1470df5c8dee1d9c2ffe248"
Added bwd data v3r1 v4r1, tweaking v1 (#10)
* Added bwd data v3r1: breaking down compute into a series of load balanced GEMM, and launch in a single kernel * Added bwd data v4r1: like v3r1, but launch GEMMs in multiple kernels * Tweaked v1r1 and v1r2 (atomic) on AMD GPU
Showing
Please register or sign in to comment