"megatron/data/gpt_dataset.py" did not exist on "6529003325ebd5c19e6aa62ef2df4fdf2676c8f4"
-
Chao Liu authored
* Added bwd data v3r1: breaking down compute into a series of load balanced GEMM, and launch in a single kernel * Added bwd data v4r1: like v3r1, but launch GEMMs in multiple kernels * Tweaked v1r1 and v1r2 (atomic) on AMD GPU
c5da0377