"...git@developer.sourcefind.cn:OpenDAS/TransformerEngine.git" did not exist on "262c184eb8331dcb50037477881900e46bd5c5f2"
Fix QKV dtype in the bwd of FP8+CP (#1134)
* fix qkv_dtype of FP8+CP Signed-off-by:Xiaowei Ren <xren@nvidia.com> * config cp correction dtype of FP8+CP Signed-off-by:
Xiaowei Ren <xren@nvidia.com> * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci * code style change Signed-off-by:
Xiaowei Ren <xren@nvidia.com> * always do FP8 CP correction in FP32 Signed-off-by:
Xiaowei Ren <xren@nvidia.com> --------- Signed-off-by:
Xiaowei Ren <xren@nvidia.com> Co-authored-by:
pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com> Co-authored-by:
Charlene Yang <8636796+cyanguwa@users.noreply.github.com>
Showing
Please register or sign in to comment