"vscode:/vscode.git/clone" did not exist on "769cddcb2cc45ec78cb27148520e10bc8d7307d9"
  1. 09 Jan, 2024 1 commit
  2. 20 Nov, 2023 1 commit
    • Hongxin Liu's avatar
      [npu] add npu support for gemini and zero (#5067) · e5ce4c8e
      Hongxin Liu authored
      * [npu] setup device utils (#5047)
      
      * [npu] add npu device support
      
      * [npu] support low level zero
      
      * [test] update npu zero plugin test
      
      * [hotfix] fix import
      
      * [test] recover tests
      
      * [npu] gemini support npu (#5052)
      
      * [npu] refactor device utils
      
      * [gemini] support npu
      
      * [example] llama2+gemini support npu
      
      * [kernel] add arm cpu adam kernel (#5065)
      
      * [kernel] add arm cpu adam
      
      * [optim] update adam optimizer
      
      * [kernel] arm cpu adam remove bf16 support
      e5ce4c8e
  3. 20 Oct, 2023 1 commit
  4. 18 Oct, 2023 1 commit
  5. 16 Oct, 2023 1 commit
  6. 12 Oct, 2023 1 commit
    • Hongxin Liu's avatar
      [gemini] support amp o3 for gemini (#4872) · df635641
      Hongxin Liu authored
      * [gemini] support no reuse fp16 chunk
      
      * [gemini] support no master weight for optim
      
      * [gemini] support no master weight for gemini ddp
      
      * [test] update gemini tests
      
      * [test] update gemini tests
      
      * [plugin] update gemini plugin
      
      * [test] fix gemini checkpointio test
      
      * [test] fix gemini checkpoint io
      df635641
  7. 19 Sep, 2023 1 commit
  8. 18 Sep, 2023 1 commit
    • Hongxin Liu's avatar
      [legacy] clean up legacy code (#4743) · b5f9e37c
      Hongxin Liu authored
      * [legacy] remove outdated codes of pipeline (#4692)
      
      * [legacy] remove cli of benchmark and update optim (#4690)
      
      * [legacy] remove cli of benchmark and update optim
      
      * [doc] fix cli doc test
      
      * [legacy] fix engine clip grad norm
      
      * [legacy] remove outdated colo tensor (#4694)
      
      * [legacy] remove outdated colo tensor
      
      * [test] fix test import
      
      * [legacy] move outdated zero to legacy (#4696)
      
      * [legacy] clean up utils (#4700)
      
      * [legacy] clean up utils
      
      * [example] update examples
      
      * [legacy] clean up amp
      
      * [legacy] fix amp module
      
      * [legacy] clean up gpc (#4742)
      
      * [legacy] clean up context
      
      * [legacy] clean core, constants and global vars
      
      * [legacy] refactor initialize
      
      * [example] fix examples ci
      
      * [example] fix examples ci
      
      * [legacy] fix tests
      
      * [example] fix gpt example
      
      * [example] fix examples ci
      
      * [devops] fix ci installation
      
      * [example] fix examples ci
      b5f9e37c
  9. 24 Aug, 2023 1 commit
    • Hongxin Liu's avatar
      [gemini] improve compatibility and add static placement policy (#4479) · 27061426
      Hongxin Liu authored
      * [gemini] remove distributed-related part from colotensor (#4379)
      
      * [gemini] remove process group dependency
      
      * [gemini] remove tp part from colo tensor
      
      * [gemini] patch inplace op
      
      * [gemini] fix param op hook and update tests
      
      * [test] remove useless tests
      
      * [test] remove useless tests
      
      * [misc] fix requirements
      
      * [test] fix model zoo
      
      * [test] fix model zoo
      
      * [test] fix model zoo
      
      * [test] fix model zoo
      
      * [test] fix model zoo
      
      * [misc] update requirements
      
      * [gemini] refactor gemini optimizer and gemini ddp (#4398)
      
      * [gemini] update optimizer interface
      
      * [gemini] renaming gemini optimizer
      
      * [gemini] refactor gemini ddp class
      
      * [example] update gemini related example
      
      * [example] update gemini related example
      
      * [plugin] fix gemini plugin args
      
      * [test] update gemini ckpt tests
      
      * [gemini] fix checkpoint io
      
      * [example] fix opt example requirements
      
      * [example] fix opt example
      
      * [example] fix opt example
      
      * [example] fix opt example
      
      * [gemini] add static placement policy (#4443)
      
      * [gemini] add static placement policy
      
      * [gemini] fix param offload
      
      * [test] update gemini tests
      
      * [plugin] update gemini plugin
      
      * [plugin] update gemini plugin docstr
      
      * [misc] fix flash attn requirement
      
      * [test] fix gemini checkpoint io test
      
      * [example] update resnet example result (#4457)
      
      * [example] update bert example result (#4458)
      
      * [doc] update gemini doc (#4468)
      
      * [example] update gemini related examples (#4473)
      
      * [example] update gpt example
      
      * [example] update dreambooth example
      
      * [example] update vit
      
      * [example] update opt
      
      * [example] update palm
      
      * [example] update vit and opt benchmark
      
      * [hotfix] fix bert in model zoo (#4480)
      
      * [hotfix] fix bert in model zoo
      
      * [test] remove chatglm gemini test
      
      * [test] remove sam gemini test
      
      * [test] remove vit gemini test
      
      * [hotfix] fix opt tutorial example (#4497)
      
      * [hotfix] fix opt tutorial example
      
      * [hotfix] fix opt tutorial example
      27061426
  10. 25 Jun, 2023 1 commit
  11. 05 Jun, 2023 1 commit
    • Hongxin Liu's avatar
      [bf16] add bf16 support (#3882) · ae02d4e4
      Hongxin Liu authored
      * [bf16] add bf16 support for fused adam (#3844)
      
      * [bf16] fused adam kernel support bf16
      
      * [test] update fused adam kernel test
      
      * [test] update fused adam test
      
      * [bf16] cpu adam and hybrid adam optimizers support bf16 (#3860)
      
      * [bf16] implement mixed precision mixin and add bf16 support for low level zero (#3869)
      
      * [bf16] add mixed precision mixin
      
      * [bf16] low level zero optim support bf16
      
      * [text] update low level zero test
      
      * [text] fix low level zero grad acc test
      
      * [bf16] add bf16 support for gemini (#3872)
      
      * [bf16] gemini support bf16
      
      * [test] update gemini bf16 test
      
      * [doc] update gemini docstring
      
      * [bf16] add bf16 support for plugins (#3877)
      
      * [bf16] add bf16 support for legacy zero (#3879)
      
      * [zero] init context support bf16
      
      * [zero] legacy zero support bf16
      
      * [test] add zero bf16 test
      
      * [doc] add bf16 related docstring for legacy zero
      ae02d4e4
  12. 06 Apr, 2023 2 commits
  13. 04 Apr, 2023 1 commit
    • ver217's avatar
      [zero] reorganize zero/gemini folder structure (#3424) · 26b7aac0
      ver217 authored
      * [zero] refactor low-level zero folder structure
      
      * [zero] fix legacy zero import path
      
      * [zero] fix legacy zero import path
      
      * [zero] remove useless import
      
      * [zero] refactor gemini folder structure
      
      * [zero] refactor gemini folder structure
      
      * [zero] refactor legacy zero import path
      
      * [zero] refactor gemini folder structure
      
      * [zero] refactor gemini folder structure
      
      * [zero] refactor gemini folder structure
      
      * [zero] refactor legacy zero import path
      
      * [zero] fix test import path
      
      * [zero] fix test
      
      * [zero] fix circular import
      
      * [zero] update import
      26b7aac0
  14. 28 Jan, 2023 1 commit
  15. 09 Jan, 2023 1 commit
  16. 26 Dec, 2022 1 commit
  17. 12 Dec, 2022 1 commit
  18. 09 Dec, 2022 1 commit
  19. 05 Dec, 2022 1 commit
  20. 30 Nov, 2022 5 commits
  21. 29 Nov, 2022 1 commit
  22. 24 Nov, 2022 1 commit
  23. 16 Nov, 2022 1 commit
  24. 02 Nov, 2022 1 commit
  25. 18 Oct, 2022 1 commit
  26. 14 Oct, 2022 1 commit
    • HELSON's avatar
      [zero] add constant placement policy (#1705) · 1468e4bc
      HELSON authored
      * fixes memory leak when paramter is in fp16 in ZeroDDP init.
      * bans chunk releasement in CUDA. Only when a chunk is about to offload, it is allowed to release.
      * adds a constant placement policy. With it, users can allocate a reserved caching memory space for parameters.
      1468e4bc
  27. 09 Oct, 2022 1 commit
  28. 26 Sep, 2022 1 commit
  29. 24 Sep, 2022 1 commit