"vscode:/vscode.git/clone" did not exist on "824a77d04d90662eeb3864d3f36e9f2458d4b9f6"
  1. 22 Apr, 2021 2 commits
  2. 19 Apr, 2021 1 commit
    • Min Xu's avatar
      FSDP: fixing training with freezing weights (#614) · 24da3b11
      Min Xu authored
      
      
      * FSDP: fixing training with freezing weights
      
      - an assert is changed to catch this case correctly
      - unit test added (based on Quentin's test code) for this case and
        compare DDP and FSDP
      
      fixes: #610
      
      * added test file to list 1
      
      * Use better and simpler code as suggested by Myle
      
      * testing both methods of freezing as well
      Co-authored-by: default avatarMin Xu <min.xu@acm.org>
      24da3b11
  3. 13 Apr, 2021 3 commits
  4. 08 Apr, 2021 1 commit
  5. 07 Apr, 2021 2 commits
  6. 06 Apr, 2021 1 commit
  7. 04 Apr, 2021 2 commits
  8. 02 Apr, 2021 1 commit
  9. 01 Apr, 2021 1 commit
  10. 31 Mar, 2021 2 commits
    • Min Xu's avatar
      [fix] FSDP: disable single rank process group for auto_wrap_bn and fixed mixed... · a0458b98
      Min Xu authored
      [fix] FSDP: disable single rank process group for auto_wrap_bn and fixed mixed precision regnet test (#556)
      
      * [fix] disable single rank process group for auto_wrap_bn
      
      - beefed up unit test with regnet-like model
      - found that single-rank process group is causing problem
      - disabled it to enable convergence tests on the vissl side
      - use `raise e from None` to get a better assertion output
        in testing.py.
      
      * [test] fix regnet test for ddp+mixed_precision
      
      - need AMP context in FSDP
      - workaround different between ddp & fsdp when bias=True
      - fixed a bug in input data generation that caused different ranks have
        the same data with wrong iteration count.
      - added TODO for need a better loss and grad_scaler and reduced
        iters so there is no nan.
      - added a (disabled) debugging code
      
      * lint
      
      * lint
      
      * add scaler
      
      * lint
      
      * scaler
      
      * add a real loss
      
      * seeding in the ranks
      
      * blance tests
      
      * run AMP DDP==FSDP test only on cuda version 11 and up
      
      * add relu inplace and comment
      
      * make wrap_bn covers more cases in full precision mode
      a0458b98
    • msbaines's avatar
      acb9ef00
  11. 30 Mar, 2021 1 commit
  12. 26 Mar, 2021 1 commit
  13. 25 Mar, 2021 2 commits
  14. 22 Mar, 2021 1 commit
  15. 20 Mar, 2021 1 commit
  16. 19 Mar, 2021 2 commits
  17. 18 Mar, 2021 4 commits
  18. 17 Mar, 2021 1 commit
  19. 12 Mar, 2021 2 commits
  20. 11 Mar, 2021 1 commit
  21. 09 Mar, 2021 2 commits
  22. 08 Mar, 2021 1 commit
    • Min Xu's avatar
      [fix]: handle inputs with containers in mixed precision (#486) · 2e9a14e7
      Min Xu authored
      * [fix]: handle inputs with containers
      
      - this is an issue surfaces by vissl as well
      - fix seems to be super simple
      - also cleaned up two tests with respect to multiple such tests
        running back to back (they don't do that presently)
      
      * cleanup
      
      * fix
      
      * lint
      2e9a14e7
  23. 06 Mar, 2021 1 commit
  24. 05 Mar, 2021 2 commits
  25. 04 Mar, 2021 2 commits