- 02 Jul, 2019 1 commit
-
-
mcarilli authored
-
- 28 Jun, 2019 2 commits
-
-
Thor Johnsen authored
Add support for fp16 update term (new UPD_T typename in template)
-
Thor Johnsen authored
-
- 27 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 25 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 24 Jun, 2019 4 commits
-
-
-
Michael Carilli authored
-
mcarilli authored
-
Michael Carilli authored
-
- 21 Jun, 2019 2 commits
-
-
Michael Carilli authored
-
Michael Carilli authored
-
- 20 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 19 Jun, 2019 2 commits
-
-
- 18 Jun, 2019 2 commits
-
-
mcarilli authored
-
Michael Carilli authored
-
- 17 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 14 Jun, 2019 4 commits
-
-
-
Michael Carilli authored
-
Thor Johnsen authored
-
Michael Carilli authored
-
- 13 Jun, 2019 2 commits
-
-
Michael Carilli authored
-
Thor Johnsen authored
-
- 11 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 07 Jun, 2019 1 commit
-
-
- 06 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 04 Jun, 2019 1 commit
-
-
Michael Carilli authored
-
- 31 May, 2019 2 commits
-
-
Thor Johnsen authored
* First draft, for discussion * Fix mistakes in LAMB equations * Add loop over chunk * Bug fix * Bug fix * Bug fix * Undo bug fix * Bug fix * Add multi tensor LAMB optimizer to setup.py * Rename step_size to learning_rate * Fix compilation errors
-
mcarilli authored
* Existing tests passing, still need to add per-tensor tests * Test is passing, still need to measure performance * ILP for l2norm functor
-
- 28 May, 2019 1 commit
-
-
- 24 May, 2019 1 commit
-
-
ptrblck authored
-
- 23 May, 2019 1 commit
-
-
Michael Carilli authored
-
- 22 May, 2019 3 commits
-
-
mcarilli authored
-
Michael Carilli authored
-
ptrblck authored
-
- 21 May, 2019 1 commit
-
-
blisc authored
* update larc Signed-off-by:
Jason <jasoli@nvidia.com> * scale_loss fix Signed-off-by:
Jason <jasoli@nvidia.com> * typo Signed-off-by:
Jason <jasoli@nvidia.com> * revert LARC
-
- 17 May, 2019 2 commits
-
-
jjsjann123 authored
update input size check to fix github issue #262 update SyncBatchNorm count check so that size 1 input with cross GPU synchronization runs fine.
-
jjsjann123 authored
resolves issue #254 Added input casting for pure python implementation, this supports mismatched input and layer dtype.
-
- 16 May, 2019 1 commit
-
-
mcarilli authored
* Support add_param_group * syntax * Test added and passing
-
- 15 May, 2019 1 commit
-
-
Michael Glass authored
-