- 21 Jun, 2023 3 commits
- 20 Jun, 2023 4 commits
-
-
lvhan028 authored
* check-in dockerfile * check-in dockerfile
-
Li Zhang authored
* add ft code * gitignore * fix lint * revert fmha
-
lvhan028 authored
* add logo * update readme
-
lvhan028 authored
* add scripts for deploying llama family models via fastertransformer * fix * fix * set symlinks True when copying triton models templates * pack model repository for triton inference server * add exception * fix * update config.pbtxt and launching scripts
-
- 18 Jun, 2023 4 commits