"vscode:/vscode.git/clone" did not exist on "ccda545c1f81d39692c93c8af33f80a86c1932df"
- 25 Jun, 2023 4 commits
-
-
tpoisonooo authored
-
tpoisonooo authored
-
tpoisonooo authored
-
lvhan028 authored
* remove constraints on model name * remove duplicate model converter * add profile * get eos and bos from server * update stop_words * update sequence_length when the last generated token is eos_id * fix * fix * check-in models * valicate model_name * make stop_words as property * debug profiling * better stats * fix assistant reponse * update profile serving * update * update
-
- 24 Jun, 2023 1 commit
-
-
Li Zhang authored
* support attention bias * fix conflict
-
- 22 Jun, 2023 2 commits
-
-
lvhan028 authored
* remove constraints on model name * remove duplicate model converter
-
q.yao authored
* update arch * clang-format * remove comment --------- Co-authored-by:yaoqian <yaoqian@localhost.localdomain>
-
- 21 Jun, 2023 3 commits
- 20 Jun, 2023 4 commits
-
-
lvhan028 authored
* check-in dockerfile * check-in dockerfile
-
Li Zhang authored
* add ft code * gitignore * fix lint * revert fmha
-
lvhan028 authored
* add logo * update readme
-
lvhan028 authored
* add scripts for deploying llama family models via fastertransformer * fix * fix * set symlinks True when copying triton models templates * pack model repository for triton inference server * add exception * fix * update config.pbtxt and launching scripts
-
- 18 Jun, 2023 4 commits