- 22 Jul, 2025 1 commit
-
-
gushiqiao authored
-
- 21 Jul, 2025 1 commit
-
-
gushiqiao authored
-
- 20 Jul, 2025 3 commits
-
-
helloyongyang authored
-
helloyongyang authored
-
helloyongyang authored
-
- 19 Jul, 2025 1 commit
-
-
gushiqiao authored
-
- 18 Jul, 2025 1 commit
-
-
gushiqiao authored
-
- 17 Jul, 2025 2 commits
-
-
wangshankun authored
-
gaclove authored
-
- 16 Jul, 2025 2 commits
-
-
Zhuguanyu Wu authored
* use CM scheduler for distill_models as default * update configs
-
helloyongyang authored
-
- 15 Jul, 2025 1 commit
-
-
wangshankun authored
-
- 14 Jul, 2025 2 commits
-
-
helloyongyang authored
-
helloyongyang authored
-
- 12 Jul, 2025 1 commit
-
-
gushiqiao authored
-
- 11 Jul, 2025 5 commits
-
-
GoatWu authored
-
helloyongyang authored
-
helloyongyang authored
-
helloyongyang authored
-
helloyongyang authored
-
- 10 Jul, 2025 2 commits
-
-
helloyongyang authored
-
Zhuguanyu Wu authored
* support dynamic cfg for cfg_distill * reformat files
-
- 09 Jul, 2025 1 commit
-
-
gushiqiao authored
-
- 04 Jul, 2025 2 commits
-
-
Yang Rongjin authored
* add readme * modify readme
-
Zhuguanyu Wu authored
* update lora keys * update lora extractor/merger tools * rename lora config and script files * bug fixed for lora tools
-
- 03 Jul, 2025 4 commits
-
-
wangshankun authored
-
wangshankun authored
-
wangshankun authored
-
wangshankun authored
-
- 02 Jul, 2025 1 commit
-
-
gushiqiao authored
Enable 720p model inference on low-spec GPUs/CPUs and accelerate T5/CLIP quantized models with vLLM operators
-
- 01 Jul, 2025 3 commits
-
-
GoatWu authored
-
gushiqiao authored
-
gushiqiao authored
Co-authored-by:gushiqiao <gushiqiao@sensetime.com>
-
- 30 Jun, 2025 2 commits
-
-
GoatWu authored
-
helloyongyang authored
-
- 16 Jun, 2025 1 commit
-
-
gushiqiao authored
-
- 12 Jun, 2025 1 commit
-
-
Zhuguanyu Wu authored
* add step & cfg distillation wan model * bug fixed
-
- 10 Jun, 2025 1 commit
-
-
gushiqiao authored
-
- 09 Jun, 2025 1 commit
-
-
gushiqiao authored
* reconstruct quantization and fix memory leak bug. * Support lazy load inference. * reconstruct quantization * Fix hunyuan bugs * deleted tmp file --------- Co-authored-by:
root <root@pt-c0b333b3a1834e81a0d4d5f412c6ffa1-worker-0.pt-c0b333b3a1834e81a0d4d5f412c6ffa1.ns-devsft-3460edd0.svc.cluster.local> Co-authored-by:
gushiqiao <gushqiaio@sensetime.com> Co-authored-by:
gushiqiao <gushiqiao@sensetime.com>
-
- 30 May, 2025 1 commit
-
-
Zhuguanyu Wu authored
* split dit server from default runner * split dit server from default runner * update loading functions * simplify loader functions and runner functions * simplify code && split dit service * simplify code && split dit service * support split server for cogvideox * clear code.
-