Merge branch 'v0.11.0-dev_tc_opt' into 'v0.11.0-dev'
perf(fused-moe): 接入 W16A16 Marlin MoE 并缓存 pack 权重 See merge request dcutoolkit/deeplearing/vllm!347
Showing
Please register or sign in to comment
perf(fused-moe): 接入 W16A16 Marlin MoE 并缓存 pack 权重 See merge request dcutoolkit/deeplearing/vllm!347