1. 19 Mar, 2026 1 commit
  2. 18 Mar, 2026 1 commit
    • laibao's avatar
      feat(moe): 增加 LightOP moe_sum+mul+add 融合并打通参数透传 · 0639678c
      laibao authored
        新增环境变量 VLLM_USE_LIGHTOP_MOE_SUM_MUL_ADD 用于控制
        fused sum+mul+add 开关。
        在 DeepseekV2MoE 中增加 fused 路径,预计算 shared_output,并下传 iqis 与 routed_scaling_factor。
        扩展 FusedMoE/SharedFusedMoE 及相关 custom op 接口,统一透传 i_q/i_s/shared_output/routed_scaling_factor。
        同步适配 Triton、Marlin W16A16、SlimQuant W4A8、CompressedTensors W8A8 等实现,支持在内核侧完成 sum+mul+add。
      0639678c
  3. 12 Mar, 2026 2 commits
  4. 07 Mar, 2026 1 commit
  5. 06 Mar, 2026 1 commit
  6. 05 Mar, 2026 1 commit
  7. 03 Mar, 2026 1 commit
  8. 02 Mar, 2026 1 commit
  9. 06 Feb, 2026 2 commits
  10. 03 Feb, 2026 1 commit
  11. 26 Jan, 2026 1 commit
  12. 21 Jan, 2026 1 commit
  13. 20 Jan, 2026 1 commit
  14. 16 Jan, 2026 1 commit
  15. 09 Jan, 2026 1 commit
  16. 07 Jan, 2026 4 commits
  17. 06 Jan, 2026 2 commits
  18. 24 Dec, 2025 2 commits
  19. 19 Dec, 2025 2 commits
  20. 18 Dec, 2025 1 commit
  21. 17 Dec, 2025 2 commits
  22. 12 Dec, 2025 1 commit
  23. 11 Dec, 2025 1 commit
  24. 08 Dec, 2025 1 commit
  25. 02 Dec, 2025 3 commits
  26. 30 Nov, 2025 1 commit
  27. 26 Nov, 2025 1 commit
  28. 24 Nov, 2025 1 commit
  29. 20 Nov, 2025 1 commit