Fix per_token_group_quant_8bit when hidden_dim // group_size is not divided by 4. (#8449)
Co-authored-by:
Zhang Kaihong <zhangkaihong.zkh@alibaba-inc.com>
Showing
Please register or sign in to comment
Co-authored-by:
Zhang Kaihong <zhangkaihong.zkh@alibaba-inc.com>