[Kernel][ROCM] Upstream prefix prefill speed up for vLLM V1 (#13305)
Signed-off-by:Sage Moore <sage@neuralmagic.com> Signed-off-by:
root <root@banff-cyxtera-s73-5.ctr.dcgpu> Signed-off-by:
Aleksandr Malyshev <maleksan@amd.com> Signed-off-by:
root <root@banff-cyxtera-s65-4.amd.com> Signed-off-by:
maleksan85 <maleksan@amd.com> Signed-off-by: <> Co-authored-by:
Sage Moore <sage@neuralmagic.com> Co-authored-by:
root <root@banff-cyxtera-s73-5.ctr.dcgpu> Co-authored-by:
Aleksandr Malyshev <maleksan@amd.com> Co-authored-by:
qli88 <qiang.li2@amd.com> Co-authored-by:
root <root@banff-cyxtera-s65-4.amd.com>
Showing
Please register or sign in to comment