- 23 Jan, 2025 1 commit
-
-
Gregory Shtrasberg authored
Signed-off-by:
Gregory Shtrasberg <Gregory.Shtrasberg@amd.com> Co-authored-by:
Micah Williamson <micah.williamson@amd.com>
-
- 20 Jan, 2025 1 commit
-
-
wangxiyuan authored
Signed-off-by:wangxiyuan <wangxiyuan1007@gmail.com>
-
- 06 Jan, 2025 1 commit
-
-
Chen Zhang authored
Signed-off-by:Chen Zhang <zhangch99@outlook.com>
-
- 02 Dec, 2024 1 commit
-
-
youkaichao authored
Signed-off-by:youkaichao <youkaichao@gmail.com>
-
- 22 Nov, 2024 1 commit
-
-
youkaichao authored
Signed-off-by:youkaichao <youkaichao@gmail.com>
-
- 20 Nov, 2024 1 commit
-
-
Woosuk Kwon authored
Signed-off-by:Woosuk Kwon <woosuk.kwon@berkeley.edu>
-
- 21 Oct, 2024 1 commit
-
-
Thomas Parnell authored
Signed-off-by:Thomas Parnell <tpa@zurich.ibm.com>
-
- 14 Oct, 2024 1 commit
-
-
Woosuk Kwon authored
-
- 27 Sep, 2024 2 commits
-
-
youkaichao authored
-
Brittany authored
-
- 30 Aug, 2024 2 commits
-
-
Woosuk Kwon authored
-
Richard Liu authored
-
- 20 Aug, 2024 1 commit
-
-
Antoni Baum authored
-
- 01 Aug, 2024 1 commit
-
-
Woosuk Kwon authored
-
- 27 Jul, 2024 2 commits
-
-
Woosuk Kwon authored
-
Woosuk Kwon authored
-
- 16 Jul, 2024 1 commit
-
-
Michael Goin authored
-
- 12 Jul, 2024 1 commit
-
-
Woosuk Kwon authored
-
- 08 Jul, 2024 1 commit
-
-
afeldman-nm authored
[Kernel] Correctly invoke prefill & decode kernels for cross-attention (towards eventual encoder/decoder model support) (#4888) Co-authored-by:Woosuk Kwon <woosuk.kwon@berkeley.edu>
-
- 28 Jun, 2024 1 commit
-
-
Woosuk Kwon authored
-
- 26 Jun, 2024 3 commits
-
-
Woosuk Kwon authored
-
Woosuk Kwon authored
-
Stephanie Wang authored
Signed-off-by:
Stephanie Wang <swang@cs.berkeley.edu> Signed-off-by:
Stephanie <swang@anyscale.com> Co-authored-by:
Stephanie <swang@anyscale.com>
-
- 14 Jun, 2024 1 commit
-
-
Woosuk Kwon authored
-
- 13 Jun, 2024 1 commit
-
-
youkaichao authored
[Core][Distributed] add coordinator to reduce code duplication in tp and pp (#5293)
-
- 12 Jun, 2024 1 commit
-
-
Woosuk Kwon authored
-