- 05 Nov, 2024 1 commit
-
-
Michael Goin authored
Signed-off-by:mgoin <michael@neuralmagic.com>
-
- 29 Oct, 2024 1 commit
-
-
wangshuai09 authored
Signed-off-by:wangshuai09 <391746016@qq.com>
-
- 28 Oct, 2024 1 commit
-
-
wangshuai09 authored
Signed-off-by:wangshuai09 <391746016@qq.com>
-
- 25 Sep, 2024 1 commit
-
-
bnellnm authored
-
- 19 Sep, 2024 1 commit
-
-
Charlie Fu authored
-
- 18 Sep, 2024 1 commit
-
-
Cyrus Leung authored
-
- 14 Sep, 2024 1 commit
-
-
Charlie Fu authored
-
- 11 Sep, 2024 1 commit
-
-
bnellnm authored
Co-authored-by:Sage Moore <sage@neuralmagic.com>
-
- 16 Aug, 2024 1 commit
-
-
jon-chuang authored
-
- 27 Jul, 2024 1 commit
-
-
Joe authored
-
- 16 Jul, 2024 1 commit
-
-
Michael Goin authored
-
- 15 Jun, 2024 1 commit
-
-
Cyrus Leung authored
-
- 31 May, 2024 1 commit
-
-
SnowDist authored
Co-authored-by:Zhuohan Li <zhuohan123@gmail.com>
-
- 13 May, 2024 1 commit
-
-
Cyrus Leung authored
Since #4335 was merged, I've noticed that the definition of ServerRunner in the tests is the same as in the test for OpenAI API. I have moved the class to the test utilities to avoid code duplication. (Although it only has been repeated twice so far, I will add another similar test suite in #4200 which would duplicate the code a third time) Also, I have moved the test utilities file (test_utils.py) to under the test directory (tests/utils.py), since none of its code is actually used in the main package. Note that I have added __init__.py to each test subpackage and updated the ray.init() call in the test utilities file in order to relative import tests/utils.py.
-
- 10 May, 2024 1 commit
-
-
Cody Yu authored
-
- 03 May, 2024 1 commit
-
-
SangBin Cho authored
-
- 11 Apr, 2024 1 commit
-
-
Kunshang Ji authored
-
- 03 Apr, 2024 1 commit
-
-
Adrian Abeyta authored
Co-authored-by:
Gregory Shtrasberg <Gregory.Shtrasberg@amd.com> Co-authored-by:
HaiShaw <hixiao@gmail.com> Co-authored-by:
AdrianAbeyta <Adrian.Abeyta@amd.com> Co-authored-by:
Matthew Wong <Matthew.Wong2@amd.com> Co-authored-by:
root <root@gt-pla-u18-08.pla.dcgpu> Co-authored-by:
mawong-amd <156021403+mawong-amd@users.noreply.github.com> Co-authored-by:
ttbachyinsda <ttbachyinsda@outlook.com> Co-authored-by:
guofangze <guofangze@kuaishou.com> Co-authored-by:
Michael Goin <mgoin64@gmail.com> Co-authored-by:
jacobthebanana <50071502+jacobthebanana@users.noreply.github.com> Co-authored-by:
Woosuk Kwon <woosuk.kwon@berkeley.edu>
-
- 25 Mar, 2024 1 commit
-
-
SangBin Cho authored
-
- 05 Feb, 2024 1 commit
-
-
Hongxia Yang authored
-
- 01 Feb, 2024 1 commit
-
-
Kunshang Ji authored
Co-authored-by:
Jiang Li <jiang1.li@intel.com> Co-authored-by:
Kunshang Ji <kunshang.ji@intel.com>
-
- 29 Jan, 2024 1 commit
-
-
zhaoyang-star authored
Co-authored-by:
zhaoyang <zhao.yang16@zte.com.cn> Co-authored-by:
Zhuohan Li <zhuohan123@gmail.com>
-
- 14 Jan, 2024 1 commit
-
-
Simon Mo authored
-
- 03 Jan, 2024 1 commit
-
-
Jee Li authored
-
- 10 Dec, 2023 1 commit
-
-
wbn authored
Co-authored-by:
wangguoya <wangguoya@baidu.com> Co-authored-by:
Yang Zhao <zhaoyangstar@foxmail.com>
-
- 24 Nov, 2023 1 commit
-
-
Yanming W authored
-
- 20 Nov, 2023 1 commit
-
-
Simon Mo authored
-
- 31 Oct, 2023 1 commit
-
-
Woosuk Kwon authored
-
- 16 Oct, 2023 1 commit
-
-
Woosuk Kwon authored
-
- 28 Sep, 2023 1 commit
-
-
Woosuk Kwon authored
-
- 27 Sep, 2023 1 commit
-
-
Antoni Baum authored
-
- 07 Sep, 2023 1 commit
-
-
Zhuohan Li authored
* [FIX] Fix Alibi implementation in PagedAttention kernel * Fix test_attention * Fix --------- Co-authored-by:
Woosuk Kwon <woosuk.kwon@berkeley.edu> Co-authored-by:
Oliver-ss <yuansongwx@outlook.com>
-
- 05 Sep, 2023 1 commit
-
-
Woosuk Kwon authored
-
- 30 Aug, 2023 1 commit
-
-
Aman Gupta Karmani authored
-
- 25 Jul, 2023 1 commit
-
-
Tao Peng authored
Signed-off-by:Tao Peng <jiankeng.pt@alibaba-inc.com>
-
- 18 Jul, 2023 1 commit
-
-
Song authored
Co-authored-by:oliveryuan <oliveryuan@basemind.com>
-
- 09 Jul, 2023 1 commit
-
-
Andre Slavescu authored
Co-authored-by:woWoosuk Kwon <woosuk.kwon@berkeley.edu>
-
- 03 Jul, 2023 2 commits
-
-
Woosuk Kwon authored
-
Zhuohan Li authored
-
- 17 Jun, 2023 1 commit
-
-
Woosuk Kwon authored
-