Logo
Explore Help
Register Sign In
wassname/vllm
Watch 1
Star 0
Fork 0
mirror of https://github.com/wassname/vllm.git synced 2026-07-23 13:10:20 +08:00
Code Issues Packages Projects Releases Wiki Activity
Files
2adb4409e0359039135b5aa6501994da12aa5a26
vllm/tests/worker
T
History
wangshuai09andGitHub 3ddbe25502 [Hardware][CPU] using current_platform.is_cpu (#9536)
2024-10-22 00:50:43 -07:00
..
__init__.py
[Speculative decoding 2/9] Multi-step worker for draft model (#2424)
2024-01-21 16:31:47 -08:00
test_encoder_decoder_model_runner.py
[Hardware][CPU] using current_platform.is_cpu (#9536)
2024-10-22 00:50:43 -07:00
test_model_input.py
[Core] Add AttentionState abstraction (#7663)
2024-08-20 18:50:45 +00:00
test_model_runner.py
[Core] Factor out common code in SequenceData and Sequence (#8675)
2024-09-21 02:30:39 +00:00
test_profile.py
🐛 fix torch memory profiling (#9516)
2024-10-18 21:25:19 -04:00
test_swap.py
[Core] Pipeline Parallel Support (#4412)
2024-07-02 10:58:08 -07:00
Powered by Gitea Version: 1.27.0 Page: 68ms Template: 1ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API