This website requires JavaScript.
Explore
Help
Register
Sign In
wassname
/
vllm
Watch
1
Star
0
Fork
0
mirror of
https://github.com/wassname/vllm.git
synced
2026-07-29 11:29:22 +08:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
59c9b6ebeba79b2d744eec86734a7e13b03dcab7
vllm
/
vllm
/
core
T
History
Sungjae Lee
and
GitHub
886936837c
[Performance][Core] Optimize the performance of evictor v1 and v2 by applying a priority queue and lazy deletion (
#7209
)
2024-12-14 11:38:10 -08:00
..
block
[Core] support LoRA and prompt adapter in content-based hashing for Block Manager v2 prefix caching (
#8240
)
2024-12-13 07:51:25 -08:00
__init__.py
Change the name to vLLM (
#150
)
2023-06-17 03:07:40 -07:00
block_manager.py
[Core] support LoRA and prompt adapter in content-based hashing for Block Manager v2 prefix caching (
#8240
)
2024-12-13 07:51:25 -08:00
evictor.py
[Performance][Core] Optimize the performance of evictor v1 and v2 by applying a priority queue and lazy deletion (
#7209
)
2024-12-14 11:38:10 -08:00
interfaces.py
Prefix Cache Aware Scheduling [1/n] (
#10128
)
2024-11-22 21:15:55 -08:00
placeholder_block_space_manager.py
[Doc] Update docs to refer to pooling models (
#11093
)
2024-12-11 13:36:27 +00:00
scheduler.py
[Misc] Split up pooling tasks (
#10820
)
2024-12-11 01:28:00 -08:00