This website requires JavaScript.
Explore
Help
Register
Sign In
wassname
/
vllm
Watch
1
Star
0
Fork
0
mirror of
https://github.com/wassname/vllm.git
synced
2026-08-15 12:54:49 +08:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
8678a69ab51956031e3bb70bdf1a781a8652e67d
vllm
/
vllm
/
model_executor
/
model_loader
T
History
Dipika Sikka
and
ElizaWszola
8678a69ab5
[Kernel] Expand MoE weight loading + Add Fused Marlin MoE Kernel (
#7527
)
...
Co-authored-by: ElizaWszola <
eliza@neuralmagic.com
>
2024-08-21 16:17:10 -07:00
..
__init__.py
[VLM] Refactor
MultiModalConfig
initialization and profiling (
#7530
)
2024-08-17 13:30:55 -07:00
loader.py
[VLM] Refactor
MultiModalConfig
initialization and profiling (
#7530
)
2024-08-17 13:30:55 -07:00
neuron.py
[Typing] Mypy typing part 2 (
#4043
)
2024-04-17 17:28:43 -07:00
openvino.py
[Hardware][Intel] OpenVINO vLLM backend (
#5379
)
2024-06-28 13:50:16 +00:00
tensorizer.py
[Frontend] Add FlexibleArgumentParser to support both underscore and dash in names (
#5718
)
2024-06-20 17:00:13 -06:00
utils.py
[Kernel] Expand MoE weight loading + Add Fused Marlin MoE Kernel (
#7527
)
2024-08-21 16:17:10 -07:00
weight_utils.py
[Model] Add AWQ quantization support for InternVL2 model (
#7187
)
2024-08-20 23:18:57 -07:00