Logo
Explore Help
Register Sign In
wassname/vllm
Watch 1
Star 0
Fork 0
mirror of https://github.com/wassname/vllm.git synced 2026-07-26 13:37:27 +08:00
Code Issues Packages Projects Releases Wiki Activity
Files
f1e15da6fe20ff17d5b8c28f37487cee38f08b83
vllm/docs/source
T
History
Cyrus LeungandGitHub ae96ef8fbd [VLM] Calculate maximum number of multi-modal tokens by model (#6121)
2024-07-04 16:37:23 -07:00
..
_templates/sections
[misc][doc] try to add warning for latest html (#5979)
2024-07-04 09:57:09 -07:00
assets
[Doc] add visualization for multi-stage dockerfile (#4456)
2024-04-30 17:41:59 +00:00
automatic_prefix_caching
[Doc] Add an automatic prefix caching section in vllm documentation (#5324)
2024-06-11 10:24:59 -07:00
community
[Docs] Add ZhenFund as a Sponsor (#5548)
2024-06-14 11:17:21 -07:00
dev
[VLM] Calculate maximum number of multi-modal tokens by model (#6121)
2024-07-04 16:37:23 -07:00
getting_started
[doc][misc] bump up py version in installation doc (#6119)
2024-07-03 15:52:04 -07:00
models
[VLM] Calculate maximum number of multi-modal tokens by model (#6121)
2024-07-04 16:37:23 -07:00
quantization
[Kernel] Expand FP8 support to Ampere GPUs using FP8 Marlin (#5975)
2024-07-03 17:38:00 +00:00
serving
[Bugfix][Doc] Fix Doc Formatting (#6048)
2024-07-01 15:09:11 -07:00
conf.py
[misc][doc] try to add warning for latest html (#5979)
2024-07-04 09:57:09 -07:00
generate_examples.py
Add example scripts to documentation (#4225)
2024-04-22 16:36:54 +00:00
index.rst
add FAQ doc under 'serving' (#5946)
2024-07-01 14:11:36 -07:00
Powered by Gitea Version: 1.27.0 Page: 53ms Template: 3ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API