Logo
Explore Help
Register Sign In
wassname/vllm
Watch 1
Star 0
Fork 0
mirror of https://github.com/wassname/vllm.git synced 2026-08-17 11:28:47 +08:00
Code Issues Packages Projects Releases Wiki Activity
Files
f49777ba62b4926d0f8c100ab06edb03c5c10098
vllm/tests/compile
T
History
Luka GovedičandVarun Sundar Rabindranath 30870b4f66 [torch.compile] Dynamic fp8 + rms_norm fusion (#10906)
Signed-off-by: luka <luka@neuralmagic.com>
Co-authored-by: Varun Sundar Rabindranath <varun@neuralmagic.com>
2024-12-13 03:19:23 +00:00
..
piecewise
[torch.compile] remove compilation_context and simplify code (#10838)
2024-12-03 06:19:02 +00:00
__init__.py
[torch.compile] register allreduce operations as custom ops (#8526)
2024-09-16 22:57:57 -07:00
backend.py
[torch.compile] Inductor code caching fix (#10273)
2024-11-20 21:44:57 -08:00
test_basic_correctness.py
[Misc] Split up pooling tasks (#10820)
2024-12-11 01:28:00 -08:00
test_full_graph.py
[2/N][torch.compile] make compilation cfg part of vllm cfg (#10383)
2024-11-16 18:02:14 -08:00
test_functionalization.py
[torch.compile] Dynamic fp8 + rms_norm fusion (#10906)
2024-12-13 03:19:23 +00:00
test_fusion.py
[torch.compile] Dynamic fp8 + rms_norm fusion (#10906)
2024-12-13 03:19:23 +00:00
test_pass_manager.py
[torch.compile] Inductor code caching fix (#10273)
2024-11-20 21:44:57 -08:00
test_wrapper.py
[2/N][torch.compile] make compilation cfg part of vllm cfg (#10383)
2024-11-16 18:02:14 -08:00
utils.py
[bugfix] fix full graph tests (#10581)
2024-11-22 10:02:14 -08:00
Powered by Gitea Version: 1.27.2 Page: 3845ms Template: 2ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API