Logo
Explore Help
Sign In
youngkingdom/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions Packages Projects Releases Wiki Activity
Files
b12518d3cf4326dfcd10a09780913b86c19fcf1a
vllm/tests/spec_decode/e2e
History
zifeitong 78687504f7 [Bugfix] AsyncLLMEngine hangs with asyncio.run (#5654)
2024-06-19 13:57:12 -07:00
..
__init__.py
[Speculative decoding 7/9] Speculative decoding end-to-end correctness tests. (#3951)
2024-04-23 08:02:36 +00:00
conftest.py
[Bugfix] AsyncLLMEngine hangs with asyncio.run (#5654)
2024-06-19 13:57:12 -07:00
test_compatibility.py
[Speculative decoding][Re-take] Enable TP>1 speculative decoding (#4840)
2024-05-16 00:53:51 -07:00
test_integration_dist.py
[Speculative decoding][Re-take] Enable TP>1 speculative decoding (#4840)
2024-05-16 00:53:51 -07:00
test_integration.py
[Speculative decoding][Re-take] Enable TP>1 speculative decoding (#4840)
2024-05-16 00:53:51 -07:00
test_logprobs.py
[Speculative decoding] Support target-model logprobs (#4378)
2024-05-03 15:52:01 -07:00
test_multistep_correctness.py
[Speculative decoding][Re-take] Enable TP>1 speculative decoding (#4840)
2024-05-16 00:53:51 -07:00
test_ngram_correctness.py
[Dynamic Spec Decoding] Minor fix for disabling speculative decoding (#5000)
2024-05-25 10:00:14 -07:00
Powered by Gitea Version: 1.24.2 Page: 106ms Template: 5ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API