Logo
Explore Help
Sign In
youngkingdom/vllm
1
0
Fork 0
You've already forked vllm
Code Issues Pull Requests Actions Packages Projects Releases Wiki Activity
Files
98aa16ff41353e3e6c8a3c2f4e933a888dbce1cb
vllm/tests/models
History
Chen Zhang 2b4fc9bd9b Support FlashAttention Backend for Hybrid SSM Models (#23299)
Signed-off-by: Chen Zhang <zhangch99@outlook.com>
2025-08-26 12:41:52 +00:00
..
fixtures
[Mistral-Small 3.1] Update docs and tests (#14977)
2025-03-18 03:29:42 -07:00
language
Support FlashAttention Backend for Hybrid SSM Models (#23299)
2025-08-26 12:41:52 +00:00
multimodal
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
quantization
[V0 Deprecation] Remove V0 FlashInfer attention backend (#22776)
2025-08-18 19:54:16 -07:00
__init__.py
[CI/Build] Move test_utils.py to tests/utils.py (#4425)
2024-05-13 23:50:09 +09:00
registry.py
[New Model]Donut model (#23229)
2025-08-24 12:52:24 +00:00
test_initialization.py
[Model] Add LFM2 architecture (#22845)
2025-08-21 09:35:07 +02:00
test_oot_registration.py
[CI/Build] Fix plugin tests (#21758)
2025-07-28 15:08:05 +00:00
test_registry.py
[Deprecation][2/N] Replace --task with --runner and --convert (#21470)
2025-07-27 19:42:40 -07:00
test_transformers.py
Enable headless models for pooling in the Transformers backend (#21767)
2025-08-01 10:31:29 -07:00
test_utils.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
test_vision.py
[Misc] Add SPDX-FileCopyrightText (#19100)
2025-06-03 11:20:17 -07:00
utils.py
[Model] Pooling models default to using chunked prefill & prefix caching if supported. (#20930)
2025-08-11 09:41:37 -07:00
Powered by Gitea Version: 1.24.2 Page: 439ms Template: 8ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API