vllm-project projects
Search results
-
- #38 updated
1 minute agoJul 29, 2026 - #17 updated
18 minutes agoJul 29, 2026 - #26 updated
22 minutes agoJul 29, 2026 - #16 updated
22 minutes agoJul 29, 2026 - #43 updated
26 minutes agoJul 29, 2026 - Work on the Transformers modeling backend: running Transformers model implementations inside vLLM.#28 updated
2 hours agoJul 29, 2026 -
-
- Maintainer's tracking board for Prometheus metrics related PRs and issues#44 updated
15 hours agoJul 29, 2026 -
-
-
DeepSeek V3/R1 Template
2025-02-25: DeepSeek V3/R1 is supported with optimized block FP8 kernels, MLA, MTP spec decode, multi-node PP, EP, and W4A16 quantization#5 updated2 days agoJul 28, 2026 -
- Main tasks for the multi-modality workstream (#4194)#8 updated
2 days agoJul 27, 2026 -
- #47 updated
4 days agoJul 26, 2026 - Community requests for multi-modal models#10 updated
4 days agoJul 25, 2026 - #57 updated
5 days agoJul 24, 2026 - #24 updated
last weekJul 24, 2026 - #51 updated
last weekJul 24, 2026 - #55 updated
last weekJul 20, 2026 - #46 updated
2 weeks agoJul 16, 2026 - #29 updated
2 weeks agoJul 14, 2026 - #33 updated
on Jun 22Jun 22, 2026 - A list of onboarding tasks for first-time contributors to get started with vLLM.#6 updated
on May 31May 31, 2026 - #45 updated
on May 17May 16, 2026 -
- #25 updated
on Mar 7Mar 7, 2026 - Tracker of known issues and bugs for serving Llama on vLLM#14 updated
on Feb 6Feb 6, 2026