Repository
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
amdblackwellcudadeepseekdeepseek-v3gptgpt-ossinference
Homepage
https://vllm.ai ↗Sourced from GitHub · Updated September 6, 2026
Data from the GitHub REST API. Not affiliated with or endorsed by GitHub.
View original source ↗Spot an error on this page? Let us know →FAQ
Common questions
What is vllm-project/vllm?
A high-throughput and memory-efficient inference and serving engine for LLMs
What license does vllm-project/vllm use?
vllm-project/vllm is licensed under Apache-2.0.
How popular is vllm-project/vllm?
vllm-project/vllm has 91.1k stars and 21.8k forks on GitHub, with 7.6k open issues.
Is vllm-project/vllm actively maintained?
The last push to vllm-project/vllm was September 6, 2026.