Skip to content
The Internet Compass

Repository

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

amdblackwellcudadeepseekdeepseek-v3gptgpt-ossinference
★ 91.1k21.8k forksPythonView on GitHub ↗

Homepage

https://vllm.ai ↗

Sourced from GitHub · Updated September 6, 2026

Data from the GitHub REST API. Not affiliated with or endorsed by GitHub.

View original source ↗Spot an error on this page? Let us know →

FAQ

Common questions

What is vllm-project/vllm?

A high-throughput and memory-efficient inference and serving engine for LLMs

What license does vllm-project/vllm use?

vllm-project/vllm is licensed under Apache-2.0.

How popular is vllm-project/vllm?

vllm-project/vllm has 91.1k stars and 21.8k forks on GitHub, with 7.6k open issues.

Is vllm-project/vllm actively maintained?

The last push to vllm-project/vllm was September 6, 2026.