vllm-project-vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
- Score
- 99.9473 signals
- Evidence
- 90,743 stars · 21,577 forks · 15 mentions
- Last commit
- as last read from GitHub; most reads are from 2 Sep 2026 or later
- Listed
Install
No one-command install. Set it up from its source.
Alternatives
- sgl-project-sglang33,202 stars · 8,472 forks · 4 mentions99.787
What it is
A high-throughput and memory-efficient inference and serving engine for LLMs
Notes
Surfaced by the Sentinel→Armory feed (practitioner mentions), 2026-09-02.