Repo Compare › LLM inference

vLLM vs llama.cpp vs Text Generation Inference

A side-by-side look at how actively vLLM, llama.cpp and Text Generation Inference are developed on GitHub — popularity, contributor base, commit activity and how quickly issues and pull requests move.

What the numbers say

Head-to-head

Snapshot taken 30 September 2026.

MetricvLLM
Active
llama.cpp
Active
Text Generation Inference
Archived
Stars93k130k11k
Forks23k24k1.3k
Contributors3.6k2.1k148
Commits, last 52 weeks12k4.6k8
Open issues2.5k871285
Open pull requests5.8k1.7k39
Closed / merged PRs34k14k1.7k
Watchers598841103
Last push30 Sep 202630 Sep 202621 Mar 2026
Latest releasev0.30.0 (Sep 2026)v0.5.0 (Sep 2026)v3.3.7 (Dec 2025)
Main languagePythonC++Python
LicenceApache-2.0MITApache-2.0
CreatedFeb 2023Mar 2023Oct 2022

Bold marks the highest value in a row. “Active / Slowing / Dormant” is a simple heuristic from how recently code was pushed and open issues per star — a starting point, not a verdict. Stars measure popularity, not quality.

Open these in the live comparer →

Related comparisons