Skip to content

LLM inference engines: star history compared

Running large language models yourself went from research stunt to commodity in about two years, and these five projects did most of the commoditizing. llama.cpp proved frontier-class models could run on a laptop, Ollama wrapped that power in a one-line install, and vLLM and SGLang turned GPU serving throughput into a public leaderboard of its own. The star histories here are effectively a timeline of the local-AI movement.

Ranked by current GitHub stars · growth = stars gained in the trailing 90 days · data updated Jul 2026

LLM inference engines — star history

Ranking

5 repositories

#RepositoryStars+90dFirst starHealth
1ollama/ollama177,460~3572023debt charts
2ggml-org/llama.cpp122,290~7812023debt charts
3vllm-project/vllm87,826~3882023debt charts
4sgl-project/sglang31,027~3532024debt charts
5huggingface/text-generation-inference10,887~222022debt charts

Decision matrix

Repository health compared

Live Postgres analysis · unavailable data stays labeled

Popularity is only one axis. Compare technical scale, maintenance cadence, ownership resilience, and change pressure using the same completed repository analysis behind each report.

:: Reach

Popularity and real-world adoption

Signalollama/ollamaggml-org/llama.cppvllm-project/vllmsgl-project/sglanghuggingface/text-generation-inference
GitHub starsCurrent public GitHub metadata total.177,460122,29087,82631,02710,887
Stars gained · trailing 90dChange in the newest 90 days covered by each cached series.+357+781+388+353+22
ForksCurrent public GitHub metadata total.Not availableNot availableNot availableNot availableNot available
Largest package audienceLargest resolved registry total; sources are shown explicitly.No resolved packageNo resolved packageNo resolved packageNo resolved packageNo resolved package
Repository ageElapsed years since the GitHub creation timestamp.3.1 years3.4 years3.5 years2.6 years3.8 years

:: Codebase

Size and technical center of gravity

Signalollama/ollamaggml-org/llama.cppvllm-project/vllmsgl-project/sglanghuggingface/text-generation-inference
Dominant languageLargest share of analyzed code lines at current HEAD.Not availableNot availableNot availableNot availableNot available
Code linesAnalyzed source lines at current HEAD, excluding blanks and comments.Not availableNot availableNot availableNot availableNot available
Languages representedLanguages with non-zero files or line counts.Not availableNot availableNot availableNot availableNot available

:: Maintenance

Commit volume and recent operating cadence

Signalollama/ollamaggml-org/llama.cppvllm-project/vllmsgl-project/sglanghuggingface/text-generation-inference
Repository commitsFull repository commit total reported by the analysis.Not availableNot availableNot availableNot availableNot available
Commits · trailing 90dCommit count in the newest 90 days covered by the repository analysis.Not availableNot availableNot availableNot availableNot available

:: Ownership

How broadly maintenance responsibility is distributed

Signalollama/ollamaggml-org/llama.cppvllm-project/vllmsgl-project/sglanghuggingface/text-generation-inference
Bus factorFewest analyzed contributors carrying at least half the commits.Not availableNot availableNot availableNot availableNot available
Top contributor concentrationShare of attributed commits carried by the most active contributor.Not availableNot availableNot availableNot availableNot available

:: Change signals

Debt markers and fix pressure in frequently changed files

Signalollama/ollamaggml-org/llama.cppvllm-project/vllmsgl-project/sglanghuggingface/text-generation-inference
Fix-labelled share · frequently changed filesFix-labelled commits divided by commits across the analyzed high-frequency file set.Not availableNot availableNot availableNot availableNot available
Recent TODO/FIXME movementLatest cumulative marker count and its change over the trailing 90 days.Not availableNot availableNot availableNot availableNot available

Frequently asked

Which of these llm inference engines has the most GitHub stars?

ollama/ollama is the most-starred of the llm inference engines gitdebt compares, with 177,460 GitHub stars as of Jul 2026. It earned its first star in 2023.

Which of these llm inference engines is growing the fastest?

ggml-org/llama.cpp added the most stars over the last 90 days of its cached history — about 781 new stars, compared with about 357 for the category leader ollama/ollama. 90-day figures are approximate for older repos, whose history is downsampled.

How is this comparison computed?

The table ranks public repository metadata by current total stars. When historical stargazer timestamps are available, gitdebt plots their cumulative series on shared axes and estimates trailing 90-day growth. Repo-health charts come independently from each repository's commit history.

Related categories

all categories

Fastest-growing repos across all of gitdebt