STACKQUADRANT

gpustack/gpustack

Inference Engines

Performance-optimized AI inference on your GPUs. Unlock superior throughput by selecting and tuning engines like vLLM or SGLang.

7.4
GitHub Metrics
Stars
5.6k
Forks
629
Open Issues
696
Watchers
43
Contributors
54
Weekly Commits
0
Language
Python
License
Apache-2.0
Last Commit
Aug 28, 2026
Created
May 11, 2024
Latest Release
v2.2.3
Release Date
Jul 31, 2026
Synced: Aug 28, 2026
Quality Scores
Documentation Qualityw: 20%
7.4

Has docs site (https://gpustack.ai). Description: 128 chars. Stars signal: 5,566. Contributors: 54. Score: 7.4/10

Community Healthw: 20%
6.4

Stars: 5,566. Contributors: 54. Watchers: 43. Forks: 628. Issue ratio: 12.5%. Score: 6.4/10

Maintenance Velocityw: 15%
9.8

Last commit: 1d ago. Weekly commits: 25. Latest release: v2.2.3. Maturity bonus: 2.3y old. Score: 9.8/10

API Design & DXw: 20%
6.4

Stars/issues ratio: 8. Dynamic language: Python. Has documentation site. Permissive license: Apache-2.0. Popularity signal: 5,566 stars. Score: 6.4/10

Production Readinessw: 15%
7.4

Battle-tested: 5,566 stars. Peer review: 54 contributors. Versioned: v2.2.3. Licensed: Apache-2.0. Age: 2.3 years. Maintenance: last commit 1d ago. Score: 7.4/10

Ecosystem Integrationw: 10%
8.0

Fork interest: 628. Major ecosystem: Python. Integration-friendly: Apache-2.0. Adoption: 5,566 stars. Has web presence. Score: 8/10

Tags
ascendcudadeepseekdistributed-inferencegenaihigh-performance-inferenceinferencellamallmllm-inference
Radar
Documentation Quality
Community Health
Maintenance Velocity
API Design & DX
Production Readiness
Ecosystem Integration