Skip to content
Fossfits
Search
Categories
Guides
Compare
Tags
openai-api.
2
listed
·
Updated Sep 11, 2026
Rank 1.
01
V
vLLM
High-throughput model server for GPU inference.
Apache 2.0 · Linux, Docker
GitHub stars:
58k
Rank 2.
02
X
Xinference
Serve LLMs, embeddings, and media models from one Apache-2.0 platform.
Apache 2.0 · Linux, macOS, +2
GitHub stars:
9.6k