[ MENU ]
LIVE
────────
·
──:──:── UTC
about
⌘
K
search
?
help
[ MENU ]
All niches
Niche
inference
Inference & Serving
[
TOOLS_TRACKED
]
35
[
ACCELERATING
]
0
[
DYING
]
0
[
AVG_VELOCITY
]
0.47
/10
See our pick → Best Inference & Serving
[
ACCELERATING
]
Accelerating
0
No accelerating tools right now.
[
STABLE
]
Stable
10
Tool
Velocity
Trend 30d
Δ 7d
Stars
Class
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
4.29
↑ +618
89k
Stable
google-ai-edge/mediapipe
Cross-platform, customizable ML solutions for live and streaming media.
2.24
↑ +116
37k
Stable
ray-project/ray
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
2.19
↑ +67
43k
Stable
xLLM-AI/xllm
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
1.22
↑ +10
1.5k
Stable
pykeio/ort
Fast ML inference & training for ONNX models in Rust
0.82
↑ +14
2.4k
Stable
Avarok-Cybersecurity/atlas
Pure Rust Inference Engine
0.82
↑ +10
634
Stable
stas00/ml-engineering
Machine Learning Engineering Open Book
0.73
↑ +56
19k
Stable
xorbitsai/inference
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
0.69
↑ +13
9.5k
Stable
OpenRLHF/OpenRLHF
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
0.63
↑ +28
9.9k
Stable
dstackai/dstack
Vendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal.
0.60
↑ +6
2.2k
Stable
LIVE
────────
·
──:──:── UTC
about
⌘
K
search
?
help