[ MENU ]
LIVE
────────
·
──:──:── UTC
about
⌘
K
search
?
help
[ MENU ]
All niches
Niche
inference
Inference & Serving
[
TOOLS_TRACKED
]
32
[
ACCELERATING
]
0
[
DYING
]
0
[
AVG_VELOCITY
]
0.56
/10
See our pick → Best Inference & Serving
[
ACCELERATING
]
Accelerating
0
No accelerating tools right now.
[
STABLE
]
Stable
12
Tool
Velocity
Trend 30d
Δ 7d
Stars
Class
ray-project/ray
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
2.27
↑ +71
44k
Stable
google-ai-edge/mediapipe
Cross-platform, customizable ML solutions for live and streaming media.
2.01
↑ +65
37k
Stable
NVIDIA/nvcf
Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.
1.93
↑ +7
221
Stable
tetherto/qvac
Open-source local AI SDK - run AI on-device with no cloud, no API keys. Supports GGUF, RAG, image, music, and video generation, speech-to-text, P2P inference, and more. Cross-platform: Linux, macOS, Windows, Android, iOS.
1.54
↑ +19
624
Stable
gpustack/gpustack
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
1.22
↑ +39
5.7k
Stable
Avarok-Cybersecurity/atlas
Pure Rust Inference Engine
1.05
↑ +18
700
Stable
OpenRLHF/OpenRLHF
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
0.87
↑ +24
10k
Stable
xorbitsai/inference
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
0.77
↑ +18
9.6k
Stable
dstackai/dstack
A unified orchestration layer for heterogeneous AI compute. It standardizes how to manage compute and run training and inference on GPU clouds, Kubernetes, VMs, or bare-metal clusters.
0.77
↑ +9
2.3k
Stable
Lightning-AI/litgpt
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
0.66
↑ +12
14k
Stable
bentoml/BentoML
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
0.66
↑ +11
8.9k
Stable
mozilla-ai/any-llm
Communicate with an LLM provider using a single interface
0.65
↑ +8
2.2k
Stable
LIVE
────────
·
──:──:── UTC
about
⌘
K
search
?
help