beam
beam
Discover
PulseActivityAnalyticsBest forMapOrgs
Niches
AgentsMCPRAGCoding AssistantsInference & ServingVector DBs
Personal
WatchlistCompareWeekly report
?
Sign in
magic link · no password
LIVE──────── · ──:──:── UTCabout
beam
beam
Discover
PulseActivityAnalyticsBest forMapOrgs
Niches
AgentsMCPRAGCoding AssistantsInference & ServingVector DBs
Personal
WatchlistCompareWeekly report
?
Sign in
magic link · no password
All niches
Nicheinference

Inference & Serving

[TOOLS_TRACKED]35
[ACCELERATING]0
[DYING]0
[AVG_VELOCITY]0.47/10
See our pick → Best Inference & Serving
[ACCELERATING]

Accelerating

0

No accelerating tools right now.

[STABLE]

Stable

10
Tool
Velocity
Trend 30d
Δ 7d
Stars
Class
  • vllm-project/vllmA high-throughput and memory-efficient inference and serving engine for LLMs
    4.29↑ +61889kStable
  • google-ai-edge/mediapipeCross-platform, customizable ML solutions for live and streaming media.
    2.24↑ +11637kStable
  • ray-project/rayRay is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
    2.19↑ +6743kStable
  • xLLM-AI/xllmA high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
    1.22↑ +101.5kStable
  • pykeio/ortFast ML inference & training for ONNX models in Rust
    0.82↑ +142.4kStable
  • Avarok-Cybersecurity/atlasPure Rust Inference Engine
    0.82↑ +10634Stable
  • stas00/ml-engineeringMachine Learning Engineering Open Book
    0.73↑ +5619kStable
  • xorbitsai/inferenceSwap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
    0.69↑ +139.5kStable
  • OpenRLHF/OpenRLHFAn Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
    0.63↑ +289.9kStable
  • dstackai/dstackVendor-agnostic orchestration for training, inference and agentic workloads across NVIDIA, AMD, TPU, and Tenstorrent on clouds, Kubernetes, and bare metal.
    0.60↑ +62.2kStable
LIVE──────── · ──:──:── UTCabout