beam
beam
Discover
PulseActivityAnalyticsBest forMapOrgs
Niches
AgentsMCPRAGCoding AssistantsInference & ServingVector DBs
Personal
WatchlistCompareWeekly report
?
Sign in
magic link · no password
LIVE──────── · ──:──:── UTCabout
beam
beam
Discover
PulseActivityAnalyticsBest forMapOrgs
Niches
AgentsMCPRAGCoding AssistantsInference & ServingVector DBs
Personal
WatchlistCompareWeekly report
?
Sign in
magic link · no password
Back to Pulse
Tool profile

JakeATX/llamAmpere

new

llama.cpp fork for significantly improved performance on Ampere (especially RTX 3090 / 3090 Ti): TurboQuant KV cache, MTP speculative decoding with a 64K draft-vocabulary shortlist, custom SM86 + Qwen kernels. 90 tok/s over a 100K-token generation at temperature 1.

ampereconsumer-gpucudaggmlggufkv-cache-quantizationllama-cppllm
Velocity score
0.00/ 10
[STARS]
105
[FORKS]
14
[CONTRIBUTORS]
2.0k
[LAST_COMMIT]
8d ago
OPEN_ON_GITHUB
Velocity class: new
Is JakeATX/llamAmpere still actively maintained?
[VR]

This tool is in the Velocity Report when it moves. Get the weekly numbers.

Score breakdown
403/ 1000
inference · JakeATX/llamAmpere
Velocity50%
Adoption30%
Maintenance15%
Community5%
[CODE_GROWTH]
223
[INSTALL_VEL]
498
[ACTIVITY]
613
[COMMUNITY_SIGNAL]
1000

Terminal score: 0–1000 raw, weighted across 4 dimensions. Public score: 0–10 normalized (shown in the 30-day stars chart above).

LIVE──────── · ──:──:── UTCabout