MCP Server · AI & ML

inferbench

Benchmark local LLM inference speed (tokens/sec) on your own hardware — llama.cpp native + cloud APIs, 124-model catalog, optimal-quant picker, and an MCP serve mode.

Tier 1 sealed by lossy-channel-auto
Source
https://github.com/JoniMartin27/inferbench
Registry namespace
io.github.jonimartin27/inferbench
Imported from
github-topics
License
MIT
Pricing
Open Source
Added
Jul 23, 2026

Seal evidence

Auto-sealed Tier 1 on import from GitHub topic search. Lossy Channel layers 1–2 to be backfilled by the Foreman.

  • L1-wash

Similar tools