io.github.RudrenduPaul/inferbench
Benchmarks local LLM inference speed (tokens/sec) on your own hardware via MCP t
Before you connect
Config example (to connect)
{
"mcpServers": {
"inferbench": {
"command": "uvx",
"args": ["--from", "inferbench-cli", "inferbench-mcp"]
}
}
}Works with these clients
- Claude Desktopuvx
- Claude Codeuvx
- Cursoruvx
- VS Codeuvx
- Windsurfuvx
Derived from the registry manifest (package type and transport) — we have not installed it ourselves. Check the official README before relying on it.
Install / Connect
claude mcp add io-github-rudrendupaul-inferbench -- uvx inferbench-cli@0.1.7Follow the official README; env and args vary per server.
🛠️ Maintenance weight 30% · 80 pts
📊 Adoption weight 25% · 6 pts
✅ Usability weight 20% · 100 pts
❤️ Health & community weight 25% · 53 pts
Sources: GitHub API · Official registry · updated 2026-08-11
90-day trend
⭐ stars
Similar / alternative servers
io.github.RudrenduPaul/workloadtruth
Classifies GPU workloads as inference or training from telemetry alone via MCP t
io.github.NORTHTEKDevs/genome
Fully local memory for AI agents: zero-LLM ingest, semantic recall, air-gapped,
io.github.RudrenduPaul/agent-eval
Statistical regression testing for LLM agents: p-value, effect size, and CI on b
io.github.Declade/lucairn-mcp-server
Pseudonymizes PII before your LLM and returns a cryptographically signed receipt
One email a week, 5 minutes to see what happened in the MCP ecosystem
New servers, breakout trends, abandonment alerts — tracked automatically, curated by hand, sent to your inbox.
One email a week · unsubscribe anytime · no data selling