MCPVault
Inferbench logo

Inferbench

Unclaimed

by JoniMartin27

Benchmark local LLM inference speed (tokens/sec) on your own hardware — llama.cpp native + cloud APIs, 124-model catalog, optimal-quant picker, and an MCP serve mode.

benchmarkelectronfastapiggufinferencellama-cppllmlocal-llmpythonai

Unclaimed listing

Is this your MCP server?

This listing was auto-indexed from the public record. Claim it to edit the page, set compatibility and unlock growth tools. Takes under two minutes.

Claim this server