trackmcp
Back to directory
RudrenduPaul

InferBench

View on GitHub

Benchmark local LLM inference engines (llama.cpp, omlx) on your own hardware with real tokens/sec, not borrowed numbers

0 stars PythonOthers Updated Aug 25, 2026
apple-siliconbenchmarkbenchmarkingclideveloper-toolsggufinferencellama-cppllmlocal-llmmlxomlxpythontokens-per-secondtypescript

No README could be loaded. View the project on GitHub for full documentation.

Frequently asked questions

What is InferBench?

InferBench is Benchmark local LLM inference engines (llama.cpp, omlx) on your own hardware with real tokens/sec, not borrowed numbers

How do I install InferBench?

Open the GitHub repository and follow its README. Most MCP servers are added to your client's MCP config, then called by your agent.

Is InferBench open source?

Yes — it is hosted on GitHub at https://github.com/RudrenduPaul/InferBench.

Related MCP tools

Run your own MCP server? See who uses it and what to fix.

Measure it with TrackMCP