InferBench
Benchmark local LLM inference engines (llama.cpp, omlx) on your own hardware with real tokens/sec, not borrowed numbers
No README could be loaded. View the project on GitHub for full documentation.
Frequently asked questions
What is InferBench?
InferBench is Benchmark local LLM inference engines (llama.cpp, omlx) on your own hardware with real tokens/sec, not borrowed numbers
How do I install InferBench?
Open the GitHub repository and follow its README. Most MCP servers are added to your client's MCP config, then called by your agent.
Is InferBench open source?
Yes — it is hosted on GitHub at https://github.com/RudrenduPaul/InferBench.
Related MCP tools
AI-powered OSINT agent with interactive REPL, MCP server, and CLI. 19 tools. Works with Claude, GPT-4, or local models. For authorized security research only.
Open-source coding agent memory. Records issues, attempts, fixes and decisions, then warns your agent before it repeats an approach that already failed. Native MCP server for Claude Code, Cursor, Antigravity and Codex. 100% local, no cloud, no telemetry. MIT.
🙌 OpenHands: Code Less, Make More for the Model Context Protocol. Enhance AI assistants with powerful integrations. Python-based implementation.
Automate browser based workflows with AI
Cut AI token costs 95%+ on code exploration. The leading MCP server for precise, symbol-level GitHub code retrieval via tree-sitter AST. Works with Claude Code, Cursor & any MCP client. 313B+ tokens saved.
Composio equips your AI agents & LLMs with 100+ high-quality integrations via function calling for the Model Context Protocol. Enhance AI assistants with powerf
Run your own MCP server? See who uses it and what to fix.
Measure it with TrackMCP