Agent Reliability
FreeNot checkedMCP server that scores tool descriptions, estimates token costs, simulates agent tool selection, and generates reliability reports to help AI agents choose the
About
MCP server that scores tool descriptions, estimates token costs, simulates agent tool selection, and generates reliability reports to help AI agents choose the right tools and reduce wasted tokens.
README
Make your AI agents more reliable.
This MCP server acts like a reliability coach for your agents.
It helps you:
- Score how clear your tool descriptions are (so the agent picks the right one)
- Estimate how many tokens your tools will cost
- Simulate which tool an agent would choose for a task
- Generate simple test prompts
- Get a full reliability report
Built for entrepreneurs and teams who are tired of agents calling the wrong tools and burning money.
Why this exists (simple story)
Imagine you give a 10-year-old child a big list of 30 toys and say “go play with the right one”.
If the labels are confusing, the child will pick the wrong toy.
AI agents are the same.
When you connect many MCP servers, the agent sees a long menu of tools.
If the descriptions are vague, it picks the wrong tool → wasted tokens → failed tasks.
This server is the “label checker” and “practice teacher” for that menu.
Quick start
# clone
git clone https://github.com/princeruhulofficial/mcp-agent-reliability.git
cd mcp-agent-reliability
# install
npm install
# build
npm run build
# run (stdio)
npm start
Add to Claude Desktop / Cursor / any MCP client
{
"mcpServers": {
"agent-reliability": {
"command": "node",
"args": ["/absolute/path/to/mcp-agent-reliability/dist/index.js"]
}
}
}
Or with npx (after publish):
{
"mcpServers": {
"agent-reliability": {
"command": "npx",
"args": ["-y", "mcp-agent-reliability"]
}
}
}
Tools
| Tool | What it does |
|---|---|
score_tool_description |
Gives a 0-100 score + reasons + suggestions for a tool description |
estimate_token_cost |
Rough token count for a list of tools |
simulate_tool_choice |
Predicts which tool an agent would pick for a prompt |
generate_agent_tests |
Creates 3 test prompts you can run against your agent |
reliability_report |
Full summary of scores + token estimates |
All tools are pure computation — no paid API keys required.
Example
Score a description:
Tool: score_tool_description
name: create_invoice
description: Create a new invoice for a customer. Requires customer_id and amount. Returns invoice_id.
You get something like:
{
"score": 85,
"reasons": ["Good length...", "Mentions inputs or outputs..."],
"suggestions": [],
"interpretation": "Excellent — agent should pick this tool reliably"
}
Tech
- TypeScript
- Official
@modelcontextprotocol/sdk - Stateless-friendly (works with 2026 MCP updates)
- Zero external cost for core features
Roadmap
- Optional LLM-backed scoring (when you want higher accuracy)
- Hosted version with dashboard
- Integration with progressive disclosure patterns
License
MIT
Made with ❤️ for the Prevalid community
Founder: Prince Ruhul
Install Agent Reliability in Claude Desktop, Claude Code & Cursor
unyly install mcp-agent-reliabilityInstalls into Claude Desktop, Claude Code, Cursor & VS Code — handles npx, uvx and build-from-source repos for you.
First time? Get the CLI: curl -fsSL https://unyly.org/install | sh
Or configure manually
Run in your terminal:
claude mcp add mcp-agent-reliability -- npx -y github:princeruhulofficial/mcp-agent-reliabilityStep-by-step: how to install Agent Reliability
FAQ
Is Agent Reliability MCP free?
Yes, Agent Reliability MCP is free — one-click install via Unyly at no cost.
Does Agent Reliability need an API key?
No, Agent Reliability runs without API keys or environment variables.
Is Agent Reliability hosted or self-hosted?
Self-hosted: the server runs locally on your machine via the install command above.
How do I install Agent Reliability in Claude Desktop, Claude Code or Cursor?
Open Agent Reliability on unyly.org, pick your client tab (Claude Desktop, Claude Code, Cursor) and press Install — the config is generated automatically, no JSON editing.
Related MCPs
Fetch
Web content fetching and conversion for efficient LLM usage.
AWS KB Retrieval
Retrieval from AWS Knowledge Base using Bedrock Agent Runtime.
by modelcontextprotocolSpring AI MCP Server
Provides auto-configuration for setting up an MCP server in Spring Boot applications.
llm-analysis-assistant
A very streamlined mcp client that supports calling and monitoring stdio/sse/streamableHttp, and can also view request responses through the /logs page. It also
by xuzexin-hzCompare Agent Reliability with
Not sure what to pick?
Find your stack in 60 seconds
Author?
Embed badge for your README
Browse similar
All ai MCPs
