mcp-agent-reliability

mcp-agent-reliability

MCP server that scores tool descriptions, estimates token costs, simulates agent tool selection, and generates reliability reports to help AI agents choose the right tools and reduce wasted tokens.

Category
Visit Server

README

mcp-agent-reliability

Make your AI agents more reliable.

This MCP server acts like a reliability coach for your agents.
It helps you:

  • Score how clear your tool descriptions are (so the agent picks the right one)
  • Estimate how many tokens your tools will cost
  • Simulate which tool an agent would choose for a task
  • Generate simple test prompts
  • Get a full reliability report

Built for entrepreneurs and teams who are tired of agents calling the wrong tools and burning money.

Why this exists (simple story)

Imagine you give a 10-year-old child a big list of 30 toys and say “go play with the right one”.
If the labels are confusing, the child will pick the wrong toy.

AI agents are the same.
When you connect many MCP servers, the agent sees a long menu of tools.
If the descriptions are vague, it picks the wrong tool → wasted tokens → failed tasks.

This server is the “label checker” and “practice teacher” for that menu.

Quick start

# clone
git clone https://github.com/princeruhulofficial/mcp-agent-reliability.git
cd mcp-agent-reliability

# install
npm install

# build
npm run build

# run (stdio)
npm start

Add to Claude Desktop / Cursor / any MCP client

{
  "mcpServers": {
    "agent-reliability": {
      "command": "node",
      "args": ["/absolute/path/to/mcp-agent-reliability/dist/index.js"]
    }
  }
}

Or with npx (after publish):

{
  "mcpServers": {
    "agent-reliability": {
      "command": "npx",
      "args": ["-y", "mcp-agent-reliability"]
    }
  }
}

Tools

Tool What it does
score_tool_description Gives a 0-100 score + reasons + suggestions for a tool description
estimate_token_cost Rough token count for a list of tools
simulate_tool_choice Predicts which tool an agent would pick for a prompt
generate_agent_tests Creates 3 test prompts you can run against your agent
reliability_report Full summary of scores + token estimates

All tools are pure computation — no paid API keys required.

Example

Score a description:

Tool: score_tool_description
name: create_invoice
description: Create a new invoice for a customer. Requires customer_id and amount. Returns invoice_id.

You get something like:

{
  "score": 85,
  "reasons": ["Good length...", "Mentions inputs or outputs..."],
  "suggestions": [],
  "interpretation": "Excellent — agent should pick this tool reliably"
}

Tech

  • TypeScript
  • Official @modelcontextprotocol/sdk
  • Stateless-friendly (works with 2026 MCP updates)
  • Zero external cost for core features

Roadmap

  • [ ] Optional LLM-backed scoring (when you want higher accuracy)
  • [ ] Hosted version with dashboard
  • [ ] Integration with progressive disclosure patterns

License

MIT


Made with ❤️ for the Prevalid community
Founder: Prince Ruhul

Recommended Servers

playwright-mcp

playwright-mcp

A Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.

Official
Featured
TypeScript
Audiense Insights MCP Server

Audiense Insights MCP Server

Enables interaction with Audiense Insights accounts via the Model Context Protocol, facilitating the extraction and analysis of marketing insights and audience data including demographics, behavior, and influencer engagement.

Official
Featured
Local
TypeScript
Magic Component Platform (MCP)

Magic Component Platform (MCP)

An AI-powered tool that generates modern UI components from natural language descriptions, integrating with popular IDEs to streamline UI development workflow.

Official
Featured
Local
TypeScript
VeyraX MCP

VeyraX MCP

Single MCP tool to connect all your favorite tools: Gmail, Calendar and 40 more.

Official
Featured
Local
Kagi MCP Server

Kagi MCP Server

An MCP server that integrates Kagi search capabilities with Claude AI, enabling Claude to perform real-time web searches when answering questions that require up-to-date information.

Official
Featured
Python
graphlit-mcp-server

graphlit-mcp-server

The Model Context Protocol (MCP) Server enables integration between MCP clients and the Graphlit service. Ingest anything from Slack to Gmail to podcast feeds, in addition to web crawling, into a Graphlit project - and then retrieve relevant contents from the MCP client.

Official
Featured
TypeScript
Neon Database

Neon Database

MCP server for interacting with Neon Management API and databases

Official
Featured
Exa Search

Exa Search

A Model Context Protocol (MCP) server lets AI assistants like Claude use the Exa AI Search API for web searches. This setup allows AI models to get real-time web information in a safe and controlled way.

Official
Featured
Qdrant Server

Qdrant Server

This repository is an example of how to create a MCP server for Qdrant, a vector search engine.

Official
Featured
E2B

E2B

Using MCP to run code via e2b.

Official
Featured