mcp-hub

mcp-hub

An extensible MCP hub that exposes internal services (chat, observability, RAG) as namespaced tools via FastMCP, with OpenAPI auto-generation, auth, and resilient error handling.

Category
Visit Server

README

Task: Build "mcp-hub" — a production-grade, extensible MCP server (hackathon project)

Context

I need an independent, network-hosted MCP server that AI agents (Claude, LangGraph, custom clients) connect to over the network. It fronts three existing internal services and must be trivially extensible when new services appear.

Services already running in the network:

  1. MiniSlack (Slack-like chat) at http://minislack:9001 — REST endpoints: POST /messages {channel, text}, GET /messages?channel=X, GET /channels
  2. Observability service at http://observability:9002 — REST endpoints: GET /metrics?service=X&window=5m, GET /alerts, GET /services
  3. Enterprise RAG (FastAPI) at http://rag:9003 — full OpenAPI spec at http://rag:9003/openapi.json; key endpoint: POST /query {question, top_k}

If an endpoint above is wrong, check the service's /docs or /openapi.json first, then adapt — do not hardcode assumptions silently; log a warning instead.

Hard requirements

  1. Framework: Python 3.12, FastMCP v3 standalone package (fastmcp, install via uv add fastmcp). Do NOT use mcp.server.fastmcp.
  2. Transport: Streamable HTTP at /mcp, stateless_http=True, host 0.0.0.0, port from env MCP_PORT (default 8000). Plus a plain /health endpoint.
  3. Connector architecture (the critical part):
    • config/services.yaml is the service registry:
      services:
        - name: minislack
          namespace: slack
          type: rest           # hand-written connector
          base_url: http://minislack:9001
          enabled: true
        - name: rag
          namespace: rag
          type: openapi        # auto-generated from OpenAPI spec
          base_url: http://rag:9003
          enabled: true
      
    • src/connectors/base.py defines a Connector protocol: name, namespace, async register(mcp: FastMCP) -> None.
    • src/core/registry.py loads services.yaml, imports the matching connector module (or builds one from OpenAPI for type: openapi), mounts each as a namespaced sub-server via FastMCP mounting so tools appear as slack.send_message, obs.get_metrics, rag.query.
    • Adding a future service = one new file in src/connectors/ + one YAML entry. Include src/connectors/template.py as a documented copy-paste starting point.
  4. RAG connector: use FastMCP v3's OpenAPI provider to auto-generate tools from http://rag:9003/openapi.json at startup. If the spec is unreachable, log the failure and continue serving the other connectors (graceful degradation — one dead upstream must never crash the hub).
  5. Resilience: shared httpx.AsyncClient with per-upstream timeout (5s), 2 retries with backoff; every tool catches upstream errors and returns a structured error object {"error": "...", "upstream": "...", "retryable": true} — never leak stack traces to the agent.
  6. Validation: validate every tool argument (Pydantic); reject empty channel names, negative top_k, etc.
  7. Auth: bearer-token middleware on /mcp — token from env MCP_AUTH_TOKEN; requests without Authorization: Bearer <token> get 401. /health stays open.
  8. Observability: structured JSON logs (structlog or stdlib json formatter); for every tool call log: tool name, namespace, latency_ms, status, upstream, error type (never log full argument values — log arg keys only).
  9. Tests: pytest with FastMCP's in-memory Client — one test file per connector mocking the upstream with respx or httpx MockTransport, plus a registry test proving all enabled services mount correctly.
  10. Docker: multi-stage Dockerfile (python:3.12-slim, non-root user, uv for deps, HEALTHCHECK on /health) and docker-compose.yaml with the hub plus stub implementations of the three services (tiny FastAPI stubs) so the whole demo runs offline.

Deliverables (in order)

  1. Project scaffold + dependency setup that runs: uv run python -m src.main serves /mcp and /health.
  2. slack.* and obs.* connectors with tests.
  3. OpenAPI-driven rag.* connector with graceful-degradation test.
  4. Auth, logging, Dockerfile, docker-compose.
  5. demo_client.py: a script that connects with FastMCP Client, lists all tools, then chains rag.query → slack.send_message → obs.get_metrics to prove end-to-end agent flow.

Build incrementally in that order, running tests after each step. Ask me for the real endpoint specs only if you cannot proceed with the ones above.

Recommended Servers

playwright-mcp

playwright-mcp

A Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.

Official
Featured
TypeScript
Audiense Insights MCP Server

Audiense Insights MCP Server

Enables interaction with Audiense Insights accounts via the Model Context Protocol, facilitating the extraction and analysis of marketing insights and audience data including demographics, behavior, and influencer engagement.

Official
Featured
Local
TypeScript
Magic Component Platform (MCP)

Magic Component Platform (MCP)

An AI-powered tool that generates modern UI components from natural language descriptions, integrating with popular IDEs to streamline UI development workflow.

Official
Featured
Local
TypeScript
VeyraX MCP

VeyraX MCP

Single MCP tool to connect all your favorite tools: Gmail, Calendar and 40 more.

Official
Featured
Local
graphlit-mcp-server

graphlit-mcp-server

The Model Context Protocol (MCP) Server enables integration between MCP clients and the Graphlit service. Ingest anything from Slack to Gmail to podcast feeds, in addition to web crawling, into a Graphlit project - and then retrieve relevant contents from the MCP client.

Official
Featured
TypeScript
Kagi MCP Server

Kagi MCP Server

An MCP server that integrates Kagi search capabilities with Claude AI, enabling Claude to perform real-time web searches when answering questions that require up-to-date information.

Official
Featured
Python
E2B

E2B

Using MCP to run code via e2b.

Official
Featured
Neon Database

Neon Database

MCP server for interacting with Neon Management API and databases

Official
Featured
Exa Search

Exa Search

A Model Context Protocol (MCP) server lets AI assistants like Claude use the Exa AI Search API for web searches. This setup allows AI models to get real-time web information in a safe and controlled way.

Official
Featured
Qdrant Server

Qdrant Server

This repository is an example of how to create a MCP server for Qdrant, a vector search engine.

Official
Featured