TachiBot MCP
Multi-model AI orchestration platform with 64 tools across 12 AI providers, enabling intelligent workflows from any MCP client.
README
<div align="center">
TachiBot MCP
Multi-Model AI Orchestration Platform
64 AI tools. 12 providers. One protocol.
Orchestrate Perplexity, Grok, GPT-5.5, Gemini, Qwen, Kimi K2.7-Code, and MiniMax M3 from Claude Code, Claude Desktop, Cursor, or any MCP client.
Get Started · View Tools · Documentation
<br>
If TachiBot helps your workflow, a star goes a long way.
</div>
What's New in v2.26.0
Prompt stack, modernized
refine_prompt(new tool) — opt-in prompt improver on a cheap/fast model: raw query → goal-first brief + what changed + open questions. Never auto-fires, never executes anything — you review, then use the brief. In Claude Code,/prompt refinepresents the open questions as clickable choices and merges your answers into a final brief.- Curated technique list —
list_prompt_techniquesnow defaults to the ~9 core techniques that still help 2026 reasoning models (output contracts likescot,pre_mortem,bdd_spec);all=truefor the full 31. technique="auto"—preview_prompt_techniquerecommends the right technique for your task, with reasons. Asktachi"improve my prompt" for the symptom-based menu.
Setup, de-mystified
tachibot init(new CLI wizard) — detects your API keys and clients, prints the exact config for Claude Code and Claude Desktop. Never writes or echoes keys.- One-click Claude Desktop install — download the
.mcpbfrom the latest release and double-click. No JSON editing. doctor— shows which keys are set, which tools are visible vs hidden and why, and what to try first.
New tools & skills (64 tools · 17 skills)
debug_triage— ranked root-cause hypotheses with the cheapest discriminating check for each (Grok 4.3)spec_writer— loose request → reviewable spec: user stories, Given/When/Then, out-of-scope, open questions (GPT-5.5)diff_review/plan_critique/testgen/security_review— multi-model diff review, adversarial plan red-team, test generation, OWASP/CWE audit- Skills:
/review,/redteam,/spec,/triage,/setup
Fixes
focusorchestration screen: 37 lines of repeated scaffolding → 10 focused linesnpm testexits 0 again (uncancelled race timers leaked past Jest teardown)- GPT-5.5 high-effort reasoning no longer cut off at 3 minutes (timeout 180s → 600s)
Skills (Claude Code)
TachiBot ships with 17 slash commands for Claude Code. These orchestrate the tools into powerful workflows:
| Skill | What it does | Example |
|---|---|---|
/setup |
Guided configuration — runs doctor, walks through keys/profiles | /setup |
/spec |
Request → reviewable spec before planning | /spec add OAuth somehow |
/blueprint |
Multi-model planning → bite-sized TDD steps | /blueprint add OAuth with refresh tokens |
/judge |
Multi-model council - parallel analysis with synthesis | /judge how to implement rate limiting |
/think |
Sequential reasoning chain with any model | /think grok,gemini design a cache layer |
/focus |
Mode-based reasoning (debate, research, analyze) | /focus architecture-debate Redis vs Pg |
/breakdown |
Strategic decomposition with pre-mortem | /breakdown refactor payment module |
/decompose |
Split into sub-problems, deep-dive each one | /decompose implement collaborative editor |
/prompt |
Recommend the right thinking technique (31 available) | /prompt why do users churn |
/algo |
Algorithm analysis with 4 specialized models (DeepSeek lead) | /algo optimize LRU cache O(1) |
/lens |
Long-context analysis over Kimi's 256K window | /lens find inconsistencies in this spec |
/reflect |
Grounded reflexion loop — critique vs external evidence | /reflect harden this auth middleware |
/tot |
Tree-of-Thought: branch → jury-prune → synthesize | /tot design a rate limiter |
/review |
Multi-model diff review — panel + Gemini judge verdict | /review (or paste a diff) |
/redteam |
Adversarial plan red-team — pre-mortem, risks, plan edits | /redteam <paste plan> |
/triage |
Ranked root-cause bug triage | /triage <paste stack trace> |
/tachi |
Help - see available skills, tools, key status | /tachi |
Skills automatically adapt to your configured API keys. Even with just 1-2 providers, all skills work.
Getting started? Type
/tachito see what's available.
Key Features
Multi-Model Intelligence
- 64 AI Tools across 12 providers — Perplexity, Grok, GPT-5, Gemini, Qwen, Kimi, MiniMax, DeepSeek, GLM (Zhipu), StepFun, ERNIE (Baidu), plus free local models (Ollama / LM Studio / llama.cpp / vLLM)
- Gemini 3.5 Flash (
gemini-3.5-flash, GA May 19 2026) — Flash/search tier; reasoning default staysgemini-3.1-pro-preview - Multi-Model Council — planner_maker synthesizes plans from 5+ models into bite-sized TDD steps
- Smart Routing — Automatic model selection for optimal results
- OpenRouter Gateway — Optional single API key for all providers
Advanced Workflows
- YAML-Based Workflows — Multi-step AI processes with dependency graphs
- Prompt Engineering — 31 research-backed techniques (including SCoT, ReAct, Reflexion)
- Verification Checkpoints — 50% / 80% / 100% with automated quality scoring
- Parallel Execution — Run multiple models simultaneously
Tool Profiles
| Profile | Tools | Best For |
|---|---|---|
| Minimal | 13 | Quick tasks, low token budget |
| Research Power | 35 | Deep investigation, multi-source |
| Code Focus | 42 | Software development, SWE tasks |
| Balanced | 53 | General-purpose, mixed workflows |
| Heavy Coding | 57 | Max code tools + agentic workflows |
| Full (default) | 64 | Everything enabled |
Developer Experience
- Claude Code — First-class support
- Claude Desktop — Full integration
- Cursor — Works seamlessly
- TypeScript — Fully typed, extensible
Quick Start
Installation
npm install -g tachibot-mcp
Setup wizard
npx -y -p tachibot-mcp tachibot init
Detects your keys and clients, then prints the exact config for Claude Code and Claude Desktop.
Claude Code (one-liner)
claude mcp add tachibot -- npx -y -p tachibot-mcp tachibot
Then verify with /mcp. Add API keys with --env, e.g. --env OPENROUTER_API_KEY=sk-or-xxx --env PERPLEXITY_API_KEY=pplx-xxx.
Setup (Claude Desktop)
One-click (easiest): download tachibot-mcp.mcpb from the latest release and double-click it — Claude Desktop installs the extension with no JSON editing. Add your API keys when prompted (or later via the extension settings).
Gateway Mode (Recommended) — 2 keys, all providers:
{
"mcpServers": {
"tachibot": {
"command": "tachibot",
"env": {
"OPENROUTER_API_KEY": "sk-or-xxx",
"PERPLEXITY_API_KEY": "pplx-xxx",
"USE_OPENROUTER_GATEWAY": "true"
}
}
}
}
Direct Mode — One key per provider:
{
"mcpServers": {
"tachibot": {
"command": "tachibot",
"env": {
"PERPLEXITY_API_KEY": "your-key",
"GROK_API_KEY": "your-key",
"OPENAI_API_KEY": "your-key",
"GOOGLE_API_KEY": "your-key",
"OPENROUTER_API_KEY": "your-key"
}
}
}
}
Get keys: OpenRouter | Perplexity
See Installation Guide for detailed instructions.
Tool Ecosystem (64 Tools)
Research & Search (5)
perplexity_ask · perplexity_reason · grok_search · openai_search · gemini_search
Reasoning & Planning (14)
grok_reason · openai_reason · qwen_reason · qwq_reason · kimi_thinking · kimi_decompose · deepseek_reason · glm_reason · stepfun_reason · ernie_reason · planner_maker · planner_runner · list_plans · spec_writer
Code Intelligence (11)
kimi_code · grok_code · grok_debug · qwen_coder · qwen_algo · qwen_competitive · deepseek_algo · minimax_code · minimax_agent · testgen · debug_triage
Analysis & Judgment (14)
gemini_analyze_text · gemini_analyze_code · gemini_judge · jury · diff_review · plan_critique · gemini_brainstorm · openai_brainstorm · openai_code_review · openai_explain · grok_brainstorm · grok_architect · security_review · kimi_long_context
Meta & Orchestration (6)
think · nextThought · focus · tachi · doctor · usage_stats
Workflows (9)
workflow · workflow_start · continue_workflow · list_workflows · create_workflow · visualize_workflow · workflow_status · validate_workflow · validate_workflow_file
Prompt Engineering (4)
list_prompt_techniques · preview_prompt_technique · execute_prompt_technique · refine_prompt
Local Models (1)
local_query — any OpenAI-compatible local server (Ollama / LM Studio / llama.cpp / vLLM). Zero-cost, offline, private; also available as the local jury juror (hermes is accepted as a legacy alias). Runs whatever LOCAL_LLM_MODEL points at — e.g. a Nous Hermes build (ollama pull hermes3). Note the Hermes agent itself is model-agnostic — it runs on 300+ backends (GPT, Claude, Gemini, DeepSeek, or self-hosted Ollama/vLLM) — so "Hermes" was never a guarantee of distinct weights.
Advanced Modes (bonus)
- Challenger — Critical analysis with multi-model fact-checking
- Verifier — Multi-model consensus verification
- Scout — Hybrid intelligence gathering
Example Usage
Multi-Model Planning
// Create a plan with multi-model council
planner_maker({ task: "Build a REST API with auth and tests", mode: "start" })
// → Grok searches → Qwen analyzes → Kimi decomposes → GPT critiques → Gemini synthesizes
// Execute with checkpoints
planner_runner({ plan: planContent, mode: "step", stepNum: 1 })
// → Automatic verification at 50%, 80% (kimi_decompose), and 100%
Task Decomposition
kimi_decompose({
task: "Migrate monolith to microservices",
depth: 3,
outputFormat: "dependencies"
})
// → Structured subtasks with IDs, parallel flags, acceptance criteria
Code Review
kimi_code({
task: "review",
code: "function processPayment(amount, card) { ... }",
language: "typescript"
})
// → SWE-Bench 76.8% quality analysis
Deep Reasoning
focus({
query: "Design a scalable event-driven architecture",
mode: "deep-reasoning",
models: ["grok", "gemini", "kimi"],
rounds: 5
})
Documentation
- Full Documentation
- Installation Guide
- Configuration
- Tools Reference
- Workflows Guide
- API Keys Guide
- Focus Modes
Setup Guides
Contributing
Contributions welcome! See CONTRIBUTING.md for guidelines.
<div align="center">
Like what you see?
Star on GitHub — it helps more than you think.
AGPL-3.0 — see LICENSE for details.
Made with care by @byPawel
Multi-model AI orchestration, unified.
</div>
Recommended Servers
playwright-mcp
A Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.
Audiense Insights MCP Server
Enables interaction with Audiense Insights accounts via the Model Context Protocol, facilitating the extraction and analysis of marketing insights and audience data including demographics, behavior, and influencer engagement.
Magic Component Platform (MCP)
An AI-powered tool that generates modern UI components from natural language descriptions, integrating with popular IDEs to streamline UI development workflow.
VeyraX MCP
Single MCP tool to connect all your favorite tools: Gmail, Calendar and 40 more.
graphlit-mcp-server
The Model Context Protocol (MCP) Server enables integration between MCP clients and the Graphlit service. Ingest anything from Slack to Gmail to podcast feeds, in addition to web crawling, into a Graphlit project - and then retrieve relevant contents from the MCP client.
Kagi MCP Server
An MCP server that integrates Kagi search capabilities with Claude AI, enabling Claude to perform real-time web searches when answering questions that require up-to-date information.
Neon Database
MCP server for interacting with Neon Management API and databases
Exa Search
A Model Context Protocol (MCP) server lets AI assistants like Claude use the Exa AI Search API for web searches. This setup allows AI models to get real-time web information in a safe and controlled way.
Qdrant Server
This repository is an example of how to create a MCP server for Qdrant, a vector search engine.
E2B
Using MCP to run code via e2b.