Command Palette

Search for a command to run...

UnylyUnyly
Весь каталог

Tachibot

БесплатноНе проверен

Multi-model orchestration running OpenAI, Google, Anthropic, xAI, Perplexity, and OpenRouter in parallel for cross-validation

GitHubEmbed

Описание

Multi-model orchestration running OpenAI, Google, Anthropic, xAI, Perplexity, and OpenRouter in parallel for cross-validation

README

TachiBot MCP

Multi-Model AI Orchestration Platform

Version Tools License Node MCP

65 AI tools. 12 providers. One protocol.

Orchestrate Perplexity, Grok 4.5, GPT-5.6, Gemini, Qwen, Kimi K3, and MiniMax M3 from Claude Code, Claude Desktop, Cursor, or any MCP client.

Get Started · View Tools · Documentation


If TachiBot helps your workflow, a star goes a long way.

GitHub stars npm downloads


What's New

v2.28.0

  • Qwen3.8 Max (Aug 3, qwen/qwen3.8-max) — Alibaba's new flagship, GA the same day, now powers qwen_algo, qwen_reason, and the qwen_reason juror. 1M context (up from 262K), multimodal (text+image+video in), and the first Qwen exposing configurable reasoning effort. $2/$6 per M.
  • Measured on a range-query algorithm problem, 3.8 Max was the only one of five Qwen candidates to surface both the offline-vs-online tradeoff and the strict-inequality-with-duplicates edge case. The outgoing qwen3-235b-a22b-thinking-2507 was correct but shallow; qwen3.7-max cost 1.6x for twice the wall time; qwen3-max-thinking was rejected outright — it returns zero reasoning tokens.
  • Effort is pinned to medium, deliberately. This model's default effort behaves like high: 302s and $0.09 on a single qwen_algo call. At medium the same call answers at equal depth in 18–48s for $0.006–0.018 — cheaper and 3.5x faster than the model it replaces, which took 169s and $0.036 for a shorter answer. low starts dropping alternatives and is not used.
  • reasoning_effort pass-through for OpenRoutercallOpenRouter now forwards the parameter; OpenRouter drops it for models that don't list it, so the quota fallback chain (3.8 Max → 3.7 Max → 235B Thinking) stays safe. Qwen3.8/3.7-Max also join the 600s extended-timeout bucket.
  • qwen_coder, qwen_competitive, testgen stay on Qwen3-Coder-Next — it is coding-specialized and ~16x cheaper ($0.12/$0.80); 3.8 Max is the reasoning tier, not the codegen tier.

v2.27.1

  • Kimi K3 (Jul 16, moonshotai/kimi-k3) now powers every Kimi tool, the kimi juror, and the Kimi seat on diff_review. 2.8T open-weight MoE — the largest open model shipped — with a 1M context (up from 262K), native multimodal input, and long-horizon agentic coding that beats Opus 4.8 and GPT-5.5 on coding/agent benchmarks. Note the price: $3/$15 per M, 4x K2.7-Code — K2.7-Code stays as the automatic fallback (K3 → K2.7-Code → K2.6).
  • Gemini 3.6 Flash (Jul 21, gemini-3.6-flash) is the new search/workhorse tier behind gemini_search — 1M context at $1.50/$7.50, down from $9 output, and ~17% fewer output tokens than 3.5 Flash. Flash-Lite moves to gemini-3.5-flash-lite ($0.30/$2.50).
  • Everything else re-verified, unchanged — a full provider audit (Jul 26) confirms GPT-5.6, Grok 4.5, DeepSeek V4 Pro, GLM-5.2, MiniMax M3, StepFun 3.7, ERNIE 4.5 VL, Qwen Coder-Next and Perplexity Sonar are all still the newest available. Gemini 3.1 Pro remains the reasoning/judge default: Google shipped three models on Jul 21 and skipped 3.5 Pro. xAI's staged rollout finished (EU access Jul 17), so the grok-4.3 fallback is now quota/region insurance rather than a rollout workaround.

v2.27.0

  • Grok 4.5 (Jul 8 flagship, "Opus-class") now powers all Grok tools, with automatic fallback to grok-4.3 while xAI's region-staged rollout completes (EU mid-July) — tools keep working everywhere, and 4.5 activates by itself.
  • grok_search_lite (new tool, 65 total) — the same Grok live search on grok-4-1-fast ($0.20/$0.50, 2M ctx), ~10x cheaper than grok_search. Use it for high-volume lookups and jury/council fan-outs.
  • GPT-5.6 tiersopenai_* tools move to gpt-5.6-sol (flagship, same $5/$30 as 5.5 but stronger), terra for code (5.5-level at half price), luna for explanations ($1/$6). The $30/$180 gpt-5.5-pro tier is replaced by sol + reasoning effort; a permission fallback (sol → terra → 5.5) covers org-gated accounts.

v2.26.1

  • /test and /audit skills (19 skills total) — /test generates runnable tests via testgen; /audit runs an OWASP/CWE security review via security_review.
  • Skill install in the wizardtachibot init now offers to install Claude Code skills with a per-skill skip choice ([Enter]=all · [s]=choose which to skip · [n]=none). Skills are opt-in — postinstall no longer writes to ~/.claude silently (npm run install-skills still installs all non-interactively).
  • Fixes — the one-click .mcpb extension now points at a valid entry point (was broken) and tracks the package version; tachibot init exits cleanly on non-interactive/CI stdin instead of hanging.

v2.26.0

Prompt stack, modernized

  • refine_prompt (new tool) — opt-in prompt improver on a cheap/fast model: raw query → goal-first brief + what changed + open questions. Never auto-fires, never executes anything — you review, then use the brief. In Claude Code, /prompt refine presents the open questions as clickable choices and merges your answers into a final brief.
  • Curated technique listlist_prompt_techniques now defaults to the ~9 core techniques that still help 2026 reasoning models (output contracts like scot, pre_mortem, bdd_spec); all=true for the full 31.
  • technique="auto"preview_prompt_technique recommends the right technique for your task, with reasons. Ask tachi "improve my prompt" for the symptom-based menu.

Setup, de-mystified

  • tachibot init (new CLI wizard) — detects your API keys and clients, prints the exact config for Claude Code and Claude Desktop. Never writes or echoes keys.
  • One-click Claude Desktop install — download the .mcpb from the latest release and double-click. No JSON editing.
  • doctor — shows which keys are set, which tools are visible vs hidden and why, and what to try first.

New tools & skills (65 tools · 19 skills)

  • debug_triage — ranked root-cause hypotheses with the cheapest discriminating check for each (Grok 4.3)
  • spec_writer — loose request → reviewable spec: user stories, Given/When/Then, out-of-scope, open questions (GPT-5.5)
  • diff_review / plan_critique / testgen / security_review — multi-model diff review, adversarial plan red-team, test generation, OWASP/CWE audit
  • Skills: /review, /redteam, /spec, /triage, /setup

Fixes

  • focus orchestration screen: 37 lines of repeated scaffolding → 10 focused lines
  • npm test exits 0 again (uncancelled race timers leaked past Jest teardown)
  • GPT-5.5 high-effort reasoning no longer cut off at 3 minutes (timeout 180s → 600s)

Skills (Claude Code)

TachiBot ships with 19 slash commands for Claude Code. These orchestrate the tools into powerful workflows:

Skill What it does Example
/setup Guided configuration — runs doctor, walks through keys/profiles /setup
/spec Request → reviewable spec before planning /spec add OAuth somehow
/blueprint Multi-model planning → bite-sized TDD steps /blueprint add OAuth with refresh tokens
/judge Multi-model council - parallel analysis with synthesis /judge how to implement rate limiting
/think Sequential reasoning chain with any model /think grok,gemini design a cache layer
/focus Mode-based reasoning (debate, research, analyze) /focus architecture-debate Redis vs Pg
/breakdown Strategic decomposition with pre-mortem /breakdown refactor payment module
/decompose Split into sub-problems, deep-dive each one /decompose implement collaborative editor
/prompt Recommend the right thinking technique (31 available) /prompt why do users churn
/algo Algorithm analysis with 4 specialized models (DeepSeek lead) /algo optimize LRU cache O(1)
/lens Long-context analysis over Kimi's 1M window /lens find inconsistencies in this spec
/reflect Grounded reflexion loop — critique vs external evidence /reflect harden this auth middleware
/tot Tree-of-Thought: branch → jury-prune → synthesize /tot design a rate limiter
/review Multi-model diff review — panel + Gemini judge verdict /review (or paste a diff)
/redteam Adversarial plan red-team — pre-mortem, risks, plan edits /redteam <paste plan>
/triage Ranked root-cause bug triage /triage <paste stack trace>
/test Generate runnable tests (edge cases first) /test src/auth.ts
/audit Security review — OWASP/CWE findings + fixes /audit the login handler
/tachi Help - see available skills, tools, key status /tachi

Skills automatically adapt to your configured API keys. Even with just 1-2 providers, all skills work.

Getting started? Type /tachi to see what's available.


Key Features

Multi-Model Intelligence

  • 65 AI Tools across 12 providers — Perplexity, Grok, GPT-5, Gemini, Qwen, Kimi, MiniMax, DeepSeek, GLM (Zhipu), StepFun, ERNIE (Baidu), plus free local models (Ollama / LM Studio / llama.cpp / vLLM)
  • Gemini 3.6 Flash (gemini-3.6-flash, GA Jul 21 2026) — Flash/search tier; reasoning default stays gemini-3.1-pro-preview (Google has still not shipped a 3.5 Pro)
  • Multi-Model Council — planner_maker synthesizes plans from 5+ models into bite-sized TDD steps
  • Smart Routing — Automatic model selection for optimal results
  • OpenRouter Gateway — Optional single API key for all providers

Advanced Workflows

  • YAML-Based Workflows — Multi-step AI processes with dependency graphs
  • Prompt Engineering — 31 research-backed techniques (including SCoT, ReAct, Reflexion)
  • Verification Checkpoints — 50% / 80% / 100% with automated quality scoring
  • Parallel Execution — Run multiple models simultaneously

Tool Profiles

Profile Tools Best For
Minimal 13 Quick tasks, low token budget
Research Power 36 Deep investigation, multi-source
Code Focus 42 Software development, SWE tasks
Balanced 54 General-purpose, mixed workflows
Heavy Coding 58 Max code tools + agentic workflows
Full (default) 65 Everything enabled

Developer Experience

  • Claude Code — First-class support
  • Claude Desktop — Full integration
  • Cursor — Works seamlessly
  • TypeScript — Fully typed, extensible

Quick Start

Installation

npm install -g tachibot-mcp

Setup wizard

npx -y -p tachibot-mcp tachibot init

Detects your keys and clients, then prints the exact config for Claude Code and Claude Desktop.

Claude Code (one-liner)

claude mcp add tachibot -- npx -y -p tachibot-mcp tachibot

Then verify with /mcp. Add API keys with --env, e.g. --env OPENROUTER_API_KEY=sk-or-xxx --env PERPLEXITY_API_KEY=pplx-xxx.

Setup (Claude Desktop)

One-click (easiest): download tachibot-mcp.mcpb from the latest release and double-click it — Claude Desktop installs the extension with no JSON editing. Add your API keys when prompted (or later via the extension settings).

Gateway Mode (Recommended) — 2 keys, all providers:

{
  "mcpServers": {
    "tachibot": {
      "command": "tachibot",
      "env": {
        "OPENROUTER_API_KEY": "sk-or-xxx",
        "PERPLEXITY_API_KEY": "pplx-xxx",
        "USE_OPENROUTER_GATEWAY": "true"
      }
    }
  }
}

Direct Mode — One key per provider:

{
  "mcpServers": {
    "tachibot": {
      "command": "tachibot",
      "env": {
        "PERPLEXITY_API_KEY": "your-key",
        "GROK_API_KEY": "your-key",
        "OPENAI_API_KEY": "your-key",
        "GOOGLE_API_KEY": "your-key",
        "OPENROUTER_API_KEY": "your-key"
      }
    }
  }
}

Get keys: OpenRouter | Perplexity

See Installation Guide for detailed instructions.


Tool Ecosystem (65 Tools)

Research & Search (6)

perplexity_ask · perplexity_reason · grok_search · grok_search_lite · openai_search · gemini_search

Reasoning & Planning (14)

grok_reason · openai_reason · qwen_reason · qwq_reason · kimi_thinking · kimi_decompose · deepseek_reason · glm_reason · stepfun_reason · ernie_reason · planner_maker · planner_runner · list_plans · spec_writer

Code Intelligence (11)

kimi_code · grok_code · grok_debug · qwen_coder · qwen_algo · qwen_competitive · deepseek_algo · minimax_code · minimax_agent · testgen · debug_triage

Analysis & Judgment (14)

gemini_analyze_text · gemini_analyze_code · gemini_judge · jury · diff_review · plan_critique · gemini_brainstorm · openai_brainstorm · openai_code_review · openai_explain · grok_brainstorm · grok_architect · security_review · kimi_long_context

Meta & Orchestration (6)

think · nextThought · focus · tachi · doctor · usage_stats

Workflows (9)

workflow · workflow_start · continue_workflow · list_workflows · create_workflow · visualize_workflow · workflow_status · validate_workflow · validate_workflow_file

Prompt Engineering (4)

list_prompt_techniques · preview_prompt_technique · execute_prompt_technique · refine_prompt

Local Models (1)

local_query — any OpenAI-compatible local server (Ollama / LM Studio / llama.cpp / vLLM). Zero-cost, offline, private; also available as the local jury juror (hermes is accepted as a legacy alias). Runs whatever LOCAL_LLM_MODEL points at — e.g. a Nous Hermes build (ollama pull hermes3). Note the Hermes agent itself is model-agnostic — it runs on 300+ backends (GPT, Claude, Gemini, DeepSeek, or self-hosted Ollama/vLLM) — so "Hermes" was never a guarantee of distinct weights.

Advanced Modes (bonus)

  • Challenger — Critical analysis with multi-model fact-checking
  • Verifier — Multi-model consensus verification
  • Scout — Hybrid intelligence gathering

Example Usage

Multi-Model Planning

// Create a plan with multi-model council
planner_maker({ task: "Build a REST API with auth and tests", mode: "start" })
// → Grok searches → Qwen analyzes → Kimi decomposes → GPT critiques → Gemini synthesizes

// Execute with checkpoints
planner_runner({ plan: planContent, mode: "step", stepNum: 1 })
// → Automatic verification at 50%, 80% (kimi_decompose), and 100%

Task Decomposition

kimi_decompose({
  task: "Migrate monolith to microservices",
  depth: 3,
  outputFormat: "dependencies"
})
// → Structured subtasks with IDs, parallel flags, acceptance criteria

Code Review

kimi_code({
  task: "review",
  code: "function processPayment(amount, card) { ... }",
  language: "typescript"
})
// → SWE-Bench 76.8% quality analysis

Deep Reasoning

focus({
  query: "Design a scalable event-driven architecture",
  mode: "deep-reasoning",
  models: ["grok", "gemini", "kimi"],
  rounds: 5
})

Documentation

Setup Guides


Contributing

Contributions welcome! See CONTRIBUTING.md for guidelines.


Like what you see?

Star on GitHub — it helps more than you think.

GitHub stars

Website · Docs · npm · Issues

AGPL-3.0 — see LICENSE for details.

Made with care by @byPawel

Multi-model AI orchestration, unified.

from github.com/bypawel/tachibot-mcp

Установить Tachibot в Claude Desktop, Claude Code, Cursor

Рекомендуется · одна команда, все IDE
unyly install tachibot

Ставит в Claude Desktop, Claude Code, Cursor и VS Code — сам разбирается с npx, uvx и сборкой из исходников.

Впервые? Поставь CLI: curl -fsSL https://unyly.org/install | sh

Или настроить вручную

Выполни в терминале:

claude mcp add tachibot -- npx -y tachibot-mcp

Пошаговые гайды: как установить Tachibot

FAQ

Tachibot MCP бесплатный?

Да, Tachibot MCP бесплатный — установка в пару кликов через Unyly без оплаты.

Нужен ли API-ключ для Tachibot?

Нет, Tachibot работает без API-ключей и переменных окружения.

Tachibot — hosted или self-hosted?

Self-hosted: сервер запускается локально на твоей машине командой из раздела установки.

Как установить Tachibot в Claude Desktop, Claude Code или Cursor?

Открой Tachibot на unyly.org, выбери вкладку своего клиента (Claude Desktop, Claude Code, Cursor) и нажми Install — конфиг сгенерируется автоматически, без правки JSON.

Похожие MCP

Compare Tachibot with

Не уверен что выбрать?

Найди свой стек за 60 секунд

Автор?

Embed-бейдж для README

Похожее

Все в категории ai