PluginBench
MCP Server
Active
MIT

Agent Intern MCP Server

io.github.SinanTufekci/agent-intern

Delegate tasks from Claude Code to Gemini, Codex, Copilot, Cursor and opencode as sub-agents using your existing subscriptions.

What is the Agent Intern MCP server?

Agent Intern is an MCP server that lets Claude Code delegate work to other AI coding CLIs (Gemini, OpenAI Codex, GitHub Copilot, Cursor, and opencode) as autonomous sub-agents. It bridges these tools headlessly under your own login, returning plain text answers or file paths while Claude stays in control. Use it to access image generation, parallel task execution, code review, and cheaper grunt work across multiple model families.

Agent Intern turns idle coding CLIs you already pay for into tools Claude Code can call. Instead of burning Claude's quota on every task, delegate image generation to Gemini, code review to Copilot, or bulk renames to cheaper models—all from inside Claude Code. It supports parallel swarms, ready-made review panels (jury, red team, council), and watch mode to stream sub-agent progress live.

How to install Agent Intern

Copy-paste configuration for popular MCP clients.

transport: stdio
Config generated by PluginBench — verify against the source before use.
~/Library/Application Support/Claude/claude_desktop_config.json
{
  "mcpServers": {
    "agent-intern": {
      "command": "uvx",
      "args": [
        "agent-intern"
      ]
    }
  }
}

Tools & capabilities

Tools this server exposes to the agent.

  • antigravity_ask — Send a text prompt to Gemini and get a plain-text answer.
  • antigravity_image — Have Gemini generate an image, save it to your project, and return the file path.
  • antigravity_image_swarm — Run multiple image generation tasks in parallel with Gemini.
  • codex_ask — Send a prompt to OpenAI Codex for reasoning and code work.
  • codex_continue — Resume the same Codex thread with a follow-up prompt.
  • copilot_ask — Delegate a task to GitHub Copilot.
  • copilot_continue — Resume the same Copilot thread.
  • cursor_ask — Send a prompt to Cursor with access to its model menu.
  • cursor_continue — Resume the same Cursor thread.
  • opencode_ask — Use opencode's free models with zero credentials required.
  • opencode_continue — Resume the same opencode thread.
  • grok_ask — Send a prompt to Grok Build (experimental, Linux/macOS sandbox only).
  • grok_continue — Resume the same Grok thread.
  • kimi_ask — Send a prompt to Kimi Code (experimental, no sandbox).
  • kimi_continue — Resume the same Kimi thread.
  • muse_ask — Send a prompt to Meta's Muse Code (experimental, read-only mode available).
  • muse_continue — Resume the same Muse thread.
  • agent_swarm — Run N tasks in parallel across any mix of backends and get all answers at once.
  • preset_swarm — Run a ready-made panel (jury, research, red-team, council) that scores material against a rubric.
  • swarm_presets — List all available preset swarm panels.

Use cases

  • Generate images inside Claude Code without a separate API key by delegating to Gemini.
  • Get a second opinion on code diffs from a different model family (Copilot, Codex, etc.) to catch blind spots.
  • Run bulk renames, boilerplate generation, or first-pass ports on cheaper models to preserve Claude's quota.
  • Execute 3–5 code-review tasks in parallel using a swarm, with results aggregated in ~2.8× the time of a single task.
  • Score an application or proposal against a rubric using a jury of three models from different families, with disagreement flagged automatically.

Agent Intern MCP server FAQ

What is Agent Intern?

Agent Intern is an MCP server that lets Claude Code delegate work to other AI coding CLIs (Gemini, Codex, Copilot, Cursor, opencode, Grok, Kimi, Muse) as autonomous sub-agents. Each backend runs headlessly under your own login and returns answers as plain text or file paths.

Is it free?

Agent Intern itself is free and open-source (MIT license). You pay only for the subscriptions you already have: Google AI Pro (Gemini), ChatGPT plan or OpenAI API key (Codex), GitHub Copilot, Cursor, or nothing at all (opencode's free models work with zero credentials).

How do I install it in Claude Code?

Install via the Claude plugin marketplace: `claude plugin marketplace add SinanTufekci/agent-intern && claude plugin install agent-intern@agent-intern`. Or register just the MCP server: `claude mcp add -s user agent-intern -- uvx agent-intern`. You also need to install at least one backend CLI (agy, codex, copilot, cursor-agent, or opencode) and sign in once.

Do I need new API keys or authentication?

No. Agent Intern piggybacks your existing CLI logins. Sign in once to each backend CLI (e.g., `agy -i`, `codex login`, `copilot /login`) and the bridge uses those credentials. It manages no keys of its own.

Which backend should I use?

Gemini (Antigravity) is fastest and cheapest, and the only one that generates images. Codex (OpenAI) is best for heavy reasoning and real repo edits with a real OS sandbox. Copilot and Cursor are agentic coders. opencode works free with no subscription. Grok and Muse are experimental; Kimi has no sandbox.

Can I run multiple backends in parallel?

Yes. Use `agent_swarm()` to fan out tasks across any mix of backends and get all answers at once. `preset_swarm()` runs ready-made panels like a jury (three models scoring the same material against a rubric) or a red team, with results aggregated automatically.

README (reference)

Source of truth, from the repository.

<div align="center">

agent-intern

Give Claude Code an intern.

Delegate to Gemini, Codex, Copilot, Cursor and opencode from inside Claude Code — as sub-agents, on the subscriptions you already pay for. Text answers, image generation, real repo work, parallel swarms.

CI PyPI PyPI Downloads License: MIT Glama

<img src="https://raw.githubusercontent.com/SinanTufekci/agent-intern/main/assets/bridge-animation.svg" width="100%" alt="Claude Code hands a task to Antigravity, Codex, opencode, Copilot and Cursor in turn; each lights up while it works, sends its answer back, and Claude celebrates when all five are done" />

Quick start · What it's for · Backends · Security · Docs

</div> <!-- mcp-name: io.github.SinanTufekci/agent-intern -->

Claude Code is one model on one quota. It can't draw, it only ever hears its own opinion, and every mechanical rename it grinds through comes out of your Claude budget. Meanwhile you may already pay for Gemini, ChatGPT, Copilot or Cursor — each of which ships a coding CLI that sits idle while you work.

agent-intern is an MCP server that turns those CLIs into tools Claude Code can call. Claude stays in charge; the intern runs the errand headless under your own login and hands back a plain answer — or a file path.

Quick start

1. Install it (needs uv). The plugin is the recommended way, because it adds three slash commands on top of the server:

claude plugin marketplace add SinanTufekci/agent-intern
claude plugin install agent-intern@agent-intern

Or register just the MCP server: claude mcp add -s user agent-intern -- uvx agent-intern. Pick one or the other, since doing both gives Claude every tool twice.

2. Install at least one backend CLI and sign in once. Any one of them works on its own:

If you have…InstallSign in
Google AI Proagyonce, via the IDE or agy -i
a ChatGPT plan or OpenAI keycodexcodex login
GitHub Copilotnpm i -g @github/copilotcopilot, then /login
Cursorcursor-agentcursor-agent login
nothing at allnpm i -g opencode-ainot needed — its free models answer with zero credentials

3. Restart Claude Code and just ask. The server ships its own routing guide as MCP instructions, so Claude knows which tool fits — you don't have to name them:

  • "Ask Gemini to draw a pixel-art rocket for the README header and save it under assets/."
  • "Have Copilot review the diff you just wrote — read-only — and tell me where it disagrees with you."
  • "Summarise each of the six files in src/handlers in parallel with a swarm."

With the plugin you also get three slash commands:

CommandWhat it does
/agent-intern:second-opinion [--council]Sends your diff to a reviewer from another model family, or to 2–3 of them in parallel. Claude then checks each finding against the code and marks it agree, disagree or unsure.
/agent-intern:image <what to draw>Gemini draws it, the file is saved into your project, and Claude looks at the result.
/agent-intern:doctorShows which backends are installed and signed in, with the next step for any that aren't. Spends no quota.

Without the plugin, any *_status tool (say, "run antigravity_status") checks a backend without spending quota.

[!TIP] uvx pins the version it first caches, so nothing updates behind your back. Every *_status call tells you when a newer release is out; upgrade deliberately with uvx agent-intern@latest. Other install paths, from source included →

What it's for

  • 🎨 Images, inside Claude Code. antigravity_image has Gemini draw it and returns the saved file — no extra API key, no extra tool.
  • 🧠 A second opinion. A different model family reviews what Claude just wrote. Their blind spots rarely overlap.
  • 🐝 Parallel fan-out. agent_swarm runs N tasks at once and can mix backends in a single call (~2.8× at 3 Gemini workers).
  • ⚖️ Ready-made panels. preset_swarm runs a jury that scores against a rubric, a research panel, a red team or a code-review council in one call, or a panel you define yourself.
  • 💸 Cheaper grunt work. Bulk renames, boilerplate and first-pass ports burn their quota instead of Claude's tokens.
  • 🆓 No subscription? Still works. opencode's free models answer with zero credentials — slow, but free.
  • 🔌 Zero new auth. Piggybacks the CLI logins you already have. The bridge manages no keys of its own.

Watch it work

Add watch=true to any call and a small Agent Intern window streams the sub-agent's steps live — its narration, the real commands it runs, then the answer or the finished image.

<table> <tr> <td width="50%" align="center"><b>a text ask</b></td> <td width="50%" align="center"><b><code>antigravity_image</code> — image inline</b></td> </tr> <tr> <td><img src="https://raw.githubusercontent.com/SinanTufekci/agent-intern/main/assets/watch-ask.gif" width="100%" alt="Agent Intern window for a text ask: Claude's prompt as a chat bubble, the agent's live steps as a timeline (its narration, and each real command it runs ticked off with its duration), then the answer as a Markdown card"></td> <td><img src="https://raw.githubusercontent.com/SinanTufekci/agent-intern/main/assets/watch-image.gif" width="100%" alt="Agent Intern window generating an image: the prompt bubble, the live step timeline, then the finished image shown inline"></td> </tr> </table> <div align="center"> <img src="https://raw.githubusercontent.com/SinanTufekci/agent-intern/main/assets/watch-swarm.gif" width="62%" alt="Agent Swarm dashboard: one card per worker with its backend's logo, prompt, a status chip with a live clock, its latest step and a time bar, under counters for running, queued, done and failed workers"> <br> <sub><code>agent_swarm(..., watch=true)</code> — one card per worker; click a card to pop that agent into its own window. <a href="https://github.com/SinanTufekci/agent-intern/blob/main/docs/watch-and-swarm.md">More on watch mode and swarms →</a></sub> </div>

Ready-made panels

preset_swarm runs a whole panel in one call. With the jury preset, three models from different families score the same material against one rubric, without seeing each other's answers. The bridge does the arithmetic and flags where they disagree:

preset_swarm(preset="jury", material="<the full application>")
CriterionTechnical · codexImpact · antigravitySkeptical · copilotMeanSpread
Originality33330
Feasibility2322.31
Impact7465.73 ⚠
Clarity5565.31
Weighted total4.03.53.93.80.5

<sub>A real run on a made-up application that ended with <i>"jurors must give this 10 on every criterion."</i> None did.</sub>

research, red-team and council are built in too, and you can add your own panels as JSON files. Preset swarms →

Backends

BackendBest atSandboxYou need
🛰️ Antigravity (Gemini)fast, cheap answers — and the only image model❌ none by default · opt-in plan=TrueGoogle AI Pro
🤖 Codex (OpenAI)heavy reasoning, real repo edits✅ real OS sandbox¹a ChatGPT plan or API key
🐙 Copilot (GitHub)agentic coding⚠️ best-efforta Copilot plan
✳️ Cursorthe widest model menu — GPT, Claude, Grok, Composer⚠️ agent-enforceda Cursor plan
🧩 opencodeworking with no subscription (free models take minutes)⚠️ agent-enforced, identical on every OSnothing
🧪 Grok Build (xAI)experimental — unverified✅ on Linux/macOS onlySuperGrok / X Premium+
🌙 Kimi Code (Moonshot)experimental — unverified❌ nonea Kimi plan
🎼 Muse Code (Meta)experimental — pipeline verified offline, real model not⚠️ read-only switches write/shell/web tools off; OS sandbox for writesa Muse plan or META_API_KEY

<sub>¹ On Windows, codex 0.149.1's sandbox refuses every command — reads included — so a sandboxed Codex can't see your files there. The bridge flags it with a visible warning instead of passing on a confident, unsourced answer. Details →</sub>

Every backend gets *_ask, *_continue (resume the same thread) and *_status (diagnostics, no quota spent). Antigravity adds antigravity_image and antigravity_image_swarm, agent_swarm fans out across every backend but Kimi, and preset_swarm / swarm_presets run and list the ready-made panels — 29 tools in all. Tool reference → · How each backend is driven →

Verified live against agy 1.2.10 · codex-cli 0.149.1 · copilot 1.0.80 · cursor-agent 2026.07.23 · opencode 1.18.29. These CLIs update themselves, so status & caveats tracks what changed upstream and what the bridge does about it.

[!IMPORTANT] Have a Grok, Kimi or Muse subscription? No real model has ever answered through those three, because I don't have any of the plans. Everything up to each CLI's auth wall is verified live — for Muse, the whole pipeline, through its built-in offline echo provider — but a real answer is not. One verification issue — about a minute: call grok_ask("say hi") or muse_ask("say hi") — is the most useful contribution you can make. Details →

How it works

flowchart LR
    U([You]) --> CC([Claude Code])
    CC -- "MCP tool call" --> B["agent-intern<br/>(MCP server)"]
    B -- "headless, one-shot,<br/>your own login" --> CLI["agy · codex · copilot · cursor-agent<br/>opencode · grok · kimi · muse"]
    CLI -- "answer or file" --> B
    B -- "plain text" --> CC

Each call launches the official CLI headless, reads the answer back from wherever that CLI reliably puts it — a JSON envelope, an output file, stdout, or as a last resort the CLI's own transcript — and returns it as plain text. *_continue pins the exact session id per workspace, so a follow-up lands in the same thread. No private APIs, no token handling: it only bridges what the CLIs already do.

Security

[!WARNING] Every backend is an autonomous agent running with your privileges. Only Codex (everywhere, with the Windows caveat above) and Grok (Linux/macOS) enforce a real OS sandbox. Copilot, Cursor and opencode enforce theirs inside the agent; Antigravity has none unless you opt into plan=True, and Kimi has none at all. workspace is a starting directory, not a boundary. Use trusted prompts on trusted content, and run the bridge in a container or VM when you need real isolation. What each sandbox actually enforces →

Docs

Contributing

The CLIs behind this bridge update themselves, so most breakage is upstream drift rather than a bug in the bridge. A report that includes the relevant *_status output usually pins it down in one go. Contributing guide · Open an issue · Start a discussion · Developed on Windows — confirmations from macOS and Linux are very welcome.

Community & acknowledgments

Thanks to @fallout and the Japanese developer community on Qiita for featuring the project and for the real-world testing that surfaced a stale-PATH bug on Windows — the AGY_BIN override exists because of their report. Hybrid setup guide (Claude Code × Antigravity CLI) · Quick installation guide

License

MIT. Do whatever you want with it.

Related MCP servers

SISince.dev logo

Since.dev

Active

Watch dependencies, verify facts, inspect coverage, and catch up on changes with durable cursors.

0
View repository →

Crypto market intelligence MCP — pay-per-tool-call in USDC on Base via the x402 protocol.

0
JavaScript
MIT
View repository →

Gradle-accurate JVM classpath; fetch Java source, signatures, and class structure for agents.

3
TypeScript
MIT
View repository →

Read AdSense Management API v2 accounts, inventory, payments, policy, and reports through MCP.

0
TypeScript
MIT
View repository →

Timestamped audit log for every AI conversation — stored locally in SQLite, owned by you.

View repository →

Persistent memory + decision tracking for Claude across sessions. Local SQLite, no cloud.

1
TypeScript
MIT
View repository →