Skip to content
The AI of MonkeysCode

Meet Capuchin.

Capuchin is the AI that powers MonkeysCode — a coding model built for one thing: shipping software. It plans, writes, and verifies across your whole codebase, thinks hard when the problem is hard, and stays instant when it isn't. It's the default the moment you open the editor — and you're never locked to it.

Included on every paid plan. Runs in the cloud, on your server, or on your machine.

Fast + Reasoning· two modes, one model
Agent-native· built for tools, not chat
Long context· whole services at once
Runs anywhere· cloud · your server · local
Built for software, not small talk

A model that does one job, exceptionally.

General chat models spread their capacity across everything. Capuchin spends it all on software engineering — reading large codebases, calling tools, editing across files, and checking its own work against your tests. That focus is why it's both sharper on real coding and lean enough to run fast and cheap, wherever you put it.

Capuchin Flash and Capuchin Reason.

One model family, two profiles — so you never trade speed for depth or depth for speed.

Capuchin Flash

instant

Best for: autocomplete, apply, quick edits

Feel: sub-second

Where it shines: the 80% of everyday work

Capuchin Reason

deliberate

Best for: planning, multi-file agents, debugging

Feel: step-by-step reasoning

Where it shines: the 20% that's actually hard

MonkeysCode routes between them automatically — or you can pin one.

What Capuchin does.

Two minds in one model

Instant when it's easy. Deliberate when it's hard.

Flash mode answers and edits land immediately for the everyday work — completions, small changes, quick questions.

Reason mode for the hard stuff, Capuchin thinks step by step before it acts — planning multi-file changes, untangling bugs.

Automatic routing MonkeysCode picks the mode by the task, so you get speed by default and depth on demand.

Budgeted thinking reasoning is capped per task, so 'thinking' never means 'spinning.'

FlashReasonAuto-routingBudgeted

Two minds in one model

Instant when it's easy. Deliberate when it's hard.

Agent-native

Built to use tools, not just talk about them.

Reliable tool use Capuchin calls your terminal, git, test runner, browser, and MCP servers — accurately, call after call.

Long-horizon stability it holds the thread across dozens of tool calls and many turns.

Plans then executes capuchin produces a plan you can approve, then carries it out across files in parallel.

Self-checking it runs your tests and fixes what it broke before handing back.

Tool useMulti-stepParallelSelf-correcting

Agent-native

Built to use tools, not just talk about them.

Sees the whole picture

Reads your codebase, not just the open file.

Massive context window hundreds of thousands of tokens, so whole services and long histories fit at once.

Repo-aware paired with MonkeysCode's on-device index and dependency graph, it reasons across files, services, and repos.

Keeps the thread long agent runs are summarized and carried forward without losing important context.

Long contextRepo-awareCross-service

Sees the whole picture

Reads your codebase, not just the open file.

Polyglot

Fluent across your stack — not just one language.

Many languages, first-class strong across PHP, Python, Go, Rust, JavaScript/TypeScript, Java, and more.

Built for MonkeysLegion extra-sharp on PHP 8.4 and the MonkeysLegion framework, where most tools are weakest.

Real-world tasks tuned on the kind of multi-file, multi-tool work engineers actually do.

PHPPythonGoRustTSJava

Polyglot

Fluent across your stack — not just one language.

Big model, lean runtime

Frontier-class capability at a cost you can serve.

Efficient by design Capuchin delivers heavyweight results while activating only a fraction of its capacity per token.

Cheap at the margin that efficiency is why we can include generous Capuchin usage on every plan.

Optimized serving speculative decoding and caching keep responses quick even under load.

EfficientFastIncluded usage

Big model, lean runtime

Frontier-class capability at a cost you can serve.

Your AI, your perimeter

Run Capuchin wherever your rules require.

Because Capuchin is built on an open foundation, it can live where your compliance team needs it to — something a closed vendor model can never offer.

Managed cloud

Included on your plan

Your server

Inside your infrastructure

Your machine

Fully air-gapped

Air-gapped

Zero data leaves

Always improving

Capuchin gets better every release.

Specialized by us

Tuned relentlessly for real software engineering inside MonkeysCode.

Learns from real work

With your permission, Capuchin improves from opt-in signals — never from private code.

Ships often

New checkpoints roll out on a steady cadence, gated against our benchmarks.

Open where it counts

Your data stays yours; improvement is consent-first by design.

The default that respects your choice.

Capuchin is what powers MonkeysCode out of the box. But the moment you want a different model — Claude, Gemini, ChatGPT, or your own — it's one switch away. Capuchin is the AI we stand behind; model freedom is the promise underneath it.

Download MonkeysCode
Included on every paid plan

Code with Capuchin.

MonkeysCode's own AI — fast, reasoning, agent-native, and yours to run anywhere.