The economical
coding agent.

Does 5× more work for the same tokens. Your subscription lasts longer. Free-tier models actually work.

╶──╮ ╶──┤ 3code v0.6.0 the economical coding agent ╶──╯ provider zaicode model glm-5.2 type a prompt. :help for commands. :q or Ctrl-D to exit.  

Never stop coding

Every other agent is incentivized to burn tokens. 3code is built to save them — aggressive caching, context compaction, chunked mode that discards irrelevant data instead of carrying it forever. The result: your API balance lasts 5× longer, and you can use free-tier models for real work.

Get started

1

Get a provider

3code also supports moonshot's Kimi (both subscriptions and api), Deepseek, OpenRouter, and many others.

More providers

2

Install 3code and enter your API key

$ curl -fsSL https://3code.capocasa.dev/install | sh
$ curl -fsSL https://3code.capocasa.dev/install | sh
> irm https://3code.capocasa.dev/install.ps1 | iex
$ mkdir myprojectdir
$ cd myprojectdir
$ 3code
  ╶──╮
  ╶──┤    3code v0.6.0   the economical coding agent
  ╶──╯

  no provider configured. let's add one. (ctrl+d to quit)
  supported: deepinfra, ovh, nvidia, nebius, fireworks

  api key              : ************************
  provider name or url : nvidia
  verifying... ok
  saved to ~/.config/3code/config
3

Run your first prompt

 Build me a Hello World program in Nim

Why 3code

Claude Code, Cursor, Copilot — they all optimize for capability. Nobody optimizes for cost. 3code is the only coding agent that treats your token budget as a first-class constraint.

That means your $20 subscription lasts a month instead of a week. It means you can use DeepSeek V4 or GLM-5.2 free tier and still get real work done. It means you don't have to think twice before asking a question.

How 3code pulls it off

We eat our own dog food — 3code is now written entirely with 3code, using GLM 5.2 on Z.ai's free tier.

Design: The 3E E's

E1
Token efficiency

There are a many, many design choices where small adjustments lead to huge savings. Then there is one unique innovation- chunked mode- that constantly extracts relevant and discards irrelevant data and keeps agentic coding with very little context- and it improves results because of intelligent discarding.

E2
Computer efficiency

3code is written as a system program and is absolutely tiny. It always loads and responds instantly- as an AI assisted programmer you are an extremely productive individual- your coding agent should not keep you waiting, even for a few seconds.

E3
Ergonomics

The idea that you have to trade efficiency for ease of use is so widespread, most designs that are efficient technically will have poor usability. But it doesn't have to be that way- Enjoy efficiency that's a joy to use.

Data to back it up

SWE-bench Verified: tasks resolved
10-task subset, both agents on Z.ai glm-5.2 via the same LiteLLM proxy, 600s per-task timeout
3code 6 / 10
opencode 5 / 10
Non-cached input tokens
batch total (9 tasks, excluding sympy-23534)
3code ~295k
opencode ~1.45M

opencode sent 4.9× the non-cached input of 3code on comparable work, driven by higher tool-call volume: more file reads, more bash, more context re-sent per turn.

Cached input tokens
batch total (9 tasks, excluding sympy-23534)
3code ~4.7M
opencode ~10.7M

On most providers cached input accounts for 90%+ of token spend, so this is where efficiency pays off most.

Output tokens
batch total (9 tasks, excluding sympy-23534)
3code ~69k
opencode ~106k
Wall time
sum of per-task agent run time across all 10 tasks
3code 55 min
opencode 65 min

Zero eval errors, zero unresolved patches; every timeout produced an empty patch. Full per-task breakdown, observations and methodology →

Technical Details