Pricing

Pay per token.

Top up, create a key, and go.

ModelInput$ per 1M tokensOutput$ per 1M tokensCache read$ per 1M tokens
DeepSeek V4 FlashFeatured

Our fastest model for agentic and coding ⚡ A 1M context window keeps long agentic runs in one pass. The default for new chats and CLI setups.

★ Founding users: $0.14 / $0.28 / $0.0028 per 1M until Monday, August 10, 2026

$0.14$0.28$0.028
Kimi K3

The first open 3T-class model, neck-and-neck with the closed frontier on agentic coding. 1M context, native vision.

$3.00$15.00$0.30
GLM 5.2

Frontier model with 400K context. Best for deep reasoning across large codebases, specs, migrations, and complex multi-step tasks.

$1.40$4.40$0.26
Umans Coder

Our router: sends your request to our top pick (today: Kimi K2.7-Code, billed at its rates).

$0.95$4.00$0.19
Kimi K2.7-CodeDeprecated · discontinued Aug 10, 2026

Moonshot’s strongest Kimi coding model with always-on reasoning. Replaced by Kimi K3.

$0.95$4.00$0.19
Umans Flash

The light, fast complement for the roles around the main coder: context gathering, scout subagents, summaries, quick edits.

$0.15$1.00$0.05

USD per 1M tokens. Top up the wallet, create a key, and every request debits exactly what the model used.

Top up & start

Labs experiments are free for seat holders while they run, and founding users have priority on seats. Watch the status page for openings and retirements.