Pricing
Pay per token.
Top up, create a key, and scale without subscription caps.
| Model | Input$ per 1M tokens | Output$ per 1M tokens | Cache read$ per 1M tokens |
|---|---|---|---|
DeepSeek V4 FlashFeatured Our fastest model for agentic and coding ⚡ A 1M context window keeps long agentic runs in one pass. The default for new chats and CLI setups. | $0.14 | $0.28 | $0.028 |
Kimi K3 The first open 3T-class model, neck-and-neck with the closed frontier on agentic coding. 1M context, native vision. | $3.00 | $15.00 | $0.30 |
GLM 5.2 Frontier model with 400K context. Best for deep reasoning across large codebases, specs, migrations, and complex multi-step tasks. | $1.40 | $4.40 | $0.26 |
Umans Coder Our router: sends your request to our top pick (today: Kimi K2.7-Code, billed at its rates). | $0.95 | $4.00 | $0.19 |
Kimi K2.7-CodeDeprecated · discontinued Aug 10, 2026 Moonshot’s strongest Kimi coding model with always-on reasoning. Replaced by Kimi K3. | $0.95 | $4.00 | $0.19 |
Umans Flash The light, fast complement for the roles around the main coder: context gathering, scout subagents, summaries, quick edits. | $0.15 | $1.00 | $0.05 |
★ Founding-user rate
Founding users pay $0.14 / $0.28 / $0.0028 per 1M tokens (input / output / cache read) on DeepSeek V4 Flash until Monday, August 10, 2026.
USD per 1M tokens. Top up the wallet, create a key, and every request debits exactly what the model used.
Top up & startLabs experiments are free for seat holders while they run, and founding users have priority on seats. Watch the status page for openings and retirements.