Credits by model: what the same task costs on every AI model
The model you choose changes a task's credit cost by up to 422× — the same median task costs 21 credits on GLM-4.7 Flash and 8,872 on GPT-5.5 Pro. This page prices one measured task on all 49 selectable models, from today's token prices.
How to read this table
Each row prices the same work on a different model. The work is measured, not invented: from 337 completed production tasks over 90 days (to 2026-10-02), ranked by how much they asked of the model. The median task used 78,098 input tokens (53,823 from the prompt cache) and 2,450 output tokens; a small task (25th percentile) 13,020 input tokens (0 from the prompt cache) and 2,078 output tokens; a large one (90th percentile) 87,357 input tokens (0 from the prompt cache) and 17,422 output tokens.
The credits are today's: each model's live per-token price from the AI Gateway, synced daily, converted exactly the way your bill is — so a cheaper model shows fewer credits for the same task. Dollars are at the Starter price of $0.001 per credit; Scale and Agency pay less per credit.
These are model-token credits only. Tools a task calls — web search, the Python sandbox, a hosted browser, image or video generation — cost the same on every model and are listed in credits by task type.
The same task on every model (cheapest first)
| Model | Provider | Small task | Median task | ≈ $ median (Starter) | Large task |
|---|---|---|---|---|---|
| GLM-4.7 Flash | Zhipu (GLM) | 12 | 21 | $0.02 | 88 |
| Qwen3.5 Flash | Alibaba (Qwen) | 15 | 27 | $0.03 | 106 |
| GPT-6 Luna | OpenAI | 16 | 29 | $0.03 | 117 |
| DeepSeek V4 Flash | DeepSeek | 15 | 31 | $0.03 | 107 |
| Grok 4.1 Fast | xAI | 25 | 48 | $0.05 | 175 |
| GPT-6 Luna Fast | OpenAI | 32 | 56 | $0.06 | 233 |
| GPT-5.6 Luna | OpenAI | 34 | 60 | $0.06 | 256 |
| Llama 4 Maverick | Meta (Llama) | 35 | 64 | $0.06 | 253 |
| Gemini 3.1 Flash Lite | 43 | 75 | $0.07 | 321 | |
| MiniMax M2.7 | MiniMax | 43 | 79 | $0.08 | 315 |
| MiniMax M3 | MiniMax | 43 | 79 | $0.08 | 315 |
| GPT-5 mini | OpenAI | 50 | 83 | $0.08 | 379 |
| Gemini 2.5 Flash | 61 | 101 | $0.10 | 466 | |
| Qwen3.7 Plus | Alibaba (Qwen) | 57 | 106 | $0.11 | 420 |
| Mistral Large 3 | Mistral | 65 | 124 | $0.12 | 466 |
| Kimi K2 Thinking | Moonshot (Kimi) | 69 | 126 | $0.13 | 506 |
| DeepSeek V3.2 | DeepSeek | 80 | 153 | $0.15 | 576 |
| DeepSeek V3.2 Thinking | DeepSeek | 80 | 153 | $0.15 | 576 |
| Llama 3.3 70B | Meta (Llama) | 73 | 155 | $0.15 | 503 |
| DeepSeek V4 Pro | DeepSeek | 85 | 163 | $0.16 | 615 |
| GLM-5.2 | Zhipu (GLM) | 105 | 200 | $0.20 | 762 |
| Gemini 3.8 Flash | 118 | 210 | $0.21 | 873 | |
| GPT-5.4 mini | OpenAI | 128 | 222 | $0.22 | 960 |
| Kimi K2.6 | Moonshot (Kimi) | 138 | 254 | $0.25 | 1,018 |
| Claude Haiku 4.5 | Anthropic | 157 | 280 | $0.28 | 1,164 |
| Grok 4.3 | xAI | 144 | 288 | $0.29 | 1,019 |
| Qwen3 Max Thinking | Alibaba (Qwen) | 188 | 336 | $0.34 | 1,396 |
| GLM-5.1 | Zhipu (GLM) | 183 | 349 | $0.35 | 1,327 |
| DeepSeek R1 | DeepSeek | 192 | 356 | $0.36 | 1,414 |
| Gemini 2.5 Pro | 247 | 411 | $0.41 | 1,890 | |
| GPT-5 | OpenAI | 247 | 411 | $0.41 | 1,890 |
| Mistral Medium 3.5 | Mistral | 235 | 420 | $0.42 | 1,745 |
| Gemini 3.5 Flash | 255 | 444 | $0.44 | 1,919 | |
| Magistral Medium | Mistral | 243 | 477 | $0.48 | 1,746 |
| Claude Sonnet 5 | Anthropic | 313 | 559 | $0.56 | 2,326 |
| GPT-6 Sol | OpenAI | 313 | 559 | $0.56 | 2,326 |
| GPT-5.2 | OpenAI | 346 | 575 | $0.58 | 2,646 |
| Gemini 3.1 Pro | 340 | 592 | $0.59 | 2,559 | |
| Qwen3.7 Max | Alibaba (Qwen) | 321 | 617 | $0.62 | 2,328 |
| GPT-5.4 | OpenAI | 425 | 740 | $0.74 | 3,199 |
| Claude Sonnet 4.6 | Anthropic | 469 | 839 | $0.84 | 3,490 |
| Claude Opus 4.6 | Anthropic | 781 | 1,397 | $1.40 | 5,816 |
| Claude Opus 4.7 | Anthropic | 781 | 1,397 | $1.40 | 5,816 |
| Claude Opus 4.8 | Anthropic | 781 | 1,397 | $1.40 | 5,816 |
| Claude Opus 5 | Anthropic | 781 | 1,397 | $1.40 | 5,816 |
| GPT-5.5 | OpenAI | 850 | 1,479 | $1.48 | 6,397 |
| Claude Fable 5 | Anthropic | 1,561 | 2,794 | $2.79 | 11,632 |
| GPT-6 Astra | OpenAI | 1,561 | 2,794 | $2.79 | 11,632 |
| GPT-5.5 Pro | OpenAI | 5,098 | 8,872 | $8.87 | 38,378 |
What this means for choosing models
- The default, GPT-6 Luna, costs 29 credits for the median task — one of the cheapest models on the list, which is why it is the default for new workspaces.
- Frontier models cost one to two orders of magnitude more for the same task: GPT-5.5 Pro takes 8,872 credits where GLM-4.7 Flash takes 21.
- The orchestrator writes the plan; the specialists execute it. Plan quality matters most, so a stronger orchestrator with cheap specialists is the usual split — set each agent's model on the Agents page.
- The plan pages budget 1,350–1,900 credits per task. That figure includes tools and covers most tasks on mid-priced models; on the default model most tasks use far less, and on a frontier model a large task can use more.
Frequently asked questions
- Which AI model is cheapest in Agent Planners?
- GLM-4.7 Flash — 21 credits for the median measured task in model tokens. The default, GPT-6 Luna, is among the cheapest too.
- How much more does a premium model cost per task?
- Up to about 422× the cheapest model for the same task: GPT-5.5 Pro takes 8,872 credits for the median task. Mid-tier models such as Claude Sonnet and Gemini Pro sit in between.
- Are these numbers measured or estimated?
- The amount of work per task is measured from 337 production tasks; the price per token is each model's live price, synced daily. The table is the product of the two, recomputed whenever the page refreshes.
- Do tools cost more on an expensive model?
- No. Web search, the sandbox, the hosted browser and image / video generation have their own per-call price, the same on every model. Only model tokens change with the model.
- How do I see what my own tasks cost?
- The Billing page shows your last 30 days of credits, broken down by model and by tool.