Billing

Credits by model: what the same task costs on every AI model

TL;DR

The model you choose changes a task's credit cost by up to 422× — the same median task costs 21 credits on GLM-4.7 Flash and 8,872 on GPT-5.5 Pro. This page prices one measured task on all 49 selectable models, from today's token prices.

How to read this table

Each row prices the same work on a different model. The work is measured, not invented: from 337 completed production tasks over 90 days (to 2026-10-02), ranked by how much they asked of the model. The median task used 78,098 input tokens (53,823 from the prompt cache) and 2,450 output tokens; a small task (25th percentile) 13,020 input tokens (0 from the prompt cache) and 2,078 output tokens; a large one (90th percentile) 87,357 input tokens (0 from the prompt cache) and 17,422 output tokens.

The credits are today's: each model's live per-token price from the AI Gateway, synced daily, converted exactly the way your bill is — so a cheaper model shows fewer credits for the same task. Dollars are at the Starter price of $0.001 per credit; Scale and Agency pay less per credit.

These are model-token credits only. Tools a task calls — web search, the Python sandbox, a hosted browser, image or video generation — cost the same on every model and are listed in credits by task type.

The same task on every model (cheapest first)

ModelProviderSmall taskMedian task≈ $ median (Starter)Large task
GLM-4.7 FlashZhipu (GLM)1221$0.0288
Qwen3.5 FlashAlibaba (Qwen)1527$0.03106
GPT-6 LunaOpenAI1629$0.03117
DeepSeek V4 FlashDeepSeek1531$0.03107
Grok 4.1 FastxAI2548$0.05175
GPT-6 Luna FastOpenAI3256$0.06233
GPT-5.6 LunaOpenAI3460$0.06256
Llama 4 MaverickMeta (Llama)3564$0.06253
Gemini 3.1 Flash LiteGoogle4375$0.07321
MiniMax M2.7MiniMax4379$0.08315
MiniMax M3MiniMax4379$0.08315
GPT-5 miniOpenAI5083$0.08379
Gemini 2.5 FlashGoogle61101$0.10466
Qwen3.7 PlusAlibaba (Qwen)57106$0.11420
Mistral Large 3Mistral65124$0.12466
Kimi K2 ThinkingMoonshot (Kimi)69126$0.13506
DeepSeek V3.2DeepSeek80153$0.15576
DeepSeek V3.2 ThinkingDeepSeek80153$0.15576
Llama 3.3 70BMeta (Llama)73155$0.15503
DeepSeek V4 ProDeepSeek85163$0.16615
GLM-5.2Zhipu (GLM)105200$0.20762
Gemini 3.8 FlashGoogle118210$0.21873
GPT-5.4 miniOpenAI128222$0.22960
Kimi K2.6Moonshot (Kimi)138254$0.251,018
Claude Haiku 4.5Anthropic157280$0.281,164
Grok 4.3xAI144288$0.291,019
Qwen3 Max ThinkingAlibaba (Qwen)188336$0.341,396
GLM-5.1Zhipu (GLM)183349$0.351,327
DeepSeek R1DeepSeek192356$0.361,414
Gemini 2.5 ProGoogle247411$0.411,890
GPT-5OpenAI247411$0.411,890
Mistral Medium 3.5Mistral235420$0.421,745
Gemini 3.5 FlashGoogle255444$0.441,919
Magistral MediumMistral243477$0.481,746
Claude Sonnet 5Anthropic313559$0.562,326
GPT-6 SolOpenAI313559$0.562,326
GPT-5.2OpenAI346575$0.582,646
Gemini 3.1 ProGoogle340592$0.592,559
Qwen3.7 MaxAlibaba (Qwen)321617$0.622,328
GPT-5.4OpenAI425740$0.743,199
Claude Sonnet 4.6Anthropic469839$0.843,490
Claude Opus 4.6Anthropic7811,397$1.405,816
Claude Opus 4.7Anthropic7811,397$1.405,816
Claude Opus 4.8Anthropic7811,397$1.405,816
Claude Opus 5Anthropic7811,397$1.405,816
GPT-5.5OpenAI8501,479$1.486,397
Claude Fable 5Anthropic1,5612,794$2.7911,632
GPT-6 AstraOpenAI1,5612,794$2.7911,632
GPT-5.5 ProOpenAI5,0988,872$8.8738,378
Credits for model tokens only, at today's gateway price. About 43% of measured input tokens were served from the prompt cache, billed at the cache-read rate — that is already in the shapes above. Your own tasks will differ with the data they read and how many steps they take; the Billing page shows your last 30 days of credits by model.

What this means for choosing models

  • The default, GPT-6 Luna, costs 29 credits for the median task — one of the cheapest models on the list, which is why it is the default for new workspaces.
  • Frontier models cost one to two orders of magnitude more for the same task: GPT-5.5 Pro takes 8,872 credits where GLM-4.7 Flash takes 21.
  • The orchestrator writes the plan; the specialists execute it. Plan quality matters most, so a stronger orchestrator with cheap specialists is the usual split — set each agent's model on the Agents page.
  • The plan pages budget 1,350–1,900 credits per task. That figure includes tools and covers most tasks on mid-priced models; on the default model most tasks use far less, and on a frontier model a large task can use more.

Frequently asked questions

Which AI model is cheapest in Agent Planners?
GLM-4.7 Flash — 21 credits for the median measured task in model tokens. The default, GPT-6 Luna, is among the cheapest too.
How much more does a premium model cost per task?
Up to about 422× the cheapest model for the same task: GPT-5.5 Pro takes 8,872 credits for the median task. Mid-tier models such as Claude Sonnet and Gemini Pro sit in between.
Are these numbers measured or estimated?
The amount of work per task is measured from 337 production tasks; the price per token is each model's live price, synced daily. The table is the product of the two, recomputed whenever the page refreshes.
Do tools cost more on an expensive model?
No. Web search, the sandbox, the hosted browser and image / video generation have their own per-call price, the same on every model. Only model tokens change with the model.
How do I see what my own tasks cost?
The Billing page shows your last 30 days of credits, broken down by model and by tool.
Pick the model per agent

A strong orchestrator, cheap specialists — set it on the Agents page. Start free, no card.