๐ฐ AI Agent Cost Calculator
Estimate monthly cost across Claude (Opus 5, Fable 5, Opus 4.8/4.7/4.6, Sonnet 5, Sonnet 4.6, Haiku 4.5), OpenAI (GPT-5.6 Sol/Terra/Luna, GPT-5.5, GPT-5.4, o4-mini), Google (Gemini 3.1 Pro/Flash, 2.5 Pro/Flash, Gemma 2 local), and open-weights models (Kimi K3, Kimi K2, Qwen 3.5/3.6, Gemma 2). API-direct, subscription, and local-Ollama paths side by side. No signup, no tracking, shareable via URL.
OpenAI GPT-5.6 rates verified 2026-08-22 โ Sol was cut from $5/$30 to $4/$20 on August 21. Claude (Opus 5, Fable 5, Opus 4.6โ4.8) rates verified 2026-07-28; Kimi K3 verified 2026-07-18; Google and other open-weights rates last verified 2026-04-26. Confirm exact rates before high-volume commitments at the official pricing pages: anthropic.com ยท openai.com ยท ai.google.dev ยท openrouter.ai. Rates are per 1M tokens, USD.
Your usage
Estimated monthly cost
How the math works
Two billing models drive every total below:
- Per-token billing (API + OpenClaw): messages ร tokens ร per-token rate. Costs scale linearly with usage; cheap at low volume, can run high at heavy usage.
- Subscription (Cowork, ChatGPT, Claude Code Pro/Max): flat per-seat fee with usage caps. Cheap above a usage threshold; potentially wasted money below it.
Published rates (USD per 1M tokens):
- Anthropic (verified 2026-09-06) โ Haiku 4.5 $1/$5 ยท Sonnet 5 $2/$10 ยท Sonnet 4.6 $3/$15 ยท Opus 4.6 / 4.7 / 4.8 / 5 $5/$25 ยท Opus 5.5 $4/$20 ยท Fable 5.1 / Mythos 5.1 / Fable 5 $10/$50 (most capable). Opus 5.5 is the current flagship โ 1M-token context as both default and maximum, 128k max output, thinking on by default, at the same $4/$20 as Opus 4.8. Opus 4.7, Opus 4.8, Opus 5.5, Opus 5, Fable 5.1, Mythos 5.1, Fable 5 scale output by the effort multiplier (low 1ร โ max 7ร).
- OpenAI (verified 2026-08-22) โ GPT-5.6 Sol $4/$20 ยท Terra $2.50/$15 ยท Luna $1/$6 ยท GPT-5.5 $4/$20 ยท GPT-5.4 $2.50/$10 ยท GPT-5.4 mini $0.30/$1.20 ยท GPT-5.4-Cyber $12/$60 ยท o4-mini $1.50/$8. GPT-5.6 Sol dropped from $5/$30 to $4/$20 on August 21, 2026 โ 20% off input, 33% off output. That is the same price as GPT-5.5, so GPT-5.5 is now strictly dominated: identical rates, older model. The Ultrafast tier for Sol has no published pricing and is not modelled here.
- Google โ Gemini 3.1 Flash $0.30/$1.20 ยท Gemini 3.1 Pro $3.50/$15 (newest, Apr 2026) ยท Gemini 2.5 Flash $0.20/$0.80 ยท Gemini 2.5 Pro $2.50/$12 ยท Gemma 2 9B local-only (Ollama)
- Open weights via API โ Kimi K3 $3/$15 (Moonshot 2.8T flagship, Jul 2026 โ premium open-weight, priced like Sonnet 4.6 and above Sonnet 5's $2/$10, not a budget tier) ยท Kimi K2 $0.60/$1.80 ยท Qwen 3.5 72B $0.40/$1.20 (typical OpenRouter / Together pricing)
- Open weights local (Ollama + GPU) โ Qwen 3.6 35B MoE, Gemma 2 9B: $0/token, ~$8โ18/mo electricity for typical home GPU usage. No data ever leaves your network.
- Subscriptions โ Cowork Pro ~$20/user ยท Business ~$30/user ยท ChatGPT Plus $20 ยท Pro $200 ยท Team $30/seat ยท OpenClaw self-hosted $0โ5/mo
Effort level multiplier (Opus 4.7, Opus 4.8, Opus 5.5, Opus 5, Fable 5.1, Mythos 5.1, Fable 5, GPT-6 Sol, GPT-6 Luna, GPT-6 Astra, Gemini 3.8 Flash): low 1ร ยท medium 1.3ร ยท high 2ร ยท xhigh 3.5ร ยท max 7ร. Multiplier applies to output tokens (where the extra reasoning work shows up).
What this calculator does not include
- Prompt-cache discounts โ can cut input cost 50โ90% for stable system prompts (see our cost optimization guide)
- Batch API discounts โ Anthropic offers 50% off for batch jobs
- Enterprise volume discounts โ negotiate above ~$50K/year spend
- Tool-call costs (web search, file uploads) โ usually small but vary by platform
- Local GPU electricity if running NemoClaw/Ollama (~$5โ20/mo for typical home use)
For most users, real-world spend lands within ยฑ25% of the estimate. The calculator's main job is to spot order-of-magnitude differences between platforms.
Common scenarios
| Scenario | Cheapest tier | Why |
|---|---|---|
| Solo developer, 40 turns/day, Sonnet | Claude Code Pro / Cowork Pro | $20 flat beats per-token at this volume |
| Heavy user, 200 turns/day, Opus 5 xhigh | API direct or Max plan | Per-token can run $350+/mo; Max plan caps it |
| 10-person team, mixed usage | Cowork Business | $300/mo total; per-token would be $500+ + dev time |
| Privacy-sensitive, any volume | OpenClaw + local Qwen 3.6 / Gemma 2 | $0/token, ~$8โ18/mo electricity; data never leaves network |
| High volume, mixed open + closed | Kimi K2 or Qwen 3.5 via OpenRouter | ~3โ5ร cheaper than GPT-5.4 / Sonnet for similar quality |
| Frontier open-weight coding, self-host optional | Kimi K3 via OpenRouter (weights out Jul 27) | $3/$15 โ matches Sonnet 4.6, a 50% premium over Sonnet 5; tops Frontend Code Arena, and open weights let you self-host at scale |
| Cheapest chat for non-coding | Gemini 2.5 Flash or GPT-5.4 mini | Sub-cent-per-1k-tokens; great for batch summarization |
| Batch processing, high volume | API with batch discount | 50% off Anthropic batch pricing beats subscriptions |
Calculator is informational only โ not financial advice. Verify current prices at anthropic.com/pricing and openai.com/pricing. See also: decision guide ยท cost optimization ยท Cowork pricing breakdown ยท effort-levels guide.