Last updated: 2026-08-22 ยท Pricing data current as of this date

๐Ÿ’ฐ AI Agent Cost Calculator

Estimate monthly cost across Claude (Opus 5, Fable 5, Opus 4.8/4.7/4.6, Sonnet 5, Sonnet 4.6, Haiku 4.5), OpenAI (GPT-5.6 Sol/Terra/Luna, GPT-5.5, GPT-5.4, o4-mini), Google (Gemini 3.1 Pro/Flash, 2.5 Pro/Flash, Gemma 2 local), and open-weights models (Kimi K3, Kimi K2, Qwen 3.5/3.6, Gemma 2). API-direct, subscription, and local-Ollama paths side by side. No signup, no tracking, shareable via URL.

๐Ÿ“… Pricing freshness

OpenAI GPT-5.6 rates verified 2026-08-22 โ€” Sol was cut from $5/$30 to $4/$20 on August 21. Claude (Opus 5, Fable 5, Opus 4.6โ€“4.8) rates verified 2026-07-28; Kimi K3 verified 2026-07-18; Google and other open-weights rates last verified 2026-04-26. Confirm exact rates before high-volume commitments at the official pricing pages: anthropic.com ยท openai.com ยท ai.google.dev ยท openrouter.ai. Rates are per 1M tokens, USD.

Your usage

Per user
3k = small repo ยท 30k = large context

Estimated monthly cost

How the math works

Two billing models drive every total below:

  • Per-token billing (API + OpenClaw): messages ร— tokens ร— per-token rate. Costs scale linearly with usage; cheap at low volume, can run high at heavy usage.
  • Subscription (Cowork, ChatGPT, Claude Code Pro/Max): flat per-seat fee with usage caps. Cheap above a usage threshold; potentially wasted money below it.

Published rates (USD per 1M tokens):

  • Anthropic (verified 2026-09-06) โ€” Haiku 4.5 $1/$5 ยท Sonnet 5 $2/$10 ยท Sonnet 4.6 $3/$15 ยท Opus 4.6 / 4.7 / 4.8 / 5 $5/$25 ยท Opus 5.5 $4/$20 ยท Fable 5.1 / Mythos 5.1 / Fable 5 $10/$50 (most capable). Opus 5.5 is the current flagship โ€” 1M-token context as both default and maximum, 128k max output, thinking on by default, at the same $4/$20 as Opus 4.8. Opus 4.7, Opus 4.8, Opus 5.5, Opus 5, Fable 5.1, Mythos 5.1, Fable 5 scale output by the effort multiplier (low 1ร— โ†’ max 7ร—).
  • OpenAI (verified 2026-08-22) โ€” GPT-5.6 Sol $4/$20 ยท Terra $2.50/$15 ยท Luna $1/$6 ยท GPT-5.5 $4/$20 ยท GPT-5.4 $2.50/$10 ยท GPT-5.4 mini $0.30/$1.20 ยท GPT-5.4-Cyber $12/$60 ยท o4-mini $1.50/$8. GPT-5.6 Sol dropped from $5/$30 to $4/$20 on August 21, 2026 โ€” 20% off input, 33% off output. That is the same price as GPT-5.5, so GPT-5.5 is now strictly dominated: identical rates, older model. The Ultrafast tier for Sol has no published pricing and is not modelled here.
  • Google โ€” Gemini 3.1 Flash $0.30/$1.20 ยท Gemini 3.1 Pro $3.50/$15 (newest, Apr 2026) ยท Gemini 2.5 Flash $0.20/$0.80 ยท Gemini 2.5 Pro $2.50/$12 ยท Gemma 2 9B local-only (Ollama)
  • Open weights via API โ€” Kimi K3 $3/$15 (Moonshot 2.8T flagship, Jul 2026 โ€” premium open-weight, priced like Sonnet 4.6 and above Sonnet 5's $2/$10, not a budget tier) ยท Kimi K2 $0.60/$1.80 ยท Qwen 3.5 72B $0.40/$1.20 (typical OpenRouter / Together pricing)
  • Open weights local (Ollama + GPU) โ€” Qwen 3.6 35B MoE, Gemma 2 9B: $0/token, ~$8โ€“18/mo electricity for typical home GPU usage. No data ever leaves your network.
  • Subscriptions โ€” Cowork Pro ~$20/user ยท Business ~$30/user ยท ChatGPT Plus $20 ยท Pro $200 ยท Team $30/seat ยท OpenClaw self-hosted $0โ€“5/mo

Effort level multiplier (Opus 4.7, Opus 4.8, Opus 5.5, Opus 5, Fable 5.1, Mythos 5.1, Fable 5, GPT-6 Sol, GPT-6 Luna, GPT-6 Astra, Gemini 3.8 Flash): low 1ร— ยท medium 1.3ร— ยท high 2ร— ยท xhigh 3.5ร— ยท max 7ร—. Multiplier applies to output tokens (where the extra reasoning work shows up).

What this calculator does not include

  • Prompt-cache discounts โ€” can cut input cost 50โ€“90% for stable system prompts (see our cost optimization guide)
  • Batch API discounts โ€” Anthropic offers 50% off for batch jobs
  • Enterprise volume discounts โ€” negotiate above ~$50K/year spend
  • Tool-call costs (web search, file uploads) โ€” usually small but vary by platform
  • Local GPU electricity if running NemoClaw/Ollama (~$5โ€“20/mo for typical home use)

For most users, real-world spend lands within ยฑ25% of the estimate. The calculator's main job is to spot order-of-magnitude differences between platforms.

Common scenarios

ScenarioCheapest tierWhy
Solo developer, 40 turns/day, SonnetClaude Code Pro / Cowork Pro$20 flat beats per-token at this volume
Heavy user, 200 turns/day, Opus 5 xhighAPI direct or Max planPer-token can run $350+/mo; Max plan caps it
10-person team, mixed usageCowork Business$300/mo total; per-token would be $500+ + dev time
Privacy-sensitive, any volumeOpenClaw + local Qwen 3.6 / Gemma 2$0/token, ~$8โ€“18/mo electricity; data never leaves network
High volume, mixed open + closedKimi K2 or Qwen 3.5 via OpenRouter~3โ€“5ร— cheaper than GPT-5.4 / Sonnet for similar quality
Frontier open-weight coding, self-host optionalKimi K3 via OpenRouter (weights out Jul 27)$3/$15 โ€” matches Sonnet 4.6, a 50% premium over Sonnet 5; tops Frontend Code Arena, and open weights let you self-host at scale
Cheapest chat for non-codingGemini 2.5 Flash or GPT-5.4 miniSub-cent-per-1k-tokens; great for batch summarization
Batch processing, high volumeAPI with batch discount50% off Anthropic batch pricing beats subscriptions

Calculator is informational only โ€” not financial advice. Verify current prices at anthropic.com/pricing and openai.com/pricing. See also: decision guide ยท cost optimization ยท Cowork pricing breakdown ยท effort-levels guide.

๐Ÿ“ฌ Weekly Digest โ€” In Your Inbox

One email a week: top news, releases, and our deepest new guide. No spam. Same content via RSS if you prefer.