Skip to content

Compare stack costs

Price AI models or tools in any stack category on the same usage. Every comparison is cost-only and states what could not be modeled.

Cost only. Every model below is priced on the exact same workload and sorted by price. This is not a quality, capability, or suitability ranking, and it does not imply these models are interchangeable.

Models priced on an identical workload, sorted by cost per request, ascending.
ModelProviderCost per requestMonthly costSource
GPT-5 nanoOpenAI$0.000550$110.00Verified 2026-08-15
DeepSeek V4 FlashDeepSeek$0.000700$140.00Verified 2026-08-15
Gemini 2.5 Flash-LiteGoogle$0.000700$140.00Verified 2026-08-15
GPT-4.1 nanoOpenAI$0.000700$140.00Verified 2026-08-15
GPT-4o miniOpenAI$0.001050$210.00Verified 2026-08-15
Mistral Small 4Mistral AI$0.001050$210.00Verified 2026-08-15
GPT-5.6 LunaOpenAI$0.001800$360.00Verified 2026-08-15
DeepSeek V4 ProDeepSeek$0.002175$435.00Verified 2026-08-15
Gemini 3.1 Flash-LiteGoogle$0.002250$450.00Verified 2026-08-15
GPT-5 miniOpenAI$0.002750$550.00Verified 2026-08-15
GPT-4.1 miniOpenAI$0.002800$560.00Verified 2026-08-15
Mistral Large 3Mistral AI$0.003000$600.00Verified 2026-08-15
Gemini 2.5 FlashGoogle$0.003400$680.00Verified 2026-08-15
Gemini 3.5 Flash-LiteGoogle$0.003400$680.00Verified 2026-08-15
Gemini 3.7 FlashGoogle$0.006000$1,200.00Verified 2026-08-15
Claude Haiku 4.5Anthropic$0.008000$1,600.00Verified 2026-08-15
Mistral Medium 3.5Mistral AI$0.012000$2,400.00Verified 2026-08-15
Gemini 3.5 FlashGoogle$0.013500$2,700.00Verified 2026-08-15
GPT-5OpenAI$0.013750$2,750.00Verified 2026-08-15
GPT-5.1OpenAI$0.013750$2,750.00Verified 2026-08-15
GPT-4.1OpenAI$0.014000$2,800.00Verified 2026-08-15
o3OpenAI$0.014000$2,800.00Verified 2026-08-15
Claude Sonnet 5Anthropic$0.016000$3,200.00Verified 2026-08-15
GPT-4oOpenAI$0.017500$3,500.00Verified 2026-08-15
GPT-5.6 TerraOpenAI$0.018000$3,600.00Verified 2026-08-15
GPT-5.2OpenAI$0.019250$3,850.00Verified 2026-08-15
Claude Sonnet 4.5Anthropic$0.024000$4,800.00Verified 2026-08-15
Claude Sonnet 4.6Anthropic$0.024000$4,800.00Verified 2026-08-15
Claude Opus 4.5Anthropic$0.040000$8,000.00Verified 2026-08-15
Claude Opus 4.6Anthropic$0.040000$8,000.00Verified 2026-08-15
Claude Opus 4.7Anthropic$0.040000$8,000.00Verified 2026-08-15
Claude Opus 4.8Anthropic$0.040000$8,000.00Verified 2026-08-15
Claude Opus 5Anthropic$0.040000$8,000.00Verified 2026-08-15
GPT-5.6 SolOpenAI$0.045000$9,000.00Verified 2026-08-15

Not included in these estimates

  • Excludes taxes, contractual or volume discounts, and regional pricing.
  • Excludes batch-processing discounts.
  • Excludes server-side tool calls, web search, and other non-token charges.
  • Text tokens only. Image, audio, and video pricing is not modeled.
  • Cached rate shown is the cache-read (hit) price. Cache-write charges are not modeled.
  • Cache writes, batch discounts, regional platform premiums, tool-specific fees, and taxes are not modeled.
  • Cache writes, batch discounts, US-only inference premiums, tool-specific fees, and taxes are not modeled.
  • Cache writes, batch discounts, data residency premiums, tool-specific fees, and taxes are not modeled.
  • Cache writes, fast mode, batch discounts, data residency premiums, tool-specific fees, and taxes are not modeled.
  • Cached rate shown is the cache-read (hit) price. Cache-write charges (1.25x for 5 minutes, 2x for 1 hour) are not modeled.
  • Excludes fast mode premium pricing and the 1.1x US data-residency multiplier.
  • Excludes the 1.1x US data-residency multiplier.
  • Input rate is the cache-miss price; the cached rate is the cache-hit price.
  • The vendor announced peak / off-peak billing effective 16 August 2026. Time-of-day pricing is not modeled; these are the rates published on the verification date.
  • Thinking tokens are billed as output tokens; audio, image and video generation, cache storage, batch pricing, grounding, maps, tool-use fees, regional pricing, and taxes are not modeled.
  • Audio, image and video generation, cache storage, batch pricing, grounding, maps, tool-use fees, regional pricing, and taxes are not modeled.
  • Rates are the promotional prices published through 31 December 2026. The page lists higher rates ($1.50 in / $7.50 out per million) from 1 January 2027; the future price is not modeled.
  • No cached-input rate is published for this model, so prompt caching is not modeled.
  • Batch, priority processing, long-context fine-tuning premiums, image input, tool-call fees, and taxes are not modeled.
  • Batch, priority processing, image input, audio, fine-tuning, tool-call fees, and taxes are not modeled.
  • Batch, flex, priority processing, image input, tool-call fees, fine-tuning, and taxes are not modeled.
  • Prompts above 272K input tokens use higher rates and are not modeled.
  • Cache writes, taxes, contractual or volume discounts, and regional pricing are not modeled.
  • Reasoning tokens are billed as output tokens; batch, flex, priority processing, image input, tool-call fees, and taxes are not modeled.