GPT vs Claude vs Gemini: What Production Actually Costs
Benchmark charts tell you which model is smartest. Your finance team asks a different question: what does a month of real traffic cost on each provider? The answer is messier than the price sheets suggest — each vendor prices its tiers differently, discounts differently, and punishes different usage patterns. Here's the comparison with live numbers (updated 2026-08-02), so it stays true after the next price cut.
The three price ladders, side by side
Every vendor sells the same ladder — a budget workhorse, a mid tier, and a flagship — but the rungs sit at very different heights (per 1M tokens, input / output):
| Vendor | Tier | Model | Input / 1M | Output / 1M |
|---|---|---|---|---|
| OpenAI | Budget | GPT-5.4 Mini | $0.75 | $4.50 |
| Mid | GPT-5.4 | $2.50 | $15.00 | |
| Flagship | GPT-5.6 Sol | $5.00 | $30.00 | |
| Anthropic | Budget | Claude Haiku 4.5 | $1.00 | $5.00 |
| Mid | Claude Sonnet 5 | $2.00 | $10.00 | |
| Flagship | Claude Fable 5 | $10.00 | $50.00 | |
| Budget | Gemini 3.5 Flash Lite | $0.3 | $2.50 | |
| Mid | Gemini 3.6 Flash | $1.50 | $7.50 | |
| Flagship | Gemini 3.1 Pro Preview | $2.00 | $12.00 |
Two patterns worth noticing. First, the budget tiers cost 5–15× less than their own vendor's flagship — routine work has no excuse to run on a flagship. Second, the flagship gap between vendors is real money: the same workload can cost 2–4× more on one provider's top model than another's, before quality even enters the conversation.
Three real workloads, priced on every tier
Monthly cost for each vendor's tiers on three representative workloads:
Support chatbot — 1,000 req/day, 1,500 in / 400 out
| Vendor | Budget | Mid | Flagship |
|---|---|---|---|
| OpenAI | $87.75/mo | $293/mo | $585/mo |
| Anthropic | $105/mo | $210/mo | $1050/mo |
| $43.50/mo | $158/mo | $234/mo |
RAG document Q&A — 500 req/day, 6,000 in / 500 out
| Vendor | Budget | Mid | Flagship |
|---|---|---|---|
| OpenAI | $101/mo | $338/mo | $675/mo |
| Anthropic | $128/mo | $255/mo | $1275/mo |
| $45.75/mo | $191/mo | $270/mo |
Coding agent (12-step loops) — 200 req/day, 8,000 in / 1,200 out
| Vendor | Budget | Mid | Flagship |
|---|---|---|---|
| OpenAI | $68.40/mo | $228/mo | $456/mo |
| Anthropic | $84.00/mo | $168/mo | $840/mo |
| $32.40/mo | $126/mo | $182/mo |
The spread within a vendor (budget → flagship) is consistently larger than the spread between vendors at the same tier. Which tier you route to matters more than which logo you pick — the full argument is in our cost-cutting guide. Run your own numbers across all 232 models in the API cost calculator.
Where the price sheets lie: discounts and fine print
- Prompt caching: OpenAI applies it automatically (~90% off cached input, prompts over ~1K tokens). Anthropic makes you mark cache breakpoints explicitly and charges ~25% extra on cache writes — more control, more foot-guns. Google offers implicit caching plus explicit cached-content with hourly storage fees. For chatbots this single feature moves the bill 30–50% — details in the caching guide.
- Batch discounts: all three sell ~50% off for asynchronous jobs. If any of your traffic can wait, the comparison above shifts uniformly down — Batch API guide.
- Long context: Google's large context windows are priced in tiers — past a threshold the per-token rate steps up. Anthropic and OpenAI keep flat rates but cap the window. If you routinely stuff 200K+ tokens, model the tier jump before committing.
- Output verbosity is a hidden multiplier: models differ in how many tokens they produce for the same task. A model that answers 20% shorter is 20% cheaper on output regardless of its rate card — measure on your own prompts.
The decision in three sentences
- Price rarely decides alone at the same tier — differences within a tier are usually smaller than quality differences on your specific task. Run an eval on your real prompts first.
- Route by tier, not by brand: send routine traffic to any budget model and reserve flagships for the hard tail — that's a 10–30× lever, versus ~2× for switching vendors.
- Re-check monthly: these vendors cut prices and ship new tiers constantly — bookmark the live price table; the numbers on this page refresh automatically.