Skip to content
CostPerPrompt

o1 API Pricing

OpenAI · context window 200K · prices updated 2026-09-15

Input per 1M tokens
$15.00
Output per 1M tokens
$60.00
Cached input per 1M
$7.50
50% cheaper than fresh input

Where o1 sits on price

At $60.00 per million output tokens, o1 is a frontier tier model — 28.8× the median output price of $2.08, which puts it in the most expensive 10% of everything we track. It sits 82th cheapest of the 92 OpenAI models we track. Output costs 4.0× input, the usual spread — trim both, starting with the answer length.

Output tokens per $1
16,667
One full context fill
$3.00
Cheaper than
5% of tracked models

Cached input is 50% cheaper. On an input-heavy workload you need roughly a 50% cache-hit rate to take 25% off the input line — reachable for chatbots and agents that resend the same system prompt and history.

What real workloads cost on o1

These three workloads are the ones teams actually run on a frontier-priced model — a $60.00/1M model is not bought for the same job as one ten times the price.

Workload Per request Per month
Small app — 1K requests/day (500 in / 150 out) The baseline most side projects actually run at. $0.0165 $495
Coding agent — 300 sessions/day (60K in / 12K out) Multi-step loops re-read the same files; caching matters more than raw price. $1.62 $14580
Support chatbot — 500 conversations/day (5K in / 1.4K out) Conversation history re-sent each turn — the classic prompt-caching win. $0.159 $2385

Model your exact traffic in the API cost calculator — it preloads o1 with caching and batch options.

o1 price history

o1's price has not moved since we began tracking it on Aug 2, 2026 — 44 days of stability in a market where 78 of the 334 models we track have repriced over the same period, including 9 of OpenAI's own 92 models.

Change-points from our daily price snapshots (tracking since Aug 2, 2026; intraday moves between snapshots are not captured).

o1 vs Claude Fable 5.1

The closest-priced alternative from another vendor is Claude Fable 5.1 (Anthropic) — 17% cheaper on output, with $5.00 less per million input tokens. When two models land this close on price, the decision is quality on your own workload, not the price sheet: run 50 real requests through both and compare.

See Claude Fable 5.1 pricing →

Cheaper alternatives

More OpenAI models

Frequently asked questions

How much does the o1 API cost?

o1 costs $15.00 per million input tokens and $60.00 per million output tokens, with cached input at $7.50 per million (50% cheaper). That works out to roughly 16,667 output tokens per dollar.

What does the small app workload cost on o1?

Small app — 1K requests/day (500 in / 150 out) costs about $0.0165 per request and $495 per month on o1. The baseline most side projects actually run at.

What does it cost to fill o1's 200K context window?

Sending 200K of input in a single request costs $3.00 at $15.00 per million tokens — before any output. With prompt caching that same fill drops to about $1.50 on repeat requests. This is why large context windows are cheap to advertise and expensive to actually use.

Is o1 worth the price?

o1 is priced in the top bracket, at $60.00 per million output tokens. It only pays off where an error is expensive — production code, legal and medical drafting, agents that take real actions. For anything routine, Claude Fable 5.1 at $50.00 does the same job for a fraction of the bill.