‹ Back to the home page

Prices and benchmarks · checked on 26 September 2026

AI cost compared: the same task costs $0.55 or $5.98

How thoroughly the AI works can be set in 5 levels. On the Opus 5.5 model the same test task costs $0.55 at the lowest level and $5.98 at the most thorough one — 11 times as much. The result only rises from 42 to 58 points. Even the “medium” level reaches 51 points: as many as the previous model, Opus 5, at its highest level, for 77% less.

In short: AI is billed by the amount of text, in so-called tokens. The price per token says little about the bill, though. What counts is how many tokens a task uses — and that is set mainly by the effort level, not by the price list. All figures come from the official price list, from independent benchmarks and from our own logs, retrieved on 26 September 2026.

11×
difference between the cheapest and the most expensive effort level — in one and the same model (Opus 5.5)
$82
is what a solved coding task costs with Fable 5 — with Opus 5.5 it is $15
+28%
more cost for no extra performance: Fable 5.1 scores the same at “max” as at “xhigh”
6.3×
as expensive would our own AI work have been if the AI had re-processed known text every time

This page is for anyone who uses AI at work and wants to know what they pay for. What does a token cost? What does a task cost? And which setting really moves the bill? The page does not replace your own costing. It shows the orders of magnitude.

Work out your own cost

Free and without sign-up. Your inputs never leave your browser.

The 4 terms on this page
Token
A text fragment of a few characters. This is what you pay for: per million tokens, separately for text read and text written.
Model
The version of the AI. Claude comes as Fable, Opus, Sonnet and Haiku; the number is the release. Fable is the most expensive, Haiku the cheapest.
Effort level
How thoroughly the model works: 5 levels from “low” to “max”. Higher levels think longer and write more tokens.
Points and tasks
The independent benchmark firm Artificial Analysis gives every model the same test tasks. Points show the result; more is better. Cost per task is the average per test task.

The biggest lever

The effort level decides: the same AI costs up to 11 times as much

Every recent Claude model has an effort setting that decides how thoroughly it works — 5 levels from “low” to “max”. It affects everything the model writes: the answer, the thinking before it and work steps such as searching or reading files. Out of the box Opus 5.5 is set to level 2 (“medium”), all other models to level 3 (“high”). If you never change it, this default decides your bill.

Cost per task and score, measured by the independent benchmark firm Artificial Analysis (Intelligence Index v4.3.2, 10 evaluations). The score compares models with each other; it is not a grade. Retrieved 26 September 2026.
Model and effortPointsCost per taskCost comparedExtra cost vs. level belowVerdict
Opus 5.5 — recommended by Anthropic, widest range
low42$0.55—baseline
medium (default)51$1.34+144%worth it
high54$1.82+36%worth it
xhigh56$3.46+90%weigh up
max58$5.98+73%weigh up
Fable 5.1 — for long, hard tasks
low47$2.37—baseline
medium49$2.98+26%worth it
high (default)51$3.91+31%worth it
xhigh53$5.98+53%weigh up
max53$7.63+28%skip
Opus 5 — the predecessor, still available
low39$1.10—baseline
medium45$2.19+99%worth it
high (default)48$3.61+65%weigh up
xhigh50$4.88+35%worth it
max51$5.86+20%weigh up
Sonnet 5 — the low-cost all-rounder
low24$0.51—baseline
medium28$1.00+96%weigh up
high (default)32$1.79+79%worth it
xhigh34$2.87+60%weigh up
max38$5.09+77%worth it
Haiku 4.5 — no effort levels
fixed setting17$0.21—baseline
The most expensive level adds nothing

Fable 5.1 reaches the same 53 points at “xhigh” and at “max”. The top level costs 28% more. This is the only place in the table where an extra charge demonstrably buys nothing.

The better model costs 89% less

Opus 5.5 at its lowest level reaches 42 points for $0.55. Sonnet 5 at its highest reaches 38 points for $5.09. The higher-priced model is better and 89% cheaper here. It writes about 10,000 tokens per task, Sonnet 5 about 118,000.

The default changes with the model

Opus 5.5 starts at “medium”, all other models at “high”. Switch from Opus 5 to Opus 5.5 without setting it, and you work one level lower. You get 51 instead of 48 points at 63% less cost per task.

How the verdict is set: Up to 20% extra cost per point gained reads “worth it”, above that “weigh up”. If a level gains no point, it reads “skip”. The threshold is our choice, not a measurement — the figures next to it are measured.

Performance

From $1.34 per task, no tested AI scores higher than Opus 5.5

A low price is no use if the result falls short. The independent benchmark firm Artificial Analysis therefore measures score, cost and time on the same tasks. The result: from $1.34 per task, no measured model from any provider scores higher than Opus 5.5. At its default level it needs 3.3 minutes per task and leads in all 6 measured fields.

From $0.55 per task, Opus 5.5 leads all other Claude modelsScore against cost per task. Each line joins the 5 effort levels of one model, from “low” on the left to “max” on the right. How to read it: the higher a point, the better the result; the further left, the cheaper. The cost axis grows in multiples: $1 to $10 is as far as $0.20 to $2.
  • Opus 5.5
  • Fable 5.1
  • Sonnet 5
  • Haiku 4.5 (no effort levels)

Opus 5.5: 42 points for $0.55 up to 58 points for $5.98. Fable 5.1: 47 to 53 points for $2.37 to $7.63. Sonnet 5: 24 to 38 points for $0.51 to $5.09. Haiku 4.5: 17 points for $0.21. All values are listed in the table in the section “The effort level decides”.

Source: Artificial Analysis, Intelligence Index v4.3.2, retrieved 26/09/2026.

Opus 5.5 needs 3.3 minutes per task — less than half the time of Opus 5Pure writing time per task, each model at its default level. Waiting time until the first character is not included.
Show values as a table
Pure writing time per task and output speed at the default effort level.
Model and effortTime per taskTokens per second
Haiku 4.5 (fixed setting)2.9 min86.2
Opus 5.5 (medium)3.3 min81.0
Sonnet 5 (high)6.6 min70.5
Fable 5.1 (high)6.9 min58.8
Opus 5 (high)8.5 min58.0

Source: Artificial Analysis, Intelligence Index v4.3.2, retrieved 26/09/2026.

Opus 5.5 leads in all 6 fieldsPoints per field, each model at its highest level. Each field combines several evaluations — finance and accounting, for example, 7. No model from another provider scores higher in any of the 6 fields either.
  • Opus 5.5
  • Fable 5.1
  • Sonnet 5
Show values as a table
Score per field, each model at its highest effort level.
FieldOpus 5.5Fable 5.1Sonnet 5
Finance and accounting60.756.440.4
Strategy and operations63.759.739.2
Law63.360.542.3
Health and medicine60.558.042.8
Engineering60.456.540.2
Economics65.662.749.4

Source: Artificial Analysis, field scores for Intelligence Index v4.3.2, retrieved 26/09/2026.

For 34 of 42 AI models there is a cheaper choice that is at least as good42 models from 15 providers: every effort level for Claude, the benchmark firm’s default selection for the other providers. The higher a point, the better; the further left, the cheaper. The blue line joins the best-value models: no other model is both cheaper and better than they are. If you use a model below the line, you pay too much for the same result. Switch providers on and off; the line is recalculated each time.
Show or hide providers (15 of 15)
  • Line of best-value models
  • Best-value model
  • Other models

Show values as a table
All models in the comparison, sorted by score. “Open weights” means anyone can download the model and run it themselves.
ProviderModel and effortPointsCost per taskTime per taskTokens per second
AnthropicOpus 5.5 · max57.6$5.98——
AnthropicOpus 5.5 · xhigh56.0$3.467.5 min91.8
AnthropicOpus 5.5 · high53.6$1.824.3 min87.4
AnthropicFable 5.1 · max53.4$7.6311.9 min69.6
AnthropicFable 5.1 · xhigh53.2$5.9810.6 min65.8
OpenAIGPT-6 Astra (max)52.7$3.267.8 min58.2
AnthropicOpus 5.5 · medium51.2$1.343.3 min81.0
AnthropicFable 5.1 · high51.2$3.916.9 min58.8
AnthropicOpus 5 · max50.8$5.8612.6 min61.0
AnthropicOpus 5 · xhigh49.7$4.8810.8 min58.2
AnthropicFable 5.1 · medium48.9$2.985.1 min58.1
MetaMuse Spark 1.3 (max)48.1$1.604.6 min226.8
AnthropicOpus 5 · high48.1$3.618.5 min58.0
OpenAIGPT-6 Sol (max)47.5$1.066.1 min86.2
AnthropicFable 5.1 · low46.8$2.374.0 min59.2
SpaceXAIGrok 4.7 (xhigh)46.4$3.7415.9 min47.5
XiaomiMiMo-V2.6-Pro (open weights)46.3$0.1323.4 min44.7
AlibabaQwen3.8 Max (0902)45.4$5.4133.3 min39.7
Z AIGLM-5.3 (max) (open weights)44.8$2.0113.6 min69.0
AnthropicOpus 5 · medium44.8$2.195.5 min58.2
SpaceXAIGrok 4.6 (high)44.3$1.868.9 min68.5
StepFunStep 5 Preview43.7$0.7212.2 min81.4
KimiKimi K3 (max) (open weights)43.6$2.0021.0 min35.2
AnthropicOpus 5.5 · low42.3$0.551.3 min80.9
Z AIGLM-5.3-Flash (open weights)41.8$0.2519.7 min44.8
GoogleGemini 3.8 Flash (high)40.9$1.243.5 min338.2
DeepSeekDeepSeek V4.1 Flash (max) (open weights)39.5$0.274.4 min236.1
AnthropicOpus 5 · low39.4$1.102.5 min59.6
AnthropicSonnet 5 · max38.2$5.0914.3 min82.1
OpenAIGPT-6 Luna (max)37.3$0.075.7 min146.2
AnthropicSonnet 5 · xhigh34.4$2.879.3 min74.1
AlibabaQwen3.8 27B (xhigh) (open weights)33.7$1.0118.4 min43.9
AnthropicSonnet 5 · high31.7$1.796.6 min70.5
MiniMaxMiniMax-M3 (open weights)29.2$0.514.4 min179.4
AnthropicSonnet 5 · medium28.1$1.004.2 min70.5
Thinking MachinesInkling (open weights)25.0$0.612.9 min167.1
AnthropicSonnet 5 · low24.3$0.512.5 min64.3
NVIDIANemotron 3 Ultra (open weights)22.9$0.553.1 min145.3
GoogleGemini 3.5 Flash-Lite22.2$0.120.8 min355.9
MetaMuse Glimmer (high) (open weights)17.5$0.062.3 min104.4
AnthropicHaiku 4.516.9$0.212.9 min86.2
MistralMistral Medium 3.5 (open weights)14.2$0.442.9 min144.1

Source: Artificial Analysis, Intelligence Index v4.3.2, default selection and model pages, retrieved 26/09/2026. Time per task = pure writing time.

Worked example: coding

Having an app built: a solved task costs $15 with Opus 5.5, $82 with Fable 5

The points above are an average across many fields. In coding the differences are especially large. If you let AI build an app, you also pay for attempts that fail. What counts is the cost per solved task. In the independent coding benchmark Terminal-Bench 4.0, Opus 5.5 at “xhigh” solves 59.6% of the tasks. That costs $14.73 per solved task on average. Opus 5 needs $39.37, Fable 5 $82.28.

How it was measured: Terminal-Bench 4.0 consists of 66 tasks, 18 of them from software development, plus science, machine learning, IT operations, hardware, security and media. Example: speed up the start of a payment service so that overdraft alerts arrive complete and within 5 seconds. The AI works on its own in an isolated test environment, through text commands only, for up to 8 hours per task. A task counts as solved only if all automated tests pass. Artificial Analysis runs each task 3 times and measures cost, tokens and time.

Opus 5.5 is ahead: more solved, faster, cheaperEach model at the effort level with its best result. For Opus 5.5, “xhigh” and “max” reach the same 59.6%; “xhigh” is cheaper. Fable 5 was only measured at “max”. One bar per model for each measure: more is better for tasks solved, less is better for cost, tokens written and time.
  • Opus 5.5 (xhigh)
  • Opus 5 (max)
  • Fable 5 (max)
Show values as a table
Terminal-Bench 4.0, values per task. Cost per solved task = cost per task divided by the share of tasks solved.
Model and effortSolvedCost per taskCost per solved taskTokens writtenTime
Opus 5.5 (xhigh)59.6%$8.78$14.73169,49719.5 min
Opus 5 (max)49.0%$19.29$39.37205,11035.5 min
Fable 5 (max)42.4%$34.91$82.28216,73432.2 min

Source: Artificial Analysis, Terminal-Bench 4.0 in Intelligence Index v4.3.2, values per task, retrieved 26/09/2026. Cost per solved task calculated by us.

Even “medium” beats Opus 5 at its highest level — for 21% of the costOne line per model through its 5 effort levels: the higher, the more tasks solved; the further left, the cheaper. Anthropic states that Opus 5.5 at its default beats Opus 5 at maximum effort for about a fifth of the cost. The independent measurement confirms it.
  • Opus 5.5
  • Opus 5
  • Fable 5 (only measured at max)

Opus 5.5 “medium”: 52.5% solved for $4.04 per task. Opus 5 “max”: 49.0% for $19.29. Fable 5 “max”: 42.4% for $34.91.

Show values as a table
Terminal-Bench 4.0 per effort level: share of tasks solved and cost per task.
Model and effortSolvedCost per task
Opus 5.5 (low)31.3%$2.08
Opus 5.5 (medium)52.5%$4.04
Opus 5.5 (high)56.6%$5.12
Opus 5.5 (xhigh)59.6%$8.78
Opus 5.5 (max)59.6%$13.11
Opus 5 (low)26.3%$5.18
Opus 5 (medium)34.3%$9.18
Opus 5 (high)46.0%$13.14
Opus 5 (xhigh)46.5%$17.35
Opus 5 (max)49.0%$19.29
Fable 5 (max)42.4%$34.91

Source: Artificial Analysis, Terminal-Bench 4.0, retrieved 26/09/2026. Vendor claim: Anthropic, Opus 5.5 announcement of 22/09/2026.

“xhigh” is enough: 33% cheaper, same number solved

Opus 5.5 solves 59.6% at both “xhigh” and “max”. “Max” costs $13.11 instead of $8.78 per task and writes 253,000 instead of 169,000 tokens.

Fable 5 costs 5.6 times as much per solved task

Fable 5 writes more tokens per task than Opus 5.5 at “xhigh” (217,000 instead of 169,000) and pays 2.5 times the token price. It solves fewer: 42.4% instead of 59.6%.

Sensitive requests are included

Requests the provider flags as sensitive, such as IT security tasks, are answered by an older model on Opus 5.5 and Fable 5. The measurements include this. At Vals.ai it affected 30 of 198 attempts; counting them as failures, Opus 5.5 still leads Opus 5 by 53.5% to 45.5%.

4 measurements, the same ranking: Opus 5.5 ahead of Opus 5 ahead of Fable 5

Share of tasks solved in Terminal-Bench 4.0. The values differ because each measurement gives the AI a different working program; the ranking stays the same. Retrieved 26/09/2026.
SourceAI working programOpus 5.5Opus 5Fable 5
Artificial Analysis (independent)mini-swe-agent59.6%49.0%42.4%
Vals.ai (independent)Terminus 261.6%45.5%—
Snorkel (independent)Claude Code—53.9%44.5%
Anthropic (vendor)not stated66.4%52.3%—

“—” means: not measured there. Snorkel measures Opus 5 at “xhigh”, Anthropic Opus 5.5 at “xhigh”; the other values apply to “max” or to the source's standard setting.

Base prices

What the AI writes costs 5 times as much as what it reads

All the costs above come from a handful of prices. Billing is per token — a piece of text, around 4 characters in English. Text the AI reads (your question, attached documents, the conversation so far) is much cheaper than text it writes. All prices per million tokens in US dollars, retrieved on 26 September 2026.

Prices per million tokens. Text that is needed several times, such as the conversation so far, is stored once (the “cache”) and then read again for a fraction of the price. Storing costs a little more than newly entered text, once. “Max. text” is how many tokens the model can handle at once.
ModelNewly enteredWritten by the AIStored for laterReusedMax. text
Fable 5.1$10.00$50.00$12.50$0.251M
Opus 5.5$4.00$20.00$5.00$0.201M
Opus 5$5.00$25.00$6.25$0.501M
Sonnet 5$2.00$10.00$2.50$0.201M
Haiku 4.5$1.00$5.00$1.25$0.10200K
Opus 5.5, fast mode$8.00$40.00$10.00$0.401M

2 discounts come on top: if you collect tasks and have them processed as a batch, you pay half. If you need the same text several times, reading it again costs only 10% of the input price. On Opus 5.5 it is 5%, on Fable 5.1 just 2.5%.

Cost traps

4 cost traps no price list mentions

Besides price and effort level, 4 things push up the bill without anyone noticing.

Same page, more tokens

Models from Claude 4.7 onwards split the same text into about 30% more tokens than their predecessors. An identical price per token therefore does not mean the same price per page of text. Anyone comparing old and new models has to allow for this.

Switching models changes the effort

Without a setting, Opus 5.5 works one level lower than Opus 5. Switch without setting effort, and you get a different bill and a different result. Set the effort level explicitly whenever you change models.

Faster means twice the price

Fast mode delivers answers up to 2.5 times faster — at twice the token price. It does not make the model smarter, only faster. For work that runs in the background it is money thrown away.

Reuse saves the most

In long sessions the AI reads known text from a store instead of processing it again. Without that, our own work would have cost 6.3 times as much. Important: changing the effort for the whole request or switching to fast mode empties this store, and you pay full price again. Changing the effort for single messages only keeps it — available on Fable 5.1, Opus 5.5 and Opus 5.

Our own figures

Our own AI use: 6.3 times cheaper by reusing known text

How big these levers are in practice is shown by our own use. Teddynews uses Claude every day. We analysed all our work sessions from 20 April to 25 September 2026. At today's price list this use would have cost $30,584 — that is what you pay if every request is billed by usage. We ourselves pay a subscription with a fixed monthly price.

The main saving comes from reuse. In long conversations the program resends the whole history with every new question. The AI then reads text it already knows from a store instead of processing it again — for 2.5 to 10% of the normal price. Without this reuse the same work would have cost $192,754, 6.3 times as much.

97
work sessions on 133 of 159 days, with 72,192 AI responses
33.2 bn
tokens processed (text fragments of a few characters), 114 million of them written by the AI itself
$30,584
is what the use would have cost at today's price list — we pay a fixed-price subscription
6.3×
as expensive would it have been if the AI had re-processed known text every time
Reused text: 97% of the volume, only 54% of the costHow the $30,584 split across the 4 items every Claude bill lists. What the AI writes itself is only 0.34% of the volume, yet 10% of the cost.
  • Reused (read from store)
  • Stored for later
  • Written by the AI
  • Newly entered
Show values as a table
The 4 items on the bill: share of the volume (tokens) and of the cost. Amounts rounded; the total is formed from unrounded values.
ItemTokensShare of tokensCostShare of cost
Reused (read from store)32,159.3M96.77%$16,50353.96%
Stored for later956.5M2.88%$10,86635.53%
Written by the AI113.8M0.34%$3,19710.45%
Newly entered3.6M0.01%$190.06%
Total33,233.2M100%$30,584100%

Source: our own Claude Code logs, 20/04–25/09/2026, valued with the provider's price list of 26/09/2026.

Our use per month: 9 times as much in September as in MayIn dollars at today's price list — what the use would have cost without a subscription. April counts from the 20th, September up to the 25th.

Total 20 April to 25 September: $30,584 · 72,192 AI responses · 33.2 bn tokens

Show values as a table
Our use per month at today's price list (Vienna time).
MonthAI responsesTokensAt list price
April3,5641.24 bn$1,093
May8,8981.22 bn$1,023
June13,6376.90 bn$6,577
July11,5515.94 bn$6,748
August12,6466.57 bn$5,924
September21,89611.37 bn$9,219
Total72,19233.23 bn$30,584

Source: our own Claude Code logs, 20/04–25/09/2026, valued with the provider's price list of 26/09/2026.

How we counted: The basis is the Claude Code logs on the Teddynews work computer; each separate session log counts as one session. Every AI response counts once, even if the log writes it several times; the final value applies. Each response is valued at today's price list of its model; storing text at the rate for the storage time shown in the log. Not included: sessions before 20 April 2026 (no logs exist for them), other tools and other computers. For 0.36% of the amount the log does not say whether the fast mode at twice the price was used; we counted the normal price.

Calculator

Calculator: what your AI use costs on 5 models

Estimate your own use with the same 4 items as in our own bill. The example is a longer working session. Everything is calculated in your browser — no input is transmitted.

Everything read for the first time
Known text read from the store — costs 2.5 to 10% of the input price, depending on the model
One-off storage, 1.25 times the input price
Answer, thinking and work steps together
Cost of this session —
The same session with other models

Questions

7 answers on subscriptions, effort levels, tokens and sources

Does this also apply to subscriptions such as ChatGPT Plus or Claude Pro?

No. A subscription has a fixed price; you do not pay per task. The figures on this page apply to usage-based billing, as it arises when AI is built into your own software. For subscribers they are still a yardstick: they show how quickly a usage allowance is used up.

Why is the most expensive level not always the best?

Because extra thinking stops adding value at some point. For Fable 5.1 this point lies between “xhigh” and “max”: same score, 28% more cost. For Opus 5.5 the top level still adds 2 points but costs 73% more. For very long tasks it can still make sense — then for stamina, not for quality.

What changes with Opus 5.5 compared with Opus 5?

3 things. The token price is 20% lower: $4 instead of $5 per million input tokens, $20 instead of $25 for output. Reused text costs 5% instead of 10% of the input price. And without a setting, Opus 5.5 works at “medium” instead of “high”. Opus 5 remains available for now.

What is a token, in one sentence?

A token is a piece of text; in English it is about 4 characters long, according to the provider. Models from Claude 4.7 onwards split the same text into about 30% more tokens.

How do I get my own usage figures?

Every bill lists the 4 items from the calculator separately: newly entered, reused, stored for later, written by the AI. Developer tools usually have their own command for it, web interfaces a usage page in the account.

Are the scores from Anthropic itself?

No. They come from Artificial Analysis, an independent benchmark firm that rates models using a fixed method of 10 evaluations. The prices come from the official price list. Both are linked below with their retrieval date.

Why does Teddynews publish this?

Because we use AI in our own tools and had to do the same calculations. Who is behind it and what they sell is stated openly in the section Who writes this page.

Disclosure

Who writes this page — and what they sell

Teddynews builds AI tools — including AI agents for businesses (German) — and calculates with the same prices. Our own figures in the section “Our own AI use” come from this work. The page costs nothing, collects no address and recommends no provider.

Anyone who publishes something like this usually has something to sell. That is true here too. So it is stated openly, with its real status:

Teddy-LMC

Your documents become a tool that pre-sorts enquiries — instead of a PDF to download.

In preparation. €149 is a proposal, not yet billable.

See Teddy-LMC

Teddy carousel

10 Instagram cards a month, built from your reader's own words. With a forecast of which ones get shared.

In preparation. €79 is a proposal, not yet billable.

See the Teddy carousel

More overviews with sources

Made the same way as this page: electricity and AI, the training market, the reasons for AI scepticism. All with sources.

Free, without sign-up.

AI as a tool · CO₂ and AI

No newsletter, no call-back, no deadline. If you have a question, write to office@teddynews.at. A person answers, not a queue.

Where the figures come from

Sources: 10 references with origin and retrieval date

All prices are official — they come from the provider. Scores, cost, tokens and time per task are independently measured — measured by independent benchmark firms, not confirmed by the provider; vendor claims are named as such. The figures in the section “Our own AI use” are our own count — counted from our logs. No value on this page is estimated; calculated values such as the cost per solved task are marked as such.

How this page was made: Research and text were created with AI assistance and checked by Teddynews (disclosure under Art. 50 of the EU AI Act). Data as of 26/09/2026; prices and measurements were retrieved the same day.

For the cost per task and effort level there is currently no second independent measurement; Artificial Analysis is the only benchmark firm measuring this way. Prices change. This page therefore gives the retrieval date for every figure. If you use it as the basis for a costing, please check the current state at the linked source.

‹ Back to the home page
Buy Teddy a Coffee