Claude Opus 5.5 and GPT-6 Sol redraw the cheapest model per score

2 hours ago

Claude Opus 5.5, GPT-6 Sol and three more models landed on September twenty-first and twenty-second. They redrew the line of cheapest models for each score. Artificial Analysis is an independent company that runs the same set of tests on every model. Artificial Analysis averages the results into one score, the Intelligence Index, and records what the whole run costs, divided per task. The line climbs in steps. Each step is the cheapest model that reaches that score. The grey line shows the cheapest models released before September eighteenth. Above fifty-three points, the cheapest of those costs five dollars ninety-eight per task. Today the cheapest model above fifty-three points costs a dollar eighty-two. Along the bottom, each gridline marks ten times the cost of the one before it.

Ask

Ask about this presentation

Answers are generated from this presentation.

Chapters

  1. 0:00The cheapest line, redrawn in two days
  2. 0:42Twice GPT-6 Sol’s token price, a third of Claude Fable 5.1’s cost per task
  3. 1:22Five releases in two days
  4. 1:52Three prices fell, two held
  5. 2:34Cost per task is price times tokens
  6. 3:19How the cheapest line is drawn
  7. 4:11Six times cheaper at 46 points, and you can run it yourself
  8. 4:32At max effort, Claude Opus 5.5 costs 2% more
  9. 5:05What to run this week
  10. 5:39Which step are you on?
Show transcript

The cheapest line, redrawn in two days

US$ per task · index points

Above 26 points, every step of the cheapest line is a model released Sep 21 or 22

Cheapest model per Intelligence Index score, on Artificial Analysis’s board

0102030405060$0.01$0.10$1$10Intelligence Index (points)Cost per task, USD · each gridline 10× the one beforeCheapest today · 53.6 points · $1.82Cheapest before Sep 18 · 53.2 points · $5.98
Released before Sep 18 · as measured todayToday

Each step: the cheapest model that reaches that score

Artificial Analysis Intelligence Index v4.3 and cost per task, read 2026-09-24 · the before line is models released before Sep 18, as measured today1 / 10

Claude Opus 5.5, GPT-6 Sol and three more models landed on September twenty-first and twenty-second. They redrew the line of cheapest models for each score. Artificial Analysis is an independent company that runs the same set of tests on every model. Artificial Analysis averages the results into one score, the Intelligence Index, and records what the whole run costs, divided per task. The line climbs in steps. Each step is the cheapest model that reaches that score. The grey line shows the cheapest models released before September eighteenth. Above fifty-three points, the cheapest of those costs five dollars ninety-eight per task. Today the cheapest model above fifty-three points costs a dollar eighty-two. Along the bottom, each gridline marks ten times the cost of the one before it.

Twice GPT-6 Sol’s token price, a third of Claude Fable 5.1’s cost per task

US$ per task · US$ per million tokens

Claude Opus 5.5 charges twice GPT-6 Sol’s token price and is the cheapest model above 53 points

$1.82 per task for 53.6 points, against Claude Fable 5.1’s $5.98 for 53.2

$0$2$4$6USD per task$5.98Claude Fable 5.1extra high · 53.2 points$1.82Claude Opus 5.5high · 53.6 points
3.3× cheaper per task · derived
Token price, USD per million tokens · input / output
Claude Opus 5.5$4 / $20
GPT-6 Sol$2 / $10
Claude Fable 5.1$10 / $50

Artificial Analysis leaderboard, 2026-09-24 · the Index is rounded to whole points here · xhigh is extra high effort

Token: a small chunk of text · Input: what you send · Output: what the model writes back

Effort: a setting on each request for how long the model thinks before answering

Artificial Analysis, read 2026-09-24 · Claude Opus 5.5 and Claude Fable 5.1 rows are “with fallback” runs, where a safeguard can hand a task to another Claude model · Rates: Anthropic and OpenAI pricing pages2 / 10

Providers charge per million tokens, small chunks of text. Input is what you send, and output is what the model writes back. Claude Opus 5.5 charges four dollars per million input tokens and twenty dollars per million output tokens, twice GPT-6 Sol's price. Effort is a setting on each request for how long the model thinks before it answers. At high effort, Claude Opus 5.5 scores fifty-three point six for that dollar eighty-two per task. The older model was Claude Fable 5.1 on extra high effort. Claude Fable 5.1 is Anthropic's pricier model, at ten dollars per million input tokens. Claude Fable 5.1 scores fifty-three point two and costs three point three times as much per task.

Five releases in two days

Price: USD per million tokens · context in tokens

Five models landed on September 21 and 22, and four of them now sit on the cheapest line

Grok 4.7 is the one on no step

Claude Opus 5.5
Anthropic · Sep 22
Context1M tokens
Price in$4.00
Price out$20.00
Speednot shown
Licenseclosed weights
GPT-6 Sol
OpenAI · Sep 22
Context872K tokens
Price in$2.00
Price out$10.00
Speednot shown
Licenseclosed weights
GPT-6 Luna
OpenAI · Sep 22
Context1M tokens
Price in$0.10
Price out$0.50
Speednot shown
Licenseclosed weights
Grok 4.7
xAI · Sep 21
Context500K tokens
Price in$2.00
Price out$6.00
Speednot shown
Licenseclosed weights
On no step of the cheapest line
MiMo-V2.6-Pro
Xiaomi · Sep 21
Context1M tokens
Price in$0.435
Price out$0.87
Speednot shown
Licenseopen weights · MIT license

Speed: release-week readings still move, so this board leaves them out

Context: how much text one request can hold · Open weights: you can download the model and run it yourself

Closed weights: the model files stay with the maker · MIT license: anyone may use and change the model, even in a paid product

Release dates, context, open-weights flag: Artificial Analysis, read 2026-09-24 · Prices: Anthropic, OpenAI, xAI and Xiaomi pricing pages, read 2026-09-24 · MiMo-V2.6-Pro license: its page on Hugging Face, a public site that hosts AI models3 / 10

On September twenty-second, Anthropic released Claude Opus 5.5, and OpenAI released GPT-6 Sol and its cheaper sibling, GPT-6 Luna. A day earlier, xAI released Grok 4.7, and Xiaomi, the phone company, released MiMo-V2.6-Pro. MiMo-V2.6-Pro is the one with open weights, so you can download the model and run it yourself. The makers of the other four keep those model files to themselves. Four of those five models now sit on the cheapest line. Grok 4.7 sits on no step of it.

Three prices fell, two held

US$ per million tokens · US$ per task

Token prices fell on three of the five releases, and Grok 4.7 costs 47% more per task

Cost per task compares each pair at the same effort setting

Older model to new modelInputOutputCost per task, same effortChange
USD per million tokensUSD per task
Claude Opus 5 to Claude Opus 5.5$5 to $4 · −20%$25 to $20 · −20%high · $3.61 to $1.82−50%
GPT-5.6 Sol to GPT-6 Sol$4 to $2$20 to $10max · $1.99 to $1.06−47%
GPT-5.6 Luna to GPT-6 Luna$0.20 to $0.10$1.20 to $0.50max · $0.18 to $0.068−62%
Grok 4.6 to Grok 4.7$2 · same$6 · samehigh · $1.86 to $2.73+47%
MiMo-V2.5-Pro to MiMo-V2.6-Pro$0.435 · same$0.87 · same26.0 to 46.3 points at the same price·

GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.

Rates: Anthropic, OpenAI, xAI and Xiaomi pricing pages, read 2026-09-24 · Cost per task and Index: Artificial Analysis, read 2026-09-24 · Claude Opus 5.5 rows are “with fallback” runs4 / 10

Token prices fell on three of the five releases. Claude Opus 5.5 cut twenty percent on input and output against Claude Opus 5, the model it replaces. GPT-6 Sol halved GPT-5.6 Sol's price. GPT-5.6 Sol's price was itself a promotional rate, and OpenAI says that promotion runs at least through November twenty-first. Grok 4.7 kept Grok 4.6's price, and Artificial Analysis measures Grok 4.7 at forty-seven percent more per task than Grok 4.6 on high effort. MiMo-V2.6-Pro kept MiMo-V2.5-Pro's price, and Artificial Analysis scores MiMo-V2.6-Pro at forty-six points against twenty-six for MiMo-V2.5-Pro.

Cost per task is price times tokens

US$ per task · index points

A model’s bill per task is its token price times the tokens it uses

For Claude Opus 5.5 at high effort, the price cut did most of the work

Cost per task=Price per token×Tokens used
0102030405060$0.01$0.10$1$10Intelligence Index (points)Cost per task, USD · each gridline 10× the one before
Released Sep 18 or laterBefore Sep 18
$0$1$2$3$4USD per task$3.61Claude Opus 5high$2.16Same tokens atClaude Opus 5.5prices · derived$1.82Claude Opus 5.5highPrice cut$1.45Fewer tokens$0.34
Cache read$0.50 to $0.20 per million tokens · −60%
Share of Claude Opus 5’s high-effort bill50% · derived
Artificial Analysis cost per task and cost split, read 2026-09-24 · $2.16 and the brackets derived from Artificial Analysis data · Claude Opus 5.5 rows are “with fallback” runs · Rates: Anthropic pricing page5 / 10

Every dot here is one model on one effort setting, placed by its score and its cost per task. Cost per task is the price per token times the number of tokens the model used. Take Claude Opus 5 against Claude Opus 5.5 at high effort. Artificial Analysis measures three dollars sixty-one per task, down to a dollar eighty-two. Price Claude Opus 5's own tokens at the new rates and the task costs two dollars sixteen. So the price cut saved a dollar forty-five, and fewer tokens saved thirty-four cents. Most of that price cut came from the cache read. A cache read is text a program sends again on every step, like the same long instructions, and Anthropic bills it at a discount. Anthropic cut that rate sixty percent, from fifty cents to twenty cents per million tokens. Cache reads were half of Claude Opus 5's high-effort bill on this test.

How the cheapest line is drawn

US$ per task · index points

Every step above 26 points sits lower than the line from before September 18

The cheapest line takes the cheapest dot, then moves right only to a higher score

0102030405060$0.01$0.10$1$10Intelligence Index (points)Cost per task, USD · each gridline 10× the one beforePays more for a lower scoreNo older model scores this highGPT-5.6 Luna
Cheapest line todayReleased before Sep 18
At the new steps, the older line costs
1.5× to 11.6× more derived
Around 50 points
GPT-6 Astra · medium · released Sep 3 · $1.54
to Claude Opus 5.5 · medium · $1.34
13% less derived
GPT-5.6 Luna · still on OpenAI’s price list
$0.0098 and $0.016 per task
One model, several effort settings, several steps
Artificial Analysis Intelligence Index v4.3 and cost per task, read 2026-09-24 · multipliers derived from Artificial Analysis data · GPT-5.6 Luna price: OpenAI pricing page, read 2026-09-24 · Claude Opus 5.5 and Claude Fable 5.1 rows are “with fallback” runs, where a safeguard can hand a task to another Claude model6 / 10

To draw the cheapest line, start at the bottom left and take the cheapest dot. Then move right only to a dot that scores higher. The grey line is the same rule, run on models released before September eighteenth. Up to fifty-three points, the older line costs one and a half to eleven point six times more at each new step. Around fifty points, GPT-6 Astra, an OpenAI model from September third, costs a dollar fifty-four per task on medium effort. Claude Opus 5.5 on medium costs a dollar thirty-four, thirteen percent less. The top three steps start at fifty-three point six points, and no older model scores that high. At twenty-one points and at twenty-five points, the cheapest model is still GPT-5.6 Luna. GPT-5.6 Luna is last generation's model, and OpenAI still sells it. GPT-6 Luna holds five steps and Claude Opus 5.5 holds the top four, one step per effort setting.

Six times cheaper at 46 points, and you can run it yourself

US$ per task · index points

MiMo-V2.6-Pro reaches 46 points for 13 cents a task, and you can download it

The highest open-weights score on Artificial Analysis’s board

$0$0.25$0.5$0.75$1USD per task$0.82GPT-6 Astralow · 45.8 points$0.13MiMo-V2.6-Pro46.3 points
6.1× cheaper per task · derived
Top open-weights scores, Intelligence Index (points)
MiMo-V2.6-Pro · Xiaomi46.3
GLM-5.3 · Z.ai44.8
Kimi K3 · Moonshot43.6
MIT license · download it from Hugging Face
Artificial Analysis, read 2026-09-24 · multiplier derived from Artificial Analysis data · License: Hugging Face card XiaomiMiMo/MiMo-V2.6-Pro-RL, license field mit7 / 10

At about forty-six points, GPT-6 Astra on low effort scores forty-five point eight for eighty-two cents per task. MiMo-V2.6-Pro scores forty-six point three for thirteen cents, six point one times less. You can download MiMo-V2.6-Pro and run it yourself. MiMo-V2.6-Pro's forty-six point three is the highest open-weights score on Artificial Analysis's board.

At max effort, Claude Opus 5.5 costs 2% more

US$ per task · tokens

At max effort Claude Opus 5.5 used about twice the tokens and costs 2% more per task

Every lower effort setting costs 29% to 50% less

$0$2$4$6USD per task$5.86Claude Opus 5max · 50.8 points$5.98Claude Opus 5.5max · 57.6 points
+2%
Tokens used on the whole Index, Claude Opus 5 to Claude Opus 5.5, max effort
Input5.5 billion to 11.0 billion
Output140 million to 260 million
Cost per task, same effort, Claude Opus 5 to 5.5
extra high −29%high −50%medium −39%low −50%
Artificial Analysis cost per task and token counts, read 2026-09-24 · Claude Opus 5.5 rows are “with fallback” runs8 / 10

At max effort, the saving disappears. Artificial Analysis measures Claude Opus 5.5 at five dollars ninety-eight per task, against five dollars eighty-six for Claude Opus 5, two percent more. On the full test set, Claude Opus 5.5 read eleven billion input tokens against five and a half billion for Claude Opus 5. Claude Opus 5.5 wrote two hundred and sixty million output tokens against a hundred and forty million for Claude Opus 5. The extra tokens cancel the price cut. At every lower setting, Claude Opus 5.5 costs twenty-nine to fifty percent less.

What to run this week

US$ per task · index points

Claude Opus 5.5 on high is this week’s default, at $1.82 per task for 53.6 points

Four picks, each the cheapest on Artificial Analysis’s board for its score

JobModel and effortIndex (points)Cost per task (USD)
Highest score, cost secondClaude Opus 5.5 · max57.6$5.98
A strong defaultClaude Opus 5.5 · high53.6$1.82
Mid-40s, cheapest, run it yourselfMiMo-V2.6-Pro46.3$0.13
High-volume cheap workGPT-6 Luna · max37.3$0.068

One averaged score · a shortlist for your own tests

Artificial Analysis Intelligence Index v4.3 and cost per task, read 2026-09-24 · Claude Opus 5.5 rows are “with fallback” runs9 / 10

Here is what I'd run this week. For the highest score, when cost comes second, Claude Opus 5.5 on max scores fifty-seven point six for five dollars ninety-eight per task. For a strong default, Claude Opus 5.5 on high scores fifty-three point six for a dollar eighty-two. For the mid-forties at the lowest cost, on hardware you control, MiMo-V2.6-Pro scores forty-six point three for thirteen cents per task. For high-volume cheap work, GPT-6 Luna on max scores thirty-seven point three for under seven cents per task. All four picks rest on one averaged score, so test them on your own work.

Which step are you on?

US$ per task · index points

This week’s four picks all sit on the new cheapest line

Find your model’s score and cost per task, then find its step

0102030405060$0.01$0.10$1$10Intelligence Index (points)Cost per task, USD · each gridline 10× the one beforeClaude Opus 5.5 · maxClaude Opus 5.5 · highMiMo-V2.6-ProGPT-6 Luna · max
Released before Sep 18 · as measured todayToday
Artificial Analysis Intelligence Index v4.3 and cost per task, read 2026-09-2410 / 10

Every step above twenty-six points on this line belongs to a model released on September twenty-first or twenty-second. Find the model you run today on Artificial Analysis's board. Is it still on a step?