Skip to content

AI data

Updated · By , AI-assisted.

Price per task: what finishing an AI task costs, model by model

Finishing one task cost from $0.00013 on GPT-6 Luna to $0.0223 on Claude Fable 5.1, 177 times as much, in our test runs of September 27, 28, 29 and October 2, 7, 8, 9, 2026. A finished task is a reply that passed its check. Each figure is what OpenRouter charged for all 50 prompts, divided by how many passed.

The price per task index

Everything each model was charged for our 50 prompts, divided by the replies that passed their check, cheapest first. “Times the cheapest” is against GPT-6 Luna.

Price per task, every model
ModelTasks passedPrice per taskTimes the cheapestSpent on all
GPT-6 Luna47 of 50$0.00013the cheapest$0.0059
GLM 5.3 Flash40 of 50$0.000241.9 times$0.0096
Claude Haiku 5.545 of 50$0.000262.1 times$0.0118
DeepSeek V4.1 Flash48 of 50$0.000544.3 times$0.0260
GLM 5.348 of 50$0.000766 times$0.0364
GPT-6.1 Sol46 of 50$0.001310 times$0.0597
Gemini 3.8 Flash46 of 50$0.001411 times$0.0644
Claude Haiku 4.543 of 50$0.001714 times$0.0741
GPT-6 Sol45 of 50$0.002923 times$0.1297
Mistral Large 447 of 50$0.003125 times$0.1453
DeepSeek V4 Pro49 of 50$0.003730 times$0.1827
Claude Sonnet 5.547 of 50$0.004132 times$0.1912
Kimi K346 of 50$0.004234 times$0.1938
Claude Sonnet 546 of 50$0.004435 times$0.2030
Grok 4.747 of 50$0.008266 times$0.3875
Gemini 3.1 Pro49 of 50$0.010987 times$0.5334
Claude Opus 5.549 of 50$0.010987 times$0.5360
GPT-6 Astra48 of 50$0.012297 times$0.5847
Claude Fable 5.145 of 50$0.0223177 times$1.0018

The cheapest way to finish each job

Each job's five prompts on their own: the model that passed all five for the least per task, and the one that cost the most doing it.

The cheapest and dearest model to pass all of each job's prompts
JobPassed all fiveCheapest per taskDearest per task
Coding17 of 19 modelsGLM 5.3 Flash, $0.00025Claude Fable 5.1, $0.0425
Writing3 of 19 modelsGPT-6 Luna, $0.000099GPT-6 Astra, $0.0111
Math18 of 19 modelsClaude Haiku 5.5, $0.00015Claude Fable 5.1, $0.0116
Summarization10 of 19 modelsDeepSeek V4.1 Flash, $0.00020GPT-6 Astra, $0.0109
Data analysis15 of 19 modelsGPT-6 Luna, $0.00018Claude Fable 5.1, $0.0283
Customer support4 of 19 modelsGLM 5.3, $0.00047Claude Opus 5.5, $0.0108
Translation18 of 19 modelsGPT-6 Luna, $0.000093Claude Fable 5.1, $0.0175
SQL19 of 19 modelsGPT-6 Luna, $0.000089Claude Fable 5.1, $0.0185
RAG and answering from documents19 of 19 modelsGPT-6 Luna, $0.000077Claude Fable 5.1, $0.0145
Agents and tool use18 of 19 modelsGPT-6 Luna, $0.000081Claude Fable 5.1, $0.0119

Every model, every job

Every model's price per task on every job; a dash means none of its replies to that job passed.

Price per task, by model and job
ModelCodingWritingMathSummarizationData analysisCustomer supportTranslationSQLRAG and answering from documentsAgents and tool use
GPT-6 Luna$0.00029$0.000099$0.00012$0.00012$0.00018$0.00011$0.000093$0.000089$0.000077$0.000081
GLM 5.3 Flash$0.00025$0.00080$0.00015$0.00019$0.00048$0.00049$0.000099$0.00013$0.00015$0.00016
Claude Haiku 5.5$0.00045$0.00027$0.00015$0.00053$0.00036$0.00029$0.00018$0.00022$0.00014$0.00014
DeepSeek V4.1 Flash$0.0018$0.00056$0.00025$0.00020$0.00085$0.00058$0.00058$0.00020$0.00017$0.00020
GLM 5.3$0.0015$0.00072$0.00035$0.00068$0.0019$0.00047$0.00070$0.00032$0.00050$0.00044
GPT-6.1 Sol$0.0023$0.0014$0.00086$0.0010$0.0016$0.0031$0.0013$0.00099$0.00079$0.00073
Gemini 3.8 Flash$0.0020$0.0022$0.0014$0.00075$0.0035$0.0020$0.00075$0.00074$0.00066$0.00087
Claude Haiku 4.5$0.0046$0.0014$0.0015$0.0010$0.0042$0.0024$0.0015$0.00099$0.0010$0.00095
GPT-6 Sol$0.0052$0.0028$0.0018$0.0027$0.0034$0.0049$0.0035$0.0021$0.0016$0.0016
Mistral Large 4$0.0091$0.0030$0.0018$0.0043$0.0037$0.0019$0.0031$0.0022$0.0014$0.0013
DeepSeek V4 Pro$0.0240$0.0011$0.00088$0.00091$0.0029$0.0019$0.0020$0.0020$0.00060$0.00066
Claude Sonnet 5.5$0.0063$0.0043$0.0025$0.0042$0.0047$0.0058$0.0042$0.0038$0.0028$0.0026
Kimi K3$0.0063$0.0048$0.0037$0.0057$0.0068$0.0057$0.0036$0.0024$0.0015$0.0027
Claude Sonnet 5$0.0102$0.0043$0.0033$0.0047$0.0086$0.0036$0.0030$0.0028$0.0022$0.0024
Grok 4.7$0.0246$0.0096$0.0072$0.0028$0.0111$0.0077$0.0082$0.0048$0.0029$0.0039
Gemini 3.1 Pro$0.0195$0.0115$0.0100$0.0088$0.0155$0.0106$0.0140$0.0077$0.0058$0.0057
Claude Opus 5.5$0.0171$0.0231$0.0060$0.0095$0.0131$0.0108$0.0108$0.0092$0.0073$0.0050
GPT-6 Astra$0.0223$0.0111$0.0085$0.0109$0.0154$0.0193$0.0126$0.0098$0.0079$0.0068
Claude Fable 5.1$0.0425$0.0287$0.0116$0.0268$0.0283$0.0276$0.0175$0.0185$0.0145$0.0119

In llmwise you pay per message instead, a count you see before you send: each model's price per message on every plan is on its page of what an AI message costs.

More AI data

Also from primary sources: OpenRouter's usage accounting docs (the cost OpenRouter reports for every request, which these figures add up). Anthropic's Claude Opus 5.5 page (the model that grades the rubric prompts). OpenAI's GPT-6 Astra docs (the model that grades Claude Opus 5.5's own replies). Read .

Bar chart: What one finished task cost, by model. GPT-6 Luna: $0.00013; GLM 5.3 Flash: $0.00024; Claude Haiku 5.5: $0.00026; DeepSeek V4.1 Flash: $0.00054; GLM 5.3: $0.00076; GPT-6.1 Sol: $0.00130; Gemini 3.8 Flash: $0.00140; Claude Haiku 4.5: $0.00172; GPT-6 Sol: $0.00288; Mistral Large 4: $0.00309; DeepSeek V4 Pro: $0.00373; Claude Sonnet 5.5: $0.00407; Kimi K3: $0.00421; Claude Sonnet 5: $0.00441; Grok 4.7: $0.00824; Gemini 3.1 Pro: $0.0109; Claude Opus 5.5: $0.0109; GPT-6 Astra: $0.0122; Claude Fable 5.1: $0.0223.
Our test runs of September 27, 28, 29 and October 2, 7, 8, 9, 2026: what OpenRouter charged for all the prompts, divided by the ones each model passed.

Price per task: how it's worked out

How is the price per task worked out?

Everything OpenRouter charged for a model's replies to a job's prompts, passed or not, divided by the replies that passed their check. A model that fails often pays for its failures: its price per task rises with them.

Is this what llmwise charges?

No. In llmwise a message is one of your plan's published count, however long the reply runs, so a model's price per message on a plan doesn't move with its tokens. This study is what the same work costs through the models' APIs.

Will my own tasks cost the same?

Our prompts start new chats and ask for one reply each: a long chat or a long document costs more, since every reply rereads it. The runs count for 45 days; we run them again after that, and when a model's price changes.

Run your own task on the models that passed

Every model in these runs is in llmwise, its count shown before you send. Start on the cheapest one that passed, and switch up only when a task needs it.