Model vs model
Claude Sonnet 5 vs GPT-6 Sol
Claude Sonnet 5 and GPT-6 Sol are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.
Model prices and specs checked against OpenRouter's Claude Sonnet 5 page, OpenRouter's GPT-6 Sol page. Updated .
Short answer
They cost the same in llmwise: up to 125 messages a month on Pro on either. Beyond that, they read the same files and neither is easier to try. In our test runs, Claude Sonnet 5 passed 46 of the 50 prompts both answered and GPT-6 Sol 45; 7 prompts split them, most on customer support (5 to 3).
Claude Sonnet 5 vs GPT-6 Sol, prompt by prompt
Every prompt Claude Sonnet 5 and GPT-6 Sol both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.
Of the 50 prompts both answered, both passed 42, only Claude Sonnet 5 passed 4, only GPT-6 Sol passed 3, and neither passed 1. Claude Sonnet 5 answered sooner on 10 of the 50 and GPT-6 Sol on 35; the rest were within 10% of each other. The 50 replies cost $0.2030 on Claude Sonnet 5 and $0.1297 on GPT-6 Sol: 1.6× less on GPT-6 Sol.
The 7 prompts only one of Claude Sonnet 5 and GPT-6 Sol passed
Announce a second bakery shop on LinkedIn (writing): GPT-6 Sol passed and Claude Sonnet 5 didn't. Claude Sonnet 5: Graded 3.5 of 5 on average (lowest 3). GPT-6 Sol: Graded 4.8 of 5 on average (lowest 4).
Rewrite corporate jargon in plain words (writing): Claude Sonnet 5 passed and GPT-6 Sol didn't. Claude Sonnet 5: Graded 4.7 of 5 on average (lowest 4). GPT-6 Sol: Graded 3.7 of 5 on average (lowest 3).
An email thread in one sentence (summarization): GPT-6 Sol passed and Claude Sonnet 5 didn't. Claude Sonnet 5: Graded 5.0 of 5 on average (lowest 5); but 32 words, over the 30 allowed. GPT-6 Sol: Graded 5.0 of 5 on average (lowest 5).
Correlation between ad spend and sign-ups (data analysis): GPT-6 Sol passed and Claude Sonnet 5 didn't. Claude Sonnet 5: Final answer 0.99; expected 0.97. GPT-6 Sol: Final answer 0.97: right.
A late order (customer support): Claude Sonnet 5 passed and GPT-6 Sol didn't. Claude Sonnet 5: Graded 4.7 of 5 on average (lowest 4). GPT-6 Sol: Graded 3.7 of 5 on average (lowest 1).
A frustrated customer (customer support): Claude Sonnet 5 passed and GPT-6 Sol didn't. Claude Sonnet 5: Graded 4.3 of 5 on average (lowest 4). GPT-6 Sol: Graded 3.0 of 5 on average (lowest 2).
Idioms into natural Japanese (translation): Claude Sonnet 5 passed and GPT-6 Sol didn't. Claude Sonnet 5: Back-translation chrF 0.42 (pass at 0.4). GPT-6 Sol: Back-translation chrF 0.40 (pass at 0.4); reads back too far from the original.
Job by job, the widest gaps first
Customer support: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 3 of 5. Their median waits were close, 3.5 s against 3.2 s. GPT-6 Sol cost 1.2× less, $0.0182 against $0.0146 for the 5 replies. Claude Sonnet 5's replies ran 150% longer, in tokens of reply, thinking not counted.
Data analysis: Claude Sonnet 5 passed 4 of 5 and GPT-6 Sol 5 of 5. GPT-6 Sol answered 1.9× sooner at the median, 5.3 s against 2.8 s. GPT-6 Sol cost 2.0× less, $0.0344 against $0.0169 for the 5 replies. Claude Sonnet 5's replies ran 86% longer, in tokens of reply, thinking not counted.
Summarization: Claude Sonnet 5 passed 3 of 5 and GPT-6 Sol 4 of 5. GPT-6 Sol answered 1.3× sooner at the median, 2.7 s against 2.0 s. GPT-6 Sol cost 1.3× less, $0.0140 against $0.0110 for the 5 replies. Claude Sonnet 5's replies ran 48% longer, in tokens of reply, thinking not counted.
Translation: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 4 of 5. Claude Sonnet 5 answered 1.2× sooner at the median, 2.8 s against 3.2 s. They cost about the same, $0.0149 against $0.0140 for the 5 replies. Claude Sonnet 5's replies ran 72% longer, in tokens of reply, thinking not counted.
Coding: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 5 of 5. Claude Sonnet 5 answered 1.8× sooner at the median, 2.7 s against 4.8 s. GPT-6 Sol cost 1.9× less, $0.0508 against $0.0262 for the 5 replies. Claude Sonnet 5's replies ran 108% longer, in tokens of reply, thinking not counted.
Math: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 5 of 5. GPT-6 Sol answered 1.8× sooner at the median, 3.9 s against 2.2 s. GPT-6 Sol cost 1.8× less, $0.0165 against $0.0090 for the 5 replies. Claude Sonnet 5's replies ran 96% longer, in tokens of reply, thinking not counted.
Writing: Claude Sonnet 5 passed 4 of 5 and GPT-6 Sol 4 of 5. GPT-6 Sol answered 1.3× sooner at the median, 4.5 s against 3.5 s. GPT-6 Sol cost 1.5× less, $0.0171 against $0.0113 for the 5 replies. Claude Sonnet 5's replies ran 120% longer, in tokens of reply, thinking not counted.
Agents and tool use: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 5 of 5. Their median waits were close, 2.0 s against 1.9 s. GPT-6 Sol cost 1.5× less, $0.0121 against $0.0081 for the 5 replies. Claude Sonnet 5's replies ran 59% longer, in tokens of reply, thinking not counted.
SQL: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 5 of 5. GPT-6 Sol answered 1.6× sooner at the median, 2.2 s against 1.3 s. GPT-6 Sol cost 1.4× less, $0.0142 against $0.0104 for the 5 replies. Claude Sonnet 5's replies ran 46% longer, in tokens of reply, thinking not counted.
RAG and answering from documents: Claude Sonnet 5 passed 5 of 5 and GPT-6 Sol 5 of 5. GPT-6 Sol answered 1.6× sooner at the median, 1.8 s against 1.1 s. GPT-6 Sol cost 1.3× less, $0.0109 against $0.0081 for the 5 replies. Claude Sonnet 5's replies ran 48% longer, in tokens of reply, thinking not counted.
All 50 prompts: who passed, who answered sooner, who cost less
| Prompt | Result | Sooner | Cheaper |
|---|---|---|---|
| Turn a title into a URL slug | Both passed | Claude Sonnet 5, 1.4×took 2.3 s and 3.3 s | Claude Sonnet 5, 1.2×cost $0.0024 and $0.0028 |
| Parse a duration like “1h 30m” | Both passed | Claude Sonnet 5, 1.9×took 2.6 s and 4.8 s | Claude Sonnet 5, 1.2×cost $0.0036 and $0.0042 |
| Merge overlapping intervals | Both passed | GPT-6 Sol, 1.6×took 2.7 s and 1.7 s | GPT-6 Sol, 1.5×cost $0.0034 and $0.0022 |
| Evaluate an arithmetic expression, no eval | Both passed | GPT-6 Sol, 1.7×took 16.9 s and 9.7 s | GPT-6 Sol, 2.5×cost $0.0226 and $0.0089 |
| Parse CSV with quoted fields | Both passed | GPT-6 Sol, 1.8×took 15.8 s and 8.9 s | GPT-6 Sol, 2.3×cost $0.0187 and $0.0081 |
| Announce a second bakery shop on LinkedIn | Only GPT-6 Sol | GPT-6 Sol, 1.3×took 4.6 s and 3.5 s | GPT-6 Sol, 1.4×cost $0.0037 and $0.0027 |
| Rewrite corporate jargon in plain words | Only Claude Sonnet 5 | Claude Sonnet 5, 1.8×took 2.7 s and 5.0 s | Claude Sonnet 5, 1.5×cost $0.0022 and $0.0033 |
| Decline a meeting and offer two times | Both passed | GPT-6 Sol, 2.1×took 2.8 s and 1.3 s | GPT-6 Sol, 1.8×cost $0.0026 and $0.0014 |
| A product announcement with five rules | Both passed | GPT-6 Sol, 2.8×took 4.5 s and 1.6 s | GPT-6 Sol, 2.4×cost $0.0035 and $0.0014 |
| Argue both sides of free buses | Both passed | GPT-6 Sol, 1.6×took 6.6 s and 4.2 s | GPT-6 Sol, 2.0×cost $0.0050 and $0.0025 |
| A discount, then sales tax | Both passed | GPT-6 Sol, 1.2×took 1.9 s and 1.7 s | GPT-6 Sol, 1.2×cost $0.0015 and $0.0013 |
| Pens at 3 for $4 | Both passed | GPT-6 Sol, 1.9×took 6.3 s and 3.3 s | GPT-6 Sol, 2.5×cost $0.0061 and $0.0024 |
| Compound interest over three years | Both passed | GPT-6 Sol, 2.6×took 3.7 s and 1.4 s | GPT-6 Sol, 1.5×cost $0.0021 and $0.0014 |
| Four-digit numbers whose digits sum to 9 | Both passed | GPT-6 Sol, 1.8×took 4.5 s and 2.5 s | GPT-6 Sol, 2.0×cost $0.0039 and $0.0019 |
| The highest of three dice is a 5 | Both passed | GPT-6 Sol, 1.8×took 3.9 s and 2.2 s | GPT-6 Sol, 1.4×cost $0.0028 and $0.0020 |
| An article in three bullets | Neither passed | GPT-6 Sol, 1.3×took 2.6 s and 2.0 s | GPT-6 Sol, 1.3×cost $0.0027 and $0.0020 |
| An email thread in one sentence | Only GPT-6 Sol | GPT-6 Sol, 1.1×took 2.3 s and 2.0 s | GPT-6 Sol, 1.4×cost $0.0019 and $0.0014 |
| Decisions and action items from a meeting | Both passed | Claude Sonnet 5, 1.2×took 2.8 s and 3.5 s | Closecost $0.0033 and $0.0034 |
| A quarterly memo for the CEO | Both passed | GPT-6 Sol, 1.1×took 3.0 s and 2.6 s | GPT-6 Sol, 1.3×cost $0.0033 and $0.0024 |
| A study with a negative result | Both passed | GPT-6 Sol, 1.6×took 2.7 s and 1.7 s | GPT-6 Sol, 1.6×cost $0.0028 and $0.0017 |
| The region with the most revenue | Both passed | GPT-6 Sol, 2.3×took 6.4 s and 2.8 s | GPT-6 Sol, 1.5×cost $0.0041 and $0.0026 |
| Average order value in August | Both passed | GPT-6 Sol, 1.4×took 3.9 s and 2.8 s | GPT-6 Sol, 1.6×cost $0.0041 and $0.0025 |
| Revenue change from July to August | Both passed | GPT-6 Sol, 1.9×took 5.3 s and 2.8 s | GPT-6 Sol, 2.5×cost $0.0062 and $0.0025 |
| A median, filtered two ways | Both passed | GPT-6 Sol, 1.9×took 4.9 s and 2.5 s | GPT-6 Sol, 1.5×cost $0.0036 and $0.0024 |
| Correlation between ad spend and sign-ups | Only GPT-6 Sol | GPT-6 Sol, 2.1×took 13.8 s and 6.6 s | GPT-6 Sol, 2.4×cost $0.0165 and $0.0069 |
| A late order | Only Claude Sonnet 5 | Closetook 3.5 s and 3.2 s | GPT-6 Sol, 1.4×cost $0.0036 and $0.0025 |
| A return inside the window | Both passed | GPT-6 Sol, 2.0×took 2.9 s and 1.5 s | GPT-6 Sol, 1.9×cost $0.0032 and $0.0017 |
| A frustrated customer | Only Claude Sonnet 5 | Claude Sonnet 5, 1.3×took 4.4 s and 5.9 s | Closecost $0.0041 and $0.0043 |
| A refund request outside the window | Both passed | Closetook 3.2 s and 3.0 s | GPT-6 Sol, 1.3×cost $0.0031 and $0.0025 |
| A message with a planted instruction | Both passed | Closetook 4.6 s and 4.4 s | Closecost $0.0041 and $0.0037 |
| A delivery message into Spanish | Both passed | Claude Sonnet 5, 1.2×took 2.7 s and 3.2 s | GPT-6 Sol, 1.1×cost $0.0029 and $0.0026 |
| A product description into French | Both passed | GPT-6 Sol, 1.2×took 2.5 s and 2.1 s | GPT-6 Sol, 1.5×cost $0.0029 and $0.0019 |
| A meeting note into German | Both passed | Claude Sonnet 5, 1.8×took 2.8 s and 5.0 s | Claude Sonnet 5, 1.3×cost $0.0027 and $0.0035 |
| Idioms into natural Japanese | Only Claude Sonnet 5 | GPT-6 Sol, 1.7×took 4.1 s and 2.4 s | GPT-6 Sol, 1.4×cost $0.0028 and $0.0021 |
| A lease clause into Brazilian Portuguese | Both passed | Claude Sonnet 5, 1.2×took 3.8 s and 4.6 s | Claude Sonnet 5, 1.1×cost $0.0036 and $0.0040 |
| Customers in one country | Both passed | GPT-6 Sol, 1.7×took 1.8 s and 1.1 s | GPT-6 Sol, 1.7×cost $0.0020 and $0.0012 |
| Count orders by status | Both passed | GPT-6 Sol, 1.8×took 2.0 s and 1.1 s | GPT-6 Sol, 1.6×cost $0.0020 and $0.0012 |
| Revenue by category | Both passed | GPT-6 Sol, 2.7×took 3.5 s and 1.3 s | GPT-6 Sol, 1.7×cost $0.0031 and $0.0018 |
| Every customer, even those without orders | Both passed | GPT-6 Sol, 1.5×took 2.2 s and 1.4 s | GPT-6 Sol, 1.6×cost $0.0030 and $0.0018 |
| Monthly revenue with a running total | Both passed | Claude Sonnet 5, 1.5×took 3.0 s and 4.5 s | Closecost $0.0042 and $0.0043 |
| A fact from one section | Both passed | GPT-6 Sol, 1.5×took 1.7 s and 1.1 s | GPT-6 Sol, 1.3×cost $0.0019 and $0.0015 |
| Core hours and start times | Both passed | GPT-6 Sol, 1.9×took 2.1 s and 1.1 s | GPT-6 Sol, 1.4×cost $0.0022 and $0.0016 |
| Two sections in one answer | Both passed | GPT-6 Sol, 1.5×took 1.8 s and 1.2 s | GPT-6 Sol, 1.4×cost $0.0022 and $0.0016 |
| A later amendment changes the answer | Both passed | Closetook 2.1 s and 2.0 s | GPT-6 Sol, 1.3×cost $0.0026 and $0.0020 |
| A question the handbook doesn't answer | Both passed | GPT-6 Sol, 2.0×took 1.8 s and 0.9 s | GPT-6 Sol, 1.3×cost $0.0019 and $0.0014 |
| Pick the tool and work out the date | Both passed | Claude Sonnet 5, 1.3×took 1.7 s and 2.2 s | GPT-6 Sol, 1.1×cost $0.0018 and $0.0016 |
| Convert a currency | Both passed | Closetook 2.0 s and 1.8 s | GPT-6 Sol, 1.5×cost $0.0018 and $0.0012 |
| Book a meeting from a sentence | Both passed | GPT-6 Sol, 3.0×took 5.7 s and 1.9 s | GPT-6 Sol, 1.8×cost $0.0037 and $0.0021 |
| Search, but don't book | Both passed | GPT-6 Sol, 1.7×took 1.9 s and 1.1 s | GPT-6 Sol, 1.5×cost $0.0020 and $0.0014 |
| Two calls with a unit conversion | Both passed | GPT-6 Sol, 1.4×took 3.8 s and 2.8 s | GPT-6 Sol, 1.6×cost $0.0028 and $0.0018 |
Claude Sonnet 5 vs GPT-6 Sol in our test runs
Claude Sonnet 5 and GPT-6 Sol on the same prompts, job by job: how many replies passed their check.
Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
| Job | Claude Sonnet 5 | GPT-6 Sol |
|---|---|---|
| Coding | 5 of 5 | 5 of 5 |
| Writing | 4 of 5 | 4 of 5 |
| Math | 5 of 5 | 5 of 5 |
| Summarization | 3 of 5 | 4 of 5 |
| Data analysis | 4 of 5 | 5 of 5 |
| Customer support | 5 of 5 | 3 of 5 |
| Translation | 5 of 5 | 4 of 5 |
| SQL | 5 of 5 | 5 of 5 |
| RAG and answering from documents | 5 of 5 | 5 of 5 |
| Agents and tool use | 5 of 5 | 5 of 5 |
| All jobs | 46 of 50 · 1 message each · $0.0041 a reply | 45 of 50 · 1 message each · $0.0026 a reply |
The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.
Where they differed most, customer support: both replies to a prompt only one of them passed.
An everyday prompt: “A late order”, in full.
Claude Sonnet 5
Passed: Graded 4.7 of 5 on average (lowest 4).
Hi there, I'm really sorry your rain jacket hasn't arrived yet — 9 business days is well past our standard 3–5 business day delivery window, so I completely understand the frustration. I'd recommend reaching out with your order number so we can track down exactly where your package is and get it moving. In the meantime, since your order is more than 7 business days late, you're entitled to free express shipping on your next order — I'll make sure that's applied to your account. If it turns out the jacket is lost in transit, we'll get a replacement or refund sorted for you right away. Thanks for your patience, and sorry again for the delay! Sam, Northwind support
713 tokens in, 213 out · 3.5 s · $0.0036 · 1 message on Pro · answered by anthropic/claude-sonnet-5 via Claude Platform on AWS ·
GPT-6 Sol
Failed: Graded 3.7 of 5 on average (lowest 1).
Hi, I’m sorry your rain jacket hasn’t arrived. Standard delivery usually takes 3–5 business days, so I understand your concern after 9 business days. Could you send me your order number? I’ll check the tracking and find out where the package is. Sam, Northwind support
511 tokens in, 144 out (78 of them reasoning) · 3.2 s · $0.0025 · 1 message on Pro · answered by openai/gpt-6-sol via OpenAI ·
Claude Sonnet 5 and GPT-6 Sol on every plan
Whether the one-time free trial reaches each model, then each paid plan's messages on it.
| Plan | Price | Claude Sonnet 5 | GPT-6 Sol |
|---|---|---|---|
| Free | $0 | In the one-time trial of 5 messages | In the one-time trial of 5 messages |
| Pro | $20 a month | Up to 125 a month | Up to 125 a month |
| Max | $50 a month | Up to 400 a month | Up to 400 a month |
| Ultra | $100 a month | Up to 900 a month | Up to 900 a month |
| Studio | $200 a month | Up to 2,000 a month | Up to 2,000 a month |
Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
What differs
Messages on Pro
They cost the same in llmwise: up to 125 messages a month on Pro on either.
Context window
Claude Sonnet 5 takes up to 1M tokens; GPT-6 Sol up to 1.05M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Images and PDFs
Both read images. Both take a PDF as the whole file, pages and all.
On Free
Both are in the free trial.
Where messages go
Claude Sonnet 5: Sent to Anthropic directly. GPT-6 Sol: Sent to OpenAI directly.
Fact by fact
| Fact | Claude Sonnet 5 | GPT-6 Sol |
|---|---|---|
| Context window | 1M tokens | 1.05M tokens |
| Reads images | Yes | Yes |
| PDFs | Whole file | Whole file |
| Reasoning | Yes | Yes |
| API price (September 2026) | $2.00 in / $10.00 out per million tokens | $2.00 in / $10.00 out per million tokens |
| A typical message at API prices (4,000 tokens in, 700 out) | $0.0150 | $0.0150 |
| A $10 top-up adds | 100 messages | 100 messages |
| Where a message goes | Sent to Anthropic directly. | Sent to OpenAI directly. |
| If the provider fails | If Anthropic fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to Claude Sonnet 5 through OpenRouter instead. | If OpenAI fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to GPT-6 Sol through OpenRouter instead. |
| Anthropic's safety fallback | Doesn't apply | Doesn't apply |
Claude Sonnet 5 or GPT-6 Sol?
From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.
On the facts llmwise keeps, Claude Sonnet 5 and GPT-6 Sol are level: same count, same files, same context. Pick the one whose answers you prefer.
Each model's page, the families, and other pairs
Claude Sonnet 5 vs GPT-6 Sol is one pair of models. The page below covers the whole families.
- Claude Sonnet 5: price, limits and messages on every plan
- GPT-6 Sol: price, limits and messages on every plan
- Claude vs ChatGPT
- Claude Opus 5.5 vs Claude Sonnet 5
- Claude Fable 5.1 vs Claude Sonnet 5
- Claude Sonnet 5 vs Claude Haiku 4.5
- Claude Opus 5.5 vs GPT-6 Sol
- GPT-6 Astra vs GPT-6 Sol
- GPT-6 Sol vs GPT-6 Luna
- Every model-vs-model page
Claude Sonnet 5 is a Claude model; GPT-6 Sol is a GPT model.
Questions
Is Claude Sonnet 5 or GPT-6 Sol cheaper in llmwise?
They cost the same in llmwise: up to 125 messages a month on Pro on either. Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.
Can I try Claude Sonnet 5 and GPT-6 Sol for free?
Yes: both are in the free trial of 5 messages.
Which has the bigger context window, Claude Sonnet 5 or GPT-6 Sol?
GPT-6 Sol: 1.05M tokens, against 1M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Can I use Claude Sonnet 5 and GPT-6 Sol in the same chat?
Yes. Pick Claude Sonnet 5 for one message and GPT-6 Sol for the next; the second sees the whole chat, including the first one's answer.
Which did better in your test runs, Claude Sonnet 5 or GPT-6 Sol?
On the same 50 prompts, run on September 27, 2026, Claude Sonnet 5 passed 46 and GPT-6 Sol passed 45. The table on this page has each job, and every reply is published.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.