Skip to content

Model vs model

Claude Sonnet 5.5 vs Gemini 3.1 Pro (preview)

Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.

Model prices and specs checked against OpenRouter's Claude Sonnet 5.5 page, OpenRouter's Gemini 3.1 Pro page. Updated .

Short answer

They cost the same in llmwise: up to 125 messages a month on Pro on either. Otherwise, only Claude Sonnet 5.5 can be passed to another model by its maker's safety system. In our test runs, Claude Sonnet 5.5 passed 47 of the 50 prompts both answered and Gemini 3.1 Pro (preview) 49; 2 prompts split them, most on summarization (4 to 5).

Claude Sonnet 5.5 vs Gemini 3.1 Pro (preview), prompt by prompt

Every prompt Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.

Of the 50 prompts both answered, both passed 47, only Claude Sonnet 5.5 passed 0, only Gemini 3.1 Pro (preview) passed 2, and neither passed 1. Claude Sonnet 5.5 answered sooner on 50 of the 50 and Gemini 3.1 Pro (preview) on 0; the rest were within 10% of each other. The 50 replies cost $0.1912 on Claude Sonnet 5.5 and $0.5334 on Gemini 3.1 Pro (preview): 2.8× less on Claude Sonnet 5.5.

The 2 prompts only one of Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) passed

  • An email thread in one sentence (summarization): Gemini 3.1 Pro (preview) passed and Claude Sonnet 5.5 didn't. Claude Sonnet 5.5: Graded 4.7 of 5 on average (lowest 4); but 35 words, over the 30 allowed. Gemini 3.1 Pro (preview): Graded 4.7 of 5 on average (lowest 4).

  • A frustrated customer (customer support): Gemini 3.1 Pro (preview) passed and Claude Sonnet 5.5 didn't. Claude Sonnet 5.5: Graded 3.7 of 5 on average (lowest 2). Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3).

Job by job, the widest gaps first

  • Summarization: Claude Sonnet 5.5 passed 4 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 4.1× sooner at the median, 1.9 s against 8.0 s. Claude Sonnet 5.5 cost 2.6× less, $0.0168 against $0.0441 for the 5 replies. Claude Sonnet 5.5's replies ran 108% longer, in tokens of reply, thinking not counted.

  • Customer support: Claude Sonnet 5.5 passed 4 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 3.4× sooner at the median, 2.5 s against 8.7 s. Claude Sonnet 5.5 cost 2.3× less, $0.0232 against $0.0529 for the 5 replies. Claude Sonnet 5.5's replies ran 61% longer, in tokens of reply, thinking not counted.

  • Math: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 5.0× sooner at the median, 1.7 s against 8.3 s. Claude Sonnet 5.5 cost 4.1× less, $0.0123 against $0.0501 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 48% longer, in tokens of reply, thinking not counted.

  • Translation: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 4.5× sooner at the median, 2.3 s against 10.3 s. Claude Sonnet 5.5 cost 3.3× less, $0.0211 against $0.0699 for the 5 replies. Claude Sonnet 5.5's replies ran 170% longer, in tokens of reply, thinking not counted.

  • Data analysis: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 3.8× sooner at the median, 2.3 s against 8.6 s. Claude Sonnet 5.5 cost 3.3× less, $0.0235 against $0.0773 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 16% longer, in tokens of reply, thinking not counted.

  • Coding: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 5.2× sooner at the median, 2.0 s against 10.3 s. Claude Sonnet 5.5 cost 3.1× less, $0.0313 against $0.0973 for the 5 replies. Claude Sonnet 5.5's replies ran 48% longer, in tokens of reply, thinking not counted.

  • Writing: Claude Sonnet 5.5 passed 4 of 5 and Gemini 3.1 Pro (preview) 4 of 5. Claude Sonnet 5.5 answered 3.9× sooner at the median, 2.3 s against 9.1 s. Claude Sonnet 5.5 cost 2.7× less, $0.0171 against $0.0459 for the 5 replies. Claude Sonnet 5.5's replies ran 70% longer, in tokens of reply, thinking not counted.

  • Agents and tool use: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 3.8× sooner at the median, 1.4 s against 5.3 s. Claude Sonnet 5.5 cost 2.2× less, $0.0129 against $0.0284 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 18% longer, in tokens of reply, thinking not counted.

  • RAG and answering from documents: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 4.1× sooner at the median, 1.4 s against 5.6 s. Claude Sonnet 5.5 cost 2.1× less, $0.0140 against $0.0290 for the 5 replies. Claude Sonnet 5.5's replies ran 165% longer, in tokens of reply, thinking not counted.

  • SQL: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 4.2× sooner at the median, 1.8 s against 7.4 s. Claude Sonnet 5.5 cost 2.0× less, $0.0190 against $0.0386 for the 5 replies. Claude Sonnet 5.5's replies ran 124% longer, in tokens of reply, thinking not counted.

All 50 prompts: who passed, who answered sooner, who cost less
Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) on each prompt of our test runs
PromptResultSoonerCheaper
Turn a title into a URL slugBoth passedClaude Sonnet 5.5, 3.3×took 1.8 s and 6.0 sClaude Sonnet 5.5, 2.0×cost $0.0028 and $0.0056
Parse a duration like “1h 30m”Both passedClaude Sonnet 5.5, 5.2×took 2.0 s and 10.3 sClaude Sonnet 5.5, 4.7×cost $0.0036 and $0.0172
Merge overlapping intervalsBoth passedClaude Sonnet 5.5, 6.2×took 1.4 s and 8.6 sClaude Sonnet 5.5, 3.1×cost $0.0033 and $0.0103
Evaluate an arithmetic expression, no evalBoth passedClaude Sonnet 5.5, 4.3×took 4.8 s and 20.6 sClaude Sonnet 5.5, 3.5×cost $0.0102 and $0.0351
Parse CSV with quoted fieldsBoth passedClaude Sonnet 5.5, 3.9×took 5.3 s and 20.6 sClaude Sonnet 5.5, 2.5×cost $0.0114 and $0.0291
Announce a second bakery shop on LinkedInBoth passedClaude Sonnet 5.5, 2.9×took 3.3 s and 9.5 sClaude Sonnet 5.5, 3.2×cost $0.0035 and $0.0113
Rewrite corporate jargon in plain wordsBoth passedClaude Sonnet 5.5, 5.5×took 1.6 s and 8.8 sClaude Sonnet 5.5, 4.0×cost $0.0025 and $0.0101
Decline a meeting and offer two timesBoth passedClaude Sonnet 5.5, 2.7×took 2.3 s and 6.2 sClaude Sonnet 5.5, 1.9×cost $0.0030 and $0.0058
A product announcement with five rulesBoth passedClaude Sonnet 5.5, 3.9×took 2.3 s and 9.2 sClaude Sonnet 5.5, 3.2×cost $0.0030 and $0.0095
Argue both sides of free busesNeither passedClaude Sonnet 5.5, 2.1×took 4.4 s and 9.1 sClaude Sonnet 5.5, 1.8×cost $0.0051 and $0.0092
A discount, then sales taxBoth passedClaude Sonnet 5.5, 2.3×took 1.6 s and 3.6 sClaude Sonnet 5.5, 2.4×cost $0.0017 and $0.0042
Pens at 3 for $4Both passedClaude Sonnet 5.5, 4.4×took 2.2 s and 9.6 sClaude Sonnet 5.5, 3.4×cost $0.0036 and $0.0121
Compound interest over three yearsBoth passedClaude Sonnet 5.5, 4.9×took 1.3 s and 6.2 sClaude Sonnet 5.5, 4.6×cost $0.0021 and $0.0099
Four-digit numbers whose digits sum to 9Both passedClaude Sonnet 5.5, 5.4×took 1.8 s and 9.6 sClaude Sonnet 5.5, 5.7×cost $0.0026 and $0.0150
The highest of three dice is a 5Both passedClaude Sonnet 5.5, 5.0×took 1.7 s and 8.3 sClaude Sonnet 5.5, 4.0×cost $0.0022 and $0.0090
An article in three bulletsBoth passedClaude Sonnet 5.5, 4.1×took 1.9 s and 8.0 sClaude Sonnet 5.5, 2.8×cost $0.0031 and $0.0088
An email thread in one sentenceOnly Gemini 3.1 Pro (preview)Claude Sonnet 5.5, 4.4×took 1.4 s and 6.3 sClaude Sonnet 5.5, 2.5×cost $0.0023 and $0.0058
Decisions and action items from a meetingBoth passedClaude Sonnet 5.5, 3.5×took 2.2 s and 7.6 sClaude Sonnet 5.5, 1.9×cost $0.0043 and $0.0081
A quarterly memo for the CEOBoth passedClaude Sonnet 5.5, 6.1×took 1.8 s and 11.2 sClaude Sonnet 5.5, 3.8×cost $0.0037 and $0.0139
A study with a negative resultBoth passedClaude Sonnet 5.5, 3.4×took 2.4 s and 8.2 sClaude Sonnet 5.5, 2.2×cost $0.0034 and $0.0075
The region with the most revenueBoth passedClaude Sonnet 5.5, 3.5×took 2.4 s and 8.5 sClaude Sonnet 5.5, 2.3×cost $0.0051 and $0.0117
Average order value in AugustBoth passedClaude Sonnet 5.5, 5.0×took 1.7 s and 8.6 sClaude Sonnet 5.5, 3.5×cost $0.0037 and $0.0130
Revenue change from July to AugustBoth passedClaude Sonnet 5.5, 3.9×took 2.3 s and 9.0 sClaude Sonnet 5.5, 2.8×cost $0.0048 and $0.0135
A median, filtered two waysBoth passedClaude Sonnet 5.5, 4.6×took 1.5 s and 7.1 sClaude Sonnet 5.5, 2.8×cost $0.0030 and $0.0082
Correlation between ad spend and sign-upsBoth passedClaude Sonnet 5.5, 3.9×took 4.2 s and 16.5 sClaude Sonnet 5.5, 4.4×cost $0.0070 and $0.0310
A late orderBoth passedClaude Sonnet 5.5, 3.4×took 2.5 s and 8.7 sClaude Sonnet 5.5, 3.3×cost $0.0038 and $0.0125
A return inside the windowBoth passedClaude Sonnet 5.5, 3.2×took 2.2 s and 6.9 sClaude Sonnet 5.5, 2.1×cost $0.0035 and $0.0076
A frustrated customerOnly Gemini 3.1 Pro (preview)Claude Sonnet 5.5, 3.6×took 2.7 s and 9.8 sClaude Sonnet 5.5, 2.6×cost $0.0042 and $0.0110
A refund request outside the windowBoth passedClaude Sonnet 5.5, 3.7×took 2.0 s and 7.4 sClaude Sonnet 5.5, 2.7×cost $0.0038 and $0.0103
A message with a planted instructionBoth passedClaude Sonnet 5.5, 1.7×took 6.1 s and 10.6 sClaude Sonnet 5.5, 1.5×cost $0.0079 and $0.0116
A delivery message into SpanishBoth passedClaude Sonnet 5.5, 7.8×took 1.5 s and 11.8 sClaude Sonnet 5.5, 5.0×cost $0.0032 and $0.0157
A product description into FrenchBoth passedClaude Sonnet 5.5, 5.0×took 1.8 s and 9.1 sClaude Sonnet 5.5, 4.5×cost $0.0032 and $0.0141
A meeting note into GermanBoth passedClaude Sonnet 5.5, 5.7×took 2.3 s and 13.0 sClaude Sonnet 5.5, 4.1×cost $0.0038 and $0.0158
Idioms into natural JapaneseBoth passedClaude Sonnet 5.5, 3.6×took 2.9 s and 10.3 sClaude Sonnet 5.5, 2.5×cost $0.0046 and $0.0118
A lease clause into Brazilian PortugueseBoth passedClaude Sonnet 5.5, 1.8×took 4.4 s and 8.1 sClaude Sonnet 5.5, 2.0×cost $0.0063 and $0.0125
Customers in one countryBoth passedClaude Sonnet 5.5, 2.8×took 1.8 s and 4.9 sClaude Sonnet 5.5, 1.2×cost $0.0024 and $0.0029
Count orders by statusBoth passedClaude Sonnet 5.5, 5.2×took 1.0 s and 5.0 sClaude Sonnet 5.5, 1.6×cost $0.0024 and $0.0039
Revenue by categoryBoth passedClaude Sonnet 5.5, 4.5×took 1.6 s and 7.4 sClaude Sonnet 5.5, 1.9×cost $0.0039 and $0.0074
Every customer, even those without ordersBoth passedClaude Sonnet 5.5, 3.5×took 2.3 s and 7.9 sClaude Sonnet 5.5, 2.1×cost $0.0043 and $0.0089
Monthly revenue with a running totalBoth passedClaude Sonnet 5.5, 4.2×took 2.7 s and 11.1 sClaude Sonnet 5.5, 2.6×cost $0.0060 and $0.0155
A fact from one sectionBoth passedClaude Sonnet 5.5, 3.9×took 1.4 s and 5.3 sClaude Sonnet 5.5, 1.7×cost $0.0027 and $0.0045
Core hours and start timesBoth passedClaude Sonnet 5.5, 4.3×took 1.4 s and 5.8 sClaude Sonnet 5.5, 1.9×cost $0.0027 and $0.0052
Two sections in one answerBoth passedClaude Sonnet 5.5, 4.2×took 1.3 s and 5.6 sClaude Sonnet 5.5, 1.8×cost $0.0027 and $0.0047
A later amendment changes the answerBoth passedClaude Sonnet 5.5, 7.0×took 1.2 s and 8.6 sClaude Sonnet 5.5, 3.2×cost $0.0031 and $0.0099
A question the handbook doesn't answerBoth passedClaude Sonnet 5.5, 4.1×took 1.4 s and 5.6 sClaude Sonnet 5.5, 1.6×cost $0.0028 and $0.0046
Pick the tool and work out the dateBoth passedClaude Sonnet 5.5, 3.4×took 1.3 s and 4.3 sClaude Sonnet 5.5, 2.3×cost $0.0021 and $0.0048
Convert a currencyBoth passedClaude Sonnet 5.5, 5.2×took 1.0 s and 5.3 sClaude Sonnet 5.5, 2.1×cost $0.0021 and $0.0044
Book a meeting from a sentenceBoth passedClaude Sonnet 5.5, 4.8×took 1.5 s and 7.1 sClaude Sonnet 5.5, 2.6×cost $0.0026 and $0.0068
Search, but don't bookBoth passedClaude Sonnet 5.5, 3.3×took 1.4 s and 4.7 sClaude Sonnet 5.5, 2.6×cost $0.0023 and $0.0061
Two calls with a unit conversionBoth passedClaude Sonnet 5.5, 1.8×took 3.7 s and 6.9 sClaude Sonnet 5.5, 1.7×cost $0.0038 and $0.0063

Claude Sonnet 5.5 vs Gemini 3.1 Pro (preview) in our test runs

Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) on the same prompts, job by job: how many replies passed their check.

Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) in our test runs, job by job
JobClaude Sonnet 5.5Gemini 3.1 Pro (preview)
Coding5 of 55 of 5
Writing4 of 54 of 5
Math5 of 55 of 5
Summarization4 of 55 of 5
Data analysis5 of 55 of 5
Customer support4 of 55 of 5
Translation5 of 55 of 5
SQL5 of 55 of 5
RAG and answering from documents5 of 55 of 5
Agents and tool use5 of 55 of 5
All jobs47 of 50 · 1 message each · $0.0038 a reply49 of 50 · 1 message each · $0.0107 a reply

The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.

Where they differed most, summarization: both replies to a prompt only one of them passed.

An everyday prompt: “An email thread in one sentence”, in full.

  • Claude Sonnet 5.5

    Failed: Graded 4.7 of 5 on average (lowest 4); but 35 words, over the 30 allowed.

    Due to a paper shortage, the brochures won't ship until the 18th, so Rosa will pick them up that day in a van Idris booked (8:00–12:00), in time for the trade fair on the 20th.

    656 tokens in, 69 out · 1.4 s · $0.0023 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·

  • Gemini 3.1 Pro

    Passed: Graded 4.7 of 5 on average (lowest 4).

    After a printing delay, Idris booked a van so Rosa can pick up the trade fair brochures on the 18th.

    529 tokens in, 397 out (372 of them reasoning) · 6.3 s · $0.0058 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) on every plan

Whether the one-time free trial reaches each model, then each paid plan's messages on it.

Messages on Claude Sonnet 5.5 and Gemini 3.1 Pro (preview), plan by plan
PlanPriceClaude Sonnet 5.5Gemini 3.1 Pro (preview)
Free$0In the one-time trial of 5 messagesIn the one-time trial of 5 messages
Pro$20 a monthUp to 125 a monthUp to 125 a month
Max$50 a monthUp to 400 a monthUp to 400 a month
Ultra$100 a monthUp to 900 a monthUp to 900 a month
Studio$200 a monthUp to 2,000 a monthUp to 2,000 a month

Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

What differs

  • Messages on Pro

    They cost the same in llmwise: up to 125 messages a month on Pro on either.

  • Context window

    Claude Sonnet 5.5 takes up to 1M tokens; Gemini 3.1 Pro (preview) up to 1.05M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

  • Images and PDFs

    Both read images. Both take a PDF as the whole file, pages and all.

  • On Free

    Both are in the free trial.

  • Where messages go

    Claude Sonnet 5.5: Sent to Anthropic directly. Gemini 3.1 Pro (preview): Sent to Google directly.

Fact by fact

Claude Sonnet 5.5 and Gemini 3.1 Pro (preview), fact by fact
FactClaude Sonnet 5.5Gemini 3.1 Pro (preview)
Context window1M tokens1.05M tokens
Reads imagesYesYes
PDFsWhole fileWhole file
ReasoningYesYes
API price (September 2026)$2.00 in / $10.00 out per million tokens$2.00 in / $12.00 out per million tokens
A typical message at API prices (4,000 tokens in, 700 out)$0.0150$0.0164
A $10 top-up adds100 messages100 messages
Where a message goesSent to Anthropic directly.Sent to Google directly.
If the provider failsIf Anthropic fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to Claude Sonnet 5.5 through OpenRouter instead.If Google fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to Gemini 3.1 Pro (preview) through OpenRouter instead.
Anthropic's safety fallbackAnthropic's safety system can pass a Claude Sonnet 5.5 message to another Claude model. The reply then names the model that answered and says what you were charged for.Doesn't apply
API prices are what our model catalog lists (Claude: Anthropic's list price; Gemini: Google's list price). In llmwise you pay per message, not per token: the counts above are what you get.

Claude Sonnet 5.5 or Gemini 3.1 Pro (preview)?

From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.

  • Pick Claude Sonnet 5.5: it costs its maker less to run ($0.0150 a typical message at API prices), though in llmwise the count is the same.

Each model's page, the families, and other pairs

Claude Sonnet 5.5 vs Gemini 3.1 Pro (preview) is one pair of models. The page below covers the whole families.

Questions

Is Claude Sonnet 5.5 or Gemini 3.1 Pro (preview) cheaper in llmwise?

They cost the same in llmwise: up to 125 messages a month on Pro on either. Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.

Can I try Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) for free?

Yes: both are in the free trial of 5 messages.

Which has the bigger context window, Claude Sonnet 5.5 or Gemini 3.1 Pro (preview)?

Gemini 3.1 Pro (preview): 1.05M tokens, against 1M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

Can I use Claude Sonnet 5.5 and Gemini 3.1 Pro (preview) in the same chat?

Yes. Pick Claude Sonnet 5.5 for one message and Gemini 3.1 Pro (preview) for the next; the second sees the whole chat, including the first one's answer.

Which did better in your test runs, Claude Sonnet 5.5 or Gemini 3.1 Pro (preview)?

On the same 50 prompts, run on September 28, 2026, Claude Sonnet 5.5 passed 47 and Gemini 3.1 Pro (preview) passed 49. The table on this page has each job, and every reply is published.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.