Skip to content

Model vs model

Claude Haiku 5.5 vs GPT-6 Luna

Claude Haiku 5.5 and GPT-6 Luna are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.

Model prices and specs checked against OpenRouter's Claude Haiku 5.5 page, OpenRouter's GPT-6 Luna page. Updated .

Short answer

Both are everyday models: each message comes from Pro's 60 a day, so they cost the same. Beyond that, they read the same files and neither is easier to try. In our test runs, Claude Haiku 5.5 passed 45 of the 50 prompts both answered and GPT-6 Luna 47; 4 prompts split them, most on writing (3 to 5).

Claude Haiku 5.5 vs GPT-6 Luna, prompt by prompt

Every prompt Claude Haiku 5.5 and GPT-6 Luna both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.

Of the 50 prompts both answered, both passed 44, only Claude Haiku 5.5 passed 1, only GPT-6 Luna passed 3, and neither passed 2. Claude Haiku 5.5 answered sooner on 17 of the 50 and GPT-6 Luna on 20; the rest were within 10% of each other. The 50 replies cost $0.0118 on Claude Haiku 5.5 and $0.0059 on GPT-6 Luna: 2.0× less on GPT-6 Luna.

The 4 prompts only one of Claude Haiku 5.5 and GPT-6 Luna passed

  • Rewrite corporate jargon in plain words (writing): GPT-6 Luna passed and Claude Haiku 5.5 didn't. Claude Haiku 5.5: Graded 3.3 of 5 on average (lowest 2). GPT-6 Luna: Graded 4.7 of 5 on average (lowest 4).

  • Argue both sides of free buses (writing): GPT-6 Luna passed and Claude Haiku 5.5 didn't. Claude Haiku 5.5: Graded 4.0 of 5 on average (lowest 4); but a paragraph of 103 words, over the 90 allowed. GPT-6 Luna: Graded 4.3 of 5 on average (lowest 4).

  • Pens at 3 for $4 (math): Claude Haiku 5.5 passed and GPT-6 Luna didn't. Claude Haiku 5.5: Final answer Buy 3 packs of 3 and 1 single pen, costing 13.50: right. GPT-6 Luna: Final answer 2 bundles and 4 singles for 12; expected 13.5.

  • An email thread in one sentence (summarization): GPT-6 Luna passed and Claude Haiku 5.5 didn't. Claude Haiku 5.5: Graded 4.3 of 5 on average (lowest 4); but 39 words, over the 30 allowed. GPT-6 Luna: Graded 4.3 of 5 on average (lowest 3).

Job by job, the widest gaps first

  • Writing: Claude Haiku 5.5 passed 3 of 5 and GPT-6 Luna 5 of 5. Claude Haiku 5.5 answered 1.6× sooner at the median, 1.5 s against 2.4 s. GPT-6 Luna cost 1.6× less, $0.0008 against $0.0005 for the 5 replies. Claude Haiku 5.5's replies ran 92% longer, in tokens of reply, thinking not counted.

  • Summarization: Claude Haiku 5.5 passed 3 of 5 and GPT-6 Luna 4 of 5. Their median waits were close, 1.4 s against 1.5 s. GPT-6 Luna cost 3.4× less, $0.0016 against $0.0005 for the 5 replies. Claude Haiku 5.5's replies ran 68% longer, in tokens of reply, thinking not counted.

  • Math: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 4 of 5. Claude Haiku 5.5 answered 1.3× sooner at the median, 1.7 s against 2.2 s. GPT-6 Luna cost 1.5× less, $0.0007 against $0.0005 for the 5 replies. Claude Haiku 5.5's replies ran 166% longer, in tokens of reply, thinking not counted.

  • Customer support: Claude Haiku 5.5 passed 4 of 5 and GPT-6 Luna 4 of 5. GPT-6 Luna answered 1.4× sooner at the median, 1.9 s against 1.4 s. GPT-6 Luna cost 2.7× less, $0.0012 against $0.0004 for the 5 replies. Claude Haiku 5.5's replies ran 213% longer, in tokens of reply, thinking not counted.

  • SQL: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 5 of 5. Their median waits were close, 1.3 s against 1.2 s. GPT-6 Luna cost 2.5× less, $0.0011 against $0.0004 for the 5 replies. Claude Haiku 5.5's replies ran 72% longer, in tokens of reply, thinking not counted.

  • Data analysis: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 5 of 5. GPT-6 Luna answered 1.3× sooner at the median, 3.3 s against 2.5 s. GPT-6 Luna cost 2.0× less, $0.0018 against $0.0009 for the 5 replies. Claude Haiku 5.5's replies ran 157% longer, in tokens of reply, thinking not counted.

  • Translation: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 5 of 5. Their median waits were close, 1.4 s against 1.6 s. GPT-6 Luna cost 1.9× less, $0.0009 against $0.0005 for the 5 replies. Claude Haiku 5.5's replies ran 115% longer, in tokens of reply, thinking not counted.

  • RAG and answering from documents: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 5 of 5. Claude Haiku 5.5 answered 1.2× sooner at the median, 1.1 s against 1.3 s. GPT-6 Luna cost 1.8× less, $0.0007 against $0.0004 for the 5 replies. Claude Haiku 5.5's replies ran 139% longer, in tokens of reply, thinking not counted.

  • Agents and tool use: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 5 of 5. Claude Haiku 5.5 answered 1.8× sooner at the median, 1.0 s against 1.9 s. GPT-6 Luna cost 1.7× less, $0.0007 against $0.0004 for the 5 replies. Claude Haiku 5.5's replies ran 59% longer, in tokens of reply, thinking not counted.

  • Coding: Claude Haiku 5.5 passed 5 of 5 and GPT-6 Luna 5 of 5. Claude Haiku 5.5 answered 1.2× sooner at the median, 3.9 s against 4.6 s. GPT-6 Luna cost 1.6× less, $0.0023 against $0.0014 for the 5 replies. Claude Haiku 5.5's replies ran 71% longer, in tokens of reply, thinking not counted.

All 50 prompts: who passed, who answered sooner, who cost less
Claude Haiku 5.5 and GPT-6 Luna on each prompt of our test runs
PromptResultSoonerCheaper
Turn a title into a URL slugBoth passedClaude Haiku 5.5, 1.6×took 1.5 s and 2.3 sGPT-6 Luna, 1.6×cost $0.0001 and $0.0001
Parse a duration like “1h 30m”Both passedGPT-6 Luna, 1.2×took 5.4 s and 4.6 sGPT-6 Luna, 2.0×cost $0.0005 and $0.0003
Merge overlapping intervalsBoth passedClaude Haiku 5.5, 1.1×took 1.3 s and 1.5 sGPT-6 Luna, 1.5×cost $0.0002 and $0.0001
Evaluate an arithmetic expression, no evalBoth passedClaude Haiku 5.5, 2.2×took 3.9 s and 8.8 sGPT-6 Luna, 1.1×cost $0.0006 and $0.0006
Parse CSV with quoted fieldsBoth passedGPT-6 Luna, 1.1×took 8.1 s and 7.2 sGPT-6 Luna, 1.9×cost $0.0008 and $0.0004
Announce a second bakery shop on LinkedInBoth passedGPT-6 Luna, 1.2×took 3.4 s and 2.9 sGPT-6 Luna, 1.4×cost $0.0002 and $0.0001
Rewrite corporate jargon in plain wordsOnly GPT-6 LunaClaude Haiku 5.5, 1.9×took 1.3 s and 2.4 sGPT-6 Luna, 1.2×cost $0.0001 and $0.0001
Decline a meeting and offer two timesBoth passedClosetook 1.3 s and 1.3 sGPT-6 Luna, 1.9×cost $0.0001 and $0.0001
A product announcement with five rulesBoth passedClosetook 1.5 s and 1.5 sGPT-6 Luna, 1.6×cost $0.0001 and $0.0001
Argue both sides of free busesOnly GPT-6 LunaGPT-6 Luna, 1.3×took 3.3 s and 2.5 sGPT-6 Luna, 2.0×cost $0.0003 and $0.0001
A discount, then sales taxBoth passedClaude Haiku 5.5, 2.5×took 1.3 s and 3.1 sGPT-6 Luna, 1.1×cost $0.0001 and $0.0001
Pens at 3 for $4Only Claude Haiku 5.5Claude Haiku 5.5, 1.6×took 1.7 s and 2.9 sGPT-6 Luna, 1.2×cost $0.0002 and $0.0001
Compound interest over three yearsBoth passedClaude Haiku 5.5, 1.6×took 1.3 s and 2.0 sGPT-6 Luna, 1.3×cost $0.0001 and $0.0001
Four-digit numbers whose digits sum to 9Both passedClosetook 1.9 s and 2.0 sGPT-6 Luna, 2.0×cost $0.0002 and $0.0001
The highest of three dice is a 5Both passedClosetook 2.1 s and 2.2 sGPT-6 Luna, 1.8×cost $0.0002 and $0.0001
An article in three bulletsNeither passedGPT-6 Luna, 4.9×took 7.2 s and 1.5 sGPT-6 Luna, 8.4×cost $0.0008 and $0.0001
An email thread in one sentenceOnly GPT-6 LunaClosetook 1.1 s and 1.1 sGPT-6 Luna, 1.8×cost $0.0001 and $0.0001
Decisions and action items from a meetingBoth passedClosetook 1.3 s and 1.5 sGPT-6 Luna, 1.8×cost $0.0002 and $0.0001
A quarterly memo for the CEOBoth passedGPT-6 Luna, 2.7×took 4.2 s and 1.5 sGPT-6 Luna, 2.7×cost $0.0003 and $0.0001
A study with a negative resultBoth passedClosetook 1.4 s and 1.3 sGPT-6 Luna, 1.8×cost $0.0001 and $0.0001
The region with the most revenueBoth passedGPT-6 Luna, 1.4×took 3.4 s and 2.5 sGPT-6 Luna, 2.0×cost $0.0003 and $0.0001
Average order value in AugustBoth passedClaude Haiku 5.5, 2.1×took 1.3 s and 2.6 sGPT-6 Luna, 1.6×cost $0.0002 and $0.0001
Revenue change from July to AugustBoth passedGPT-6 Luna, 1.3×took 3.3 s and 2.5 sGPT-6 Luna, 2.9×cost $0.0004 and $0.0001
A median, filtered two waysBoth passedClosetook 2.6 s and 2.5 sGPT-6 Luna, 1.8×cost $0.0002 and $0.0001
Correlation between ad spend and sign-upsBoth passedClaude Haiku 5.5, 1.2×took 5.7 s and 6.7 sGPT-6 Luna, 2.0×cost $0.0007 and $0.0004
A late orderBoth passedGPT-6 Luna, 1.1×took 1.8 s and 1.6 sGPT-6 Luna, 2.1×cost $0.0002 and $0.0001
A return inside the windowBoth passedGPT-6 Luna, 1.1×took 1.4 s and 1.2 sGPT-6 Luna, 2.0×cost $0.0002 and $0.0001
A frustrated customerNeither passedGPT-6 Luna, 1.5×took 2.1 s and 1.4 sGPT-6 Luna, 2.5×cost $0.0002 and $0.0001
A refund request outside the windowBoth passedGPT-6 Luna, 1.5×took 1.9 s and 1.3 sGPT-6 Luna, 2.5×cost $0.0002 and $0.0001
A message with a planted instructionBoth passedGPT-6 Luna, 3.0×took 4.6 s and 1.5 sGPT-6 Luna, 4.2×cost $0.0004 and $0.0001
A delivery message into SpanishBoth passedClaude Haiku 5.5, 1.9×took 1.4 s and 2.6 sGPT-6 Luna, 1.7×cost $0.0002 and $0.0001
A product description into FrenchBoth passedClosetook 1.4 s and 1.4 sGPT-6 Luna, 1.8×cost $0.0002 and $0.0001
A meeting note into GermanBoth passedClaude Haiku 5.5, 1.3×took 1.3 s and 1.7 sGPT-6 Luna, 1.9×cost $0.0002 and $0.0001
Idioms into natural JapaneseBoth passedGPT-6 Luna, 1.3×took 2.0 s and 1.6 sGPT-6 Luna, 1.7×cost $0.0002 and $0.0001
A lease clause into Brazilian PortugueseBoth passedGPT-6 Luna, 1.5×took 2.3 s and 1.5 sGPT-6 Luna, 2.6×cost $0.0003 and $0.0001
Customers in one countryBoth passedClosetook 1.0 s and 1.0 sGPT-6 Luna, 2.0×cost $0.0001 and $0.0001
Count orders by statusBoth passedGPT-6 Luna, 1.1×took 1.0 s and 0.9 sGPT-6 Luna, 2.1×cost $0.0001 and $0.0001
Revenue by categoryBoth passedClosetook 1.3 s and 1.2 sGPT-6 Luna, 1.9×cost $0.0002 and $0.0001
Every customer, even those without ordersBoth passedClaude Haiku 5.5, 1.9×took 1.3 s and 2.4 sGPT-6 Luna, 1.9×cost $0.0002 and $0.0001
Monthly revenue with a running totalBoth passedGPT-6 Luna, 1.8×took 3.2 s and 1.8 sGPT-6 Luna, 3.7×cost $0.0005 and $0.0001
A fact from one sectionBoth passedClaude Haiku 5.5, 1.3×took 1.0 s and 1.3 sGPT-6 Luna, 1.8×cost $0.0001 and $0.0001
Core hours and start timesBoth passedClosetook 1.0 s and 1.0 sGPT-6 Luna, 1.6×cost $0.0001 and $0.0001
Two sections in one answerBoth passedClosetook 1.4 s and 1.4 sGPT-6 Luna, 1.8×cost $0.0001 and $0.0001
A later amendment changes the answerBoth passedGPT-6 Luna, 1.3×took 1.3 s and 1.0 sGPT-6 Luna, 2.0×cost $0.0002 and $0.0001
A question the handbook doesn't answerBoth passedClaude Haiku 5.5, 1.8×took 1.1 s and 1.9 sGPT-6 Luna, 2.0×cost $0.0001 and $0.0001
Pick the tool and work out the dateBoth passedClaude Haiku 5.5, 2.0×took 1.0 s and 2.1 sGPT-6 Luna, 1.4×cost $0.0001 and $0.0001
Convert a currencyBoth passedClaude Haiku 5.5, 1.4×took 0.8 s and 1.1 sGPT-6 Luna, 1.6×cost $0.0001 and $0.0001
Book a meeting from a sentenceBoth passedGPT-6 Luna, 2.4×took 3.0 s and 1.2 sGPT-6 Luna, 2.2×cost $0.0002 and $0.0001
Search, but don't bookBoth passedClaude Haiku 5.5, 2.2×took 1.0 s and 2.2 sGPT-6 Luna, 1.4×cost $0.0001 and $0.0001
Two calls with a unit conversionBoth passedGPT-6 Luna, 1.2×took 2.3 s and 1.9 sGPT-6 Luna, 1.9×cost $0.0002 and $0.0001

Claude Haiku 5.5 vs GPT-6 Luna in our test runs

Claude Haiku 5.5 and GPT-6 Luna on the same prompts, job by job: how many replies passed their check.

Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Claude Haiku 5.5 and GPT-6 Luna in our test runs, job by job
JobClaude Haiku 5.5GPT-6 Luna
Coding5 of 55 of 5
Writing3 of 55 of 5
Math5 of 54 of 5
Summarization3 of 54 of 5
Data analysis5 of 55 of 5
Customer support4 of 54 of 5
Translation5 of 55 of 5
SQL5 of 55 of 5
RAG and answering from documents5 of 55 of 5
Agents and tool use5 of 55 of 5
All jobs45 of 50 · 1 message each · $0.0002 a reply47 of 50 · 1 message each · $0.0001 a reply

The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.

Where they differed most, writing: both replies to a prompt only one of them passed.

An everyday prompt: “Rewrite corporate jargon in plain words”, in full.

  • Claude Haiku 5.5

    Failed: Graded 3.3 of 5 on average (lowest 2).

    Through teamwork across departments, our Q3 initiative shifted our approach to focus more on customers. Customer retention rose 12%, and customer loss slowed by 30 basis points across our enterprise segment. We also adjusted our multichannel customer contact points to use resources more efficiently.

    600 tokens in, 95 out · 1.3 s · $0.0001 · 1 message on Pro · answered by anthropic/claude-haiku-5.5 via Anthropic ·

  • GPT-6 Luna

    Passed: Graded 4.7 of 5 on average (lowest 4).

    In Q3, teams across the company worked together to focus more on customers. Retention measures improved by 12%, and the rate of customer loss fell by 0.3 percentage points in the enterprise segment. The company also adjusted its customer contact channels to make better use of available capacity.

    417 tokens in, 124 out (59 of them reasoning) · 2.4 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·

Claude Haiku 5.5 and GPT-6 Luna on every plan

Whether the one-time free trial reaches each model, then each paid plan's messages on it.

Messages on Claude Haiku 5.5 and GPT-6 Luna, plan by plan
PlanPriceClaude Haiku 5.5GPT-6 Luna
Free$0In the one-time trial of 5 messagesIn the one-time trial of 5 messages
Pro$20 a month60 a day60 a day
Max$50 a month120 a day120 a day
Ultra$100 a month200 a day200 a day
Studio$200 a month200 a day200 a day

Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

What differs

  • Messages on Pro

    Both are everyday models: each message comes from Pro's 60 a day, so they cost the same.

  • Context window

    Claude Haiku 5.5 takes up to 1M tokens; GPT-6 Luna up to 1.05M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

  • Images and PDFs

    Both read images. Both take a PDF as the whole file, pages and all.

  • On Free

    Both are in the free trial.

  • Where messages go

    Claude Haiku 5.5: Sent to Anthropic directly. GPT-6 Luna: Sent to OpenAI directly.

Fact by fact

Claude Haiku 5.5 and GPT-6 Luna, fact by fact
FactClaude Haiku 5.5GPT-6 Luna
Context window1M tokens1.05M tokens
Reads imagesYesYes
PDFsWhole fileWhole file
ReasoningYesYes
API price (October 2026)$0.10 in / $0.50 out per million tokens$0.10 in / $0.50 out per million tokens
A typical message at API prices (4,000 tokens in, 700 out)$0.0008$0.0008
A $10 top-up addsNothing: an everyday model's count is dailyNothing: an everyday model's count is daily
Where a message goesSent to Anthropic directly.Sent to OpenAI directly.
If the provider failsIf Anthropic fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to Claude Haiku 5.5 through OpenRouter instead.If OpenAI fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to GPT-6 Luna through OpenRouter instead.
Anthropic's safety fallbackDoesn't applyDoesn't apply
API prices are what our model catalog lists (Claude: Anthropic's list price; GPT: OpenAI's list price). In llmwise you pay per message, not per token: the counts above are what you get.

Claude Haiku 5.5 or GPT-6 Luna?

From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.

  • On the facts llmwise keeps, Claude Haiku 5.5 and GPT-6 Luna are level: same count, same files, same context. Pick the one whose answers you prefer.

Each model's page, the families, and other pairs

Claude Haiku 5.5 vs GPT-6 Luna is one pair of models. The page below covers the whole families.

Questions

Is Claude Haiku 5.5 or GPT-6 Luna cheaper in llmwise?

Both are everyday models: each message comes from Pro's 60 a day, so they cost the same. Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.

Can I try Claude Haiku 5.5 and GPT-6 Luna for free?

Yes: both are in the free trial of 5 messages.

Which has the bigger context window, Claude Haiku 5.5 or GPT-6 Luna?

GPT-6 Luna: 1.05M tokens, against 1M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

Can I use Claude Haiku 5.5 and GPT-6 Luna in the same chat?

Yes. Pick Claude Haiku 5.5 for one message and GPT-6 Luna for the next; the second sees the whole chat, including the first one's answer.

Which did better in your test runs, Claude Haiku 5.5 or GPT-6 Luna?

On the same 50 prompts, run on October 7, 2026, Claude Haiku 5.5 passed 45 and GPT-6 Luna passed 47. The table on this page has each job, and every reply is published.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.