Skip to content

Comparison

Grok vs Claude

In our test runs on October 7, 2026, the same 50 prompts across 10 jobs: Grok's one model passed 47 of 50 replies and Claude's 6 models passed 275 of 300 replies. On llmwise Pro, Grok 4.7 up to 250 messages a month and Claude Sonnet 5.5 up to 125.

Based on 350 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Short answer

Job by job, both families' best models shared the top result on 8 of the 10 jobs; Claude's alone had it on writing and customer support.

Grok vs Claude, job by job

On each job, Grok's pick against Claude's: the model of each family that passed the most of the job's 5 prompts (then the most hard ones, then the cheaper). The jobs where they differ most come first.

  • Writing: Grok 4.7 passed 3 of 5 and Claude Sonnet 5 4 of 5. Claude Sonnet 5 answered 1.3× sooner at the median, 6.0 s against 4.5 s. Claude Sonnet 5 cost 1.7× less, $0.0288 against $0.0171 for the 5 replies. Claude Sonnet 5's replies ran 97% longer, in tokens of reply, thinking not counted.

  • Customer support: Grok 4.7 passed 4 of 5 and Claude Sonnet 5 5 of 5. Claude Sonnet 5 answered 3.3× sooner at the median, 11.6 s against 3.5 s. Claude Sonnet 5 cost 1.7× less, $0.0306 against $0.0182 for the 5 replies. Claude Sonnet 5's replies ran 99% longer, in tokens of reply, thinking not counted.

  • Coding: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 8.0× sooner at the median, 31.6 s against 3.9 s. Claude Haiku 5.5 cost 54.3× less, $0.1229 against $0.0023 for the 5 replies. Claude Haiku 5.5's replies ran 42% longer, in tokens of reply, thinking not counted.

  • Math: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 6.7× sooner at the median, 11.7 s against 1.7 s. Claude Haiku 5.5 cost 49.2× less, $0.0362 against $0.0007 for the 5 replies. Claude Haiku 5.5's replies ran 33% longer, in tokens of reply, thinking not counted.

  • Summarization: Grok 4.7 passed 5 of 5 and Claude Haiku 4.5 5 of 5. Claude Haiku 4.5 answered 1.7× sooner at the median, 2.8 s against 1.7 s. Claude Haiku 4.5 cost 2.8× less, $0.0141 against $0.0051 for the 5 replies. Claude Haiku 4.5's replies ran 12% longer, in tokens of reply, thinking not counted.

  • Data analysis: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 2.7× sooner at the median, 8.9 s against 3.3 s. Claude Haiku 5.5 cost 31.0× less, $0.0556 against $0.0018 for the 5 replies. Claude Haiku 5.5's replies ran 55% longer, in tokens of reply, thinking not counted.

  • Translation: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 7.5× sooner at the median, 10.7 s against 1.4 s. Claude Haiku 5.5 cost 45.5× less, $0.0410 against $0.0009 for the 5 replies. Claude Haiku 5.5's replies ran 132% longer, in tokens of reply, thinking not counted.

  • SQL: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 2.4× sooner at the median, 3.0 s against 1.3 s. Claude Haiku 5.5 cost 21.8× less, $0.0240 against $0.0011 for the 5 replies. Claude Haiku 5.5's replies ran 89% longer, in tokens of reply, thinking not counted.

  • RAG and answering from documents: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 1.8× sooner at the median, 1.9 s against 1.1 s. Claude Haiku 5.5 cost 20.6× less, $0.0146 against $0.0007 for the 5 replies. Claude Haiku 5.5's replies ran 198% longer, in tokens of reply, thinking not counted.

  • Agents and tool use: Grok 4.7 passed 5 of 5 and Claude Haiku 5.5 5 of 5. Claude Haiku 5.5 answered 5.0× sooner at the median, 5.2 s against 1.0 s. Claude Haiku 5.5 cost 28.6× less, $0.0197 against $0.0007 for the 5 replies. Claude Haiku 5.5's replies ran 48% longer, in tokens of reply, thinking not counted.

The 4 prompts only one of Grok and Claude passed

Where one family's pick passed a prompt and the other's didn't, in each check's own words.

  • Announce a second bakery shop on LinkedIn (writing): Grok 4.7 passed and Claude Sonnet 5 didn't. Grok 4.7: Graded 4.8 of 5 on average (lowest 4). Claude Sonnet 5: Graded 3.5 of 5 on average (lowest 3).

  • Rewrite corporate jargon in plain words (writing): Claude Sonnet 5 passed and Grok 4.7 didn't. Grok 4.7: Graded 3.7 of 5 on average (lowest 3). Claude Sonnet 5: Graded 4.7 of 5 on average (lowest 4).

  • Argue both sides of free buses (writing): Claude Sonnet 5 passed and Grok 4.7 didn't. Grok 4.7: Graded 4.7 of 5 on average (lowest 4); but a paragraph of 100 words, over the 90 allowed. Claude Sonnet 5: Graded 4.0 of 5 on average (lowest 4).

  • A frustrated customer (customer support): Claude Sonnet 5 passed and Grok 4.7 didn't. Grok 4.7: Graded 3.7 of 5 on average (lowest 3). Claude Sonnet 5: Graded 4.3 of 5 on average (lowest 4).

Each Grok model against each Claude model

Every Grok model against every Claude model on the same 50 prompts: each one's passes, how many prompts split them, and who was cheaper and quicker.

  • Grok 4.7 vs Claude Fable 5.1: 47 and 45 of 50; 6 prompts split them; Grok 4.7's replies cost 2.6× less in all, and Claude Fable 5.1 answered sooner on 29, Grok 4.7 on 18. Claude Fable 5.1 vs Grok 4.7.

  • Grok 4.7 vs Claude Opus 5.5: 47 and 49 of 50; 2 prompts split them; Grok 4.7's replies cost 1.4× less in all, and Claude Opus 5.5 answered sooner on 29, Grok 4.7 on 17.

  • Grok 4.7 vs Claude Sonnet 5.5: 47 and 47 of 50; 2 prompts split them; Claude Sonnet 5.5's replies cost 2.0× less in all, and Claude Sonnet 5.5 answered sooner on 47, Grok 4.7 on 1.

  • Grok 4.7 vs Claude Sonnet 5: 47 and 46 of 50; 7 prompts split them; Claude Sonnet 5's replies cost 1.9× less in all, and Claude Sonnet 5 answered sooner on 36, Grok 4.7 on 5. Claude Sonnet 5 vs Grok 4.7.

  • Grok 4.7 vs Claude Haiku 4.5: 47 and 43 of 50; 8 prompts split them; Claude Haiku 4.5's replies cost 5.2× less in all, and Claude Haiku 4.5 answered sooner on 49, Grok 4.7 on 1. Claude Haiku 4.5 vs Grok 4.7.

  • Grok 4.7 vs Claude Haiku 5.5: 47 and 45 of 50; 2 prompts split them; Claude Haiku 5.5's replies cost 33.0× less in all, and Claude Haiku 5.5 answered sooner on 48, Grok 4.7 on 1.

One writing prompt, both replies

Grok 4.7 and Claude Sonnet 5 on a writing prompt, the job where they differed most: both replies as they came.

An everyday prompt: “Announce a second bakery shop on LinkedIn”, in full.

  • Grok 4.7

    Passed: Graded 4.8 of 5 on average (lowest 4).

    We're happy to tell you that Crumb & Co. is opening a second shop in Leeds on 14 November. Our small bakery has been shaped by the customers who drop in for fresh bread, a slice of cake, and a friendly word at the counter. That support means a great deal to us, and we're glad to bring the same simple welcome to a new street. Whether you already know us or this will be your first visit, you'll find the same ovens, the same recipes, and the same quiet care.
    
    Opening day is 14 November, and coffee will be free for everyone who comes through the door. Please stop by, have a look around, and taste something straight from the oven. We would love to meet you there.

    1,625 tokens in, 1,299 out (1,149 of them reasoning) · 16.1 s · $0.0093 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·

  • Claude Sonnet 5

    Failed: Graded 3.5 of 5 on average (lowest 3).

    We have some exciting news to share: Crumb & Co. is opening our second shop, this time in Leeds!
    
    From our very first bakery, we've believed in doing simple things well — proper sourdough, honest pastries, and coffee worth stopping for. Now we get to bring that same spirit to a new neighbourhood and, hopefully, a whole new group of familiar faces.
    
    Our Leeds shop opens its doors on 14 November, and we'd love for you to be there. To celebrate, every coffee on opening day is on us — no catch, just our way of saying thank you for supporting us along the way.
    
    This has been months of planning, a lot of taste-testing, and more excitement than we know what to do with. We can't wait to welcome you in, share some fresh bakes, and start building new memories together.
    
    See you there — first coffee's on the house.

    538 tokens in, 267 out · 4.6 s · $0.0037 · 1 message on Pro · answered by anthropic/claude-sonnet-5 via Claude Platform on AWS ·

Every model, every job

All 7 Grok and Claude models in llmwise across the 50 prompts, with what a reply cost to run and each model's count on Pro.

Every Grok and Claude model across our test runs
ModelPassedHard onesCost per replyOn Pro
Grok 4.7xAI47 of 5019 of 20$0.0077Up to 250 a month
Claude Fable 5.1Anthropic45 of 5017 of 20$0.0200Up to 31 a month
Claude Opus 5.5Anthropic49 of 5019 of 20$0.0107Up to 62 a month
Claude Sonnet 5.5Anthropic47 of 5019 of 20$0.0038Up to 125 a month
Claude Sonnet 5Anthropic46 of 5019 of 20$0.0041Up to 125 a month
Claude Haiku 4.5Anthropic43 of 5015 of 20$0.0015Up to 250 a month
Claude Haiku 5.5Anthropic45 of 5019 of 20$0.00024Up to 60 a day

Passed: replies that passed their check, of those scored. Cost: what OpenRouter charged us per reply, on average; in llmwise you pay per message, not per token. Every prompt and how it's scored.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

Their own subscriptions

Each company's own plan at about llmwise Pro's price ($20 a month), in its own words, dated. We drop a plan here when its facts are more than 45 days old.

llmwise Pro, $20 a month, has all 7 of these models in one chat, on one monthly allowance. On it: Grok 4.7 up to 250 messages a month and Claude Sonnet 5.5 up to 125.

The lineups at a glance

What follows from each model's facts in our catalog.

  • The lineups

    Grok: one model, Grok 4.7. Claude: 6 models, Claude Fable 5.1, Claude Opus 5.5, Claude Sonnet 5.5, Claude Sonnet 5, Claude Haiku 4.5, and Claude Haiku 5.5.

  • Price per message

    Grok's one model is Grok 4.7 (250 messages a month on Pro); the least expensive Claude model is Claude Haiku 5.5 (60 messages a day on Pro).

  • Context window

    Grok goes up to 500K tokens (Grok 4.7); Claude up to 1M tokens (Claude Fable 5.1).

  • Images and PDFs

    Every model here reads images. Every model here takes a PDF as a whole file.

  • On the Free plan

    Free's one-time trial of 5 messages covers Grok 4.7, Claude Sonnet 5.5, Claude Sonnet 5, Claude Haiku 4.5, and Claude Haiku 5.5, and Claude Opus 5.5 for 1 message. Paid plans have every model, with messages every month.

Model by model

Two named models side by side, prompt by prompt, each with its messages on every plan.

More head-to-heads

Each of Grok and Claude against the other families, every job from the same test runs.

Where your messages go

In llmwise, a message to Claude goes to its maker, Anthropic, or through OpenRouter when llmwise can't reach the maker directly. Grok models are served only through OpenRouter, by endpoints that don't store or train on prompts. The Privacy Policy has the details.

Bar chart: Prompts passed in our test runs, all 10 jobs. Claude Opus 5.5: 49 of 50; Grok 4.7: 47 of 50; Claude Sonnet 5.5: 47 of 50; Claude Sonnet 5: 46 of 50; Claude Haiku 5.5: 45 of 50; Claude Fable 5.1: 45 of 50; Claude Haiku 4.5: 43 of 50.
Our test runs of September 27, 28, 29 and October 2, 7, 8, 9, 2026: the same prompts for every model, each reply checked the same way.

Questions

Which is better, Grok or Claude?

In our test runs on October 7, 2026, the same 50 prompts across 10 jobs: Grok's one model passed 47 of 50 replies and Claude's 6 models passed 275 of 300 replies. Job by job, both families' best models shared the top result on 8 of the 10 jobs; Claude's alone had it on writing and customer support.

Which is cheaper, Grok or Claude?

In llmwise, the least expensive Grok model is Grok 4.7 (250 messages a month on Pro), and the least expensive Claude model is Claude Haiku 5.5 (60 messages a day on Pro). At API list prices (October 2026), a typical message of 4,000 tokens in and 700 out costs $0.0122 on Grok 4.7 and $0.0008 on Claude Haiku 5.5.

Can I use Grok and Claude in the same chat?

Yes. Pick a model for each message; when you switch, the next model sees the whole conversation, including the other one's answers.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, GLM, and Mistral, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.