Skip to content

Comparison · 4 families

ChatGPT vs Claude vs Gemini vs Grok

On llmwise Pro, GPT-6 Sol up to 125 messages a month, Claude Sonnet 5.5 up to 125, Gemini 3.1 Pro (preview) up to 125, and Grok 4.7 up to 250. ChatGPT is OpenAI's own app for its GPT models; llmwise has GPT models in its own chat, not ChatGPT itself. We ran the same 50 prompts across 10 jobs on every GPT, Claude, Gemini, and Grok model and published every reply: each job with the families' picks ranked, then each one's own subscription.

Based on 550 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Short answer

In our test runs on September 28, 2026, the same 50 prompts across 10 jobs: GPT's 3 models passed 140 of 150 replies, Claude's 5 models passed 230 of 250 replies, Gemini's 2 models passed 95 of 100 replies, and Grok's one model passed 47 of 50 replies. Job by job, all 4 families' best models shared the top result on 8 of the 10 jobs; GPT's alone had it on writing; GPT's trailed on customer support; Claude's trailed on writing; Gemini's trailed on writing; Grok's trailed on writing and customer support.

GPT, Claude, Gemini, Grok: each job ranked

Each family's pick on each job (its model that passed the most of the job's prompts, then the most hard ones, then the cheaper), ranked by passes and then by what a reply cost.

  • Coding (all 4 picks level on passes, GPT's the cheapest): GPT's GPT-6 Luna, 5 of 5 at $0.0003 a reply in 4.9 s; Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0020 a reply in 4.4 s; Claude's Claude Sonnet 5.5, 5 of 5 at $0.0063 a reply in 3.0 s; Grok's Grok 4.7, 5 of 5 at $0.0220 a reply in 52.2 s.

  • Writing (GPT on top): GPT's GPT-6 Luna, 5 of 5 at $0.0001 a reply in 2.1 s; Claude's Claude Sonnet 5, 4 of 5 at $0.0034 a reply in 4.2 s; Gemini's Gemini 3.1 Pro, 4 of 5 at $0.0092 a reply in 8.6 s; Grok's Grok 4.7, 3 of 5 at $0.0042 a reply in 8.0 s.

  • Math (all 4 picks level on passes, Gemini's the cheapest): Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0014 a reply in 3.8 s; Claude's Claude Haiku 4.5, 5 of 5 at $0.0015 a reply in 2.4 s; GPT's GPT-6 Sol, 5 of 5 at $0.0018 a reply in 2.2 s; Grok's Grok 4.7, 5 of 5 at $0.0050 a reply in 14.0 s.

  • Summarization (all 4 picks level on passes, Gemini's the cheapest): Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0008 a reply in 3.7 s; Claude's Claude Haiku 4.5, 5 of 5 at $0.0010 a reply in 1.9 s; Grok's Grok 4.7, 5 of 5 at $0.0032 a reply in 5.1 s; GPT's GPT-6 Astra, 5 of 5 at $0.0109 a reply in 3.0 s.

  • Data analysis (all 4 picks level on passes, GPT's the cheapest): GPT's GPT-6 Luna, 5 of 5 at $0.0002 a reply in 3.4 s; Claude's Claude Sonnet 5.5, 5 of 5 at $0.0047 a reply in 2.4 s; Grok's Grok 4.7, 5 of 5 at $0.0085 a reply in 14.9 s; Gemini's Gemini 3.1 Pro, 5 of 5 at $0.0155 a reply in 9.9 s.

  • Customer support (Claude and Gemini on top): Claude's Claude Sonnet 5, 5 of 5 at $0.0036 a reply in 3.7 s; Gemini's Gemini 3.1 Pro, 5 of 5 at $0.0106 a reply in 8.7 s; GPT's GPT-6 Luna, 4 of 5 at $0.0001 a reply in 1.4 s; Grok's Grok 4.7, 4 of 5 at $0.0048 a reply in 9.4 s.

  • Translation (all 4 picks level on passes, GPT's the cheapest): GPT's GPT-6 Luna, 5 of 5 at $0.0001 a reply in 1.7 s; Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0007 a reply in 3.1 s; Claude's Claude Haiku 4.5, 5 of 5 at $0.0015 a reply in 2.3 s; Grok's Grok 4.7, 5 of 5 at $0.0069 a reply in 14.8 s.

  • SQL (all 4 picks level on passes, GPT's the cheapest): GPT's GPT-6 Luna, 5 of 5 at $0.0001 a reply in 1.5 s; Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0007 a reply in 3.6 s; Claude's Claude Haiku 4.5, 5 of 5 at $0.0010 a reply in 1.4 s; Grok's Grok 4.7, 5 of 5 at $0.0038 a reply in 6.1 s.

  • RAG and answering from documents (all 4 picks level on passes, GPT's the cheapest): GPT's GPT-6 Luna, 5 of 5 at $0.0001 a reply in 1.3 s; Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0007 a reply in 3.2 s; Claude's Claude Haiku 4.5, 5 of 5 at $0.0010 a reply in 1.3 s; Grok's Grok 4.7, 5 of 5 at $0.0023 a reply in 2.7 s.

  • Agents and tool use (all 4 picks level on passes, GPT's the cheapest): GPT's GPT-6 Luna, 5 of 5 at $0.0001 a reply in 1.7 s; Gemini's Gemini 3.8 Flash, 5 of 5 at $0.0009 a reply in 4.6 s; Claude's Claude Haiku 4.5, 5 of 5 at $0.0009 a reply in 1.1 s; Grok's Grok 4.7, 5 of 5 at $0.0026 a reply in 3.8 s.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

Their own subscriptions

Each company's own plan at about llmwise Pro's price ($20 a month), in its own words, dated. We drop a plan here when its facts are more than 45 days old.

llmwise Pro, $20 a month, has all 11 of these models in one chat, from one monthly allowance: GPT-6 Sol up to 125 messages a month, Claude Sonnet 5.5 up to 125, Gemini 3.1 Pro (preview) up to 125, and Grok 4.7 up to 250.

More head-to-heads

These families two at a time, and the other pages about several.

Where your messages go

In llmwise, a message to GPT, Claude, or Gemini goes to the model's maker, or through OpenRouter when llmwise can't reach the maker directly. Grok models are served only through OpenRouter, by endpoints that don't store or train on prompts. The Privacy Policy has the details.

Questions

Which is better, ChatGPT, Claude, Gemini, or Grok?

In our test runs on September 28, 2026, the same 50 prompts across 10 jobs: GPT's 3 models passed 140 of 150 replies, Claude's 5 models passed 230 of 250 replies, Gemini's 2 models passed 95 of 100 replies, and Grok's one model passed 47 of 50 replies. Job by job, all 4 families' best models shared the top result on 8 of the 10 jobs; GPT's alone had it on writing; GPT's trailed on customer support; Claude's trailed on writing; Gemini's trailed on writing; Grok's trailed on writing and customer support.

Which is cheaper, GPT, Claude, Gemini, or Grok?

At API list prices (September 2026), the least expensive model of each family for a typical message is GPT-6 Luna ($0.0008), Claude Haiku 4.5 ($0.0075), Gemini 3.8 Flash ($0.0056), and Grok 4.7 ($0.0098). In llmwise you pay per message, and every one of them is on the same plan.

Can I use GPT, Claude, Gemini, and Grok in the same chat?

Yes. Pick GPT, Claude, Gemini, and Grok models message by message in one chat; each model sees the whole conversation, the others' answers included.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.