Skip to content

Comparison · Data analysis

Claude vs Gemini for data analysis

On llmwise Pro, Claude Sonnet 5.5 gets up to 125 messages a month and Gemini 3.1 Pro (preview) up to 125 messages a month. We ran the same 5 data analysis prompts on all 5 Claude models and all 2 Gemini models and published every reply: the results, then Claude Sonnet 5.5 against Gemini 3.1 Pro prompt by prompt, then how Claude and Gemini compare on price per message, context and files.

Based on 35 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Short answer

In our data analysis test runs on September 28, 2026, Claude's 5 models passed 22 of 25; Claude Fable 5.1, Claude Opus 5.5, and Claude Sonnet 5.5 each passed 5 of 5. Gemini's 2 models passed 9 of 10; its best, Gemini 3.1 Pro, passed 5 of 5 ($0.0155 a reply). Claude Sonnet 5.5 and Gemini 3.1 Pro each passed 5 of the 5 prompts, so these data analysis prompts don't split Claude and Gemini; the replies on this page show how they differ.

Claude and Gemini on our data analysis test runs

Every Claude and Gemini model in llmwise on our 5 data analysis prompts: how many replies passed, what each counted as on Pro, and what it cost to run.

Claude and Gemini on our data analysis test runs
ModelPassedHard onesMessages used on ProCost per replyTime per reply
Claude Fable 5.1Anthropic5 of 52 of 21 each, of 31 a month on Pro$0.02836.7 s
Claude Opus 5.5Anthropic5 of 52 of 21 each, of 62 a month on Pro$0.01315.9 s
Claude Sonnet 5.5Anthropic5 of 52 of 21 each, of 125 a month on Pro$0.00472.4 s
Claude Sonnet 5Anthropic4 of 51 of 21 each, of 125 a month on Pro$0.00696.9 s
Claude Haiku 4.5Anthropic3 of 51 of 21 each, of 250 a month on Pro$0.00253.2 s
Gemini 3.1 Pro (preview)Google5 of 52 of 21 each, of 125 a month on Pro$0.01559.9 s
Gemini 3.8 FlashGoogle4 of 51 of 21 each, of 250 a month on Pro$0.00285.4 s
Claude: 22 of 25 replies passed; Gemini: 9 of 10 replies passed. Messages used is what each reply counted as on Pro; cost is what OpenRouter charged us per reply.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

Claude vs Gemini, prompt by prompt

Each data analysis prompt with Claude Sonnet 5.5 and Gemini 3.1 Pro side by side, as they replied: for each side, the model that passed the most data analysis prompts (then the most hard ones, then the cheaper).

Data analysis: Claude Sonnet 5.5 passed 5 of 5 and Gemini 3.1 Pro (preview) 5 of 5. Claude Sonnet 5.5 answered 3.8× sooner at the median, 2.3 s against 8.6 s. Claude Sonnet 5.5 cost 3.3× less, $0.0235 against $0.0773 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 16% longer, in tokens of reply, thinking not counted.

  1. 1. The region with the most revenue

    Everyday. Claude: 5 of 5 models passed, Claude Haiku 4.5 the cheapest at $0.0026 · Gemini: 2 of 2 models passed, Gemini 3.8 Flash the cheapest at $0.0038. The prompt in full.

    Claude Sonnet 5.5 and Gemini 3.1 Pro both passed this one. Claude Sonnet 5.5 answered sooner (2.4 s to 8.5 s), and Claude Sonnet 5.5's reply cost less ($0.0051 to $0.0117); Claude Sonnet 5.5 wrote 307 tokens of reply, Gemini 3.1 Pro 255.

    • Claude Sonnet 5.5

      Passed: Final answer North, with 1,085.00 in revenue: right.

      Revenue per order (quantity × price):
      
      - 1001 West: 2×45 = 90
      - 1002 East: 1×120 = 120
      - 1003 West: 1×310 = 310
      - 1004 North: 4×45 = 180
      - 1005 East: 3×45 = 135
      - 1006 North: 2×120 = 240
      - 1007 West: 2×120 = 240
      - 1008 East: 1×310 = 310
      - 1009 North: 2×310 = 620
      - 1010 West: 5×45 = 225
      - 1011 East: 3×120 = 360
      - 1012 North: 1×45 = 45
      
      Totals by region:
      
      - West: 90 + 310 + 240 + 225 = 865
      - East: 120 + 135 + 310 + 360 = 925
      - North: 180 + 240 + 620 + 45 = 1085
      
      North is highest.
      
      Final answer: North, with $1,085.00 in revenue

      817 tokens in, 307 out · 2.4 s · $0.0051 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·

    • Gemini 3.1 Pro

      Passed: Final answer North, 1085: right.

      Here is the revenue calculation for each region (quantity × price):
      
      *   **West**: 
          *   2 × $45 = $90
          *   1 × $310 = $310
          *   2 × $120 = $240
          *   5 × $45 = $225
          *   *Total* = $865
      *   **East**: 
          *   1 × $120 = $120
          *   3 × $45 = $135
          *   1 × $310 = $310
          *   3 × $120 = $360
          *   *Total* = $925
      *   **North**: 
          *   4 × $45 = $180
          *   2 × $120 = $240
          *   2 × $310 = $620
          *   1 × $45 = $45
          *   *Total* = $1085
      
      Final answer: North, $1085

      770 tokens in, 849 out (594 of them reasoning) · 8.5 s · $0.0117 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

  2. 2. Average order value in August

    Everyday. Claude: 4 of 5 models passed, Claude Sonnet 5.5 the cheapest at $0.0037 · Gemini: 2 of 2 models passed, Gemini 3.8 Flash the cheapest at $0.0016. The prompt in full.

    Claude Sonnet 5.5 and Gemini 3.1 Pro both passed this one. Claude Sonnet 5.5 answered sooner (1.7 s to 8.6 s), and Claude Sonnet 5.5's reply cost less ($0.0037 to $0.0130); Claude Sonnet 5.5 wrote 170 tokens of reply, Gemini 3.1 Pro 273.

    • Claude Sonnet 5.5

      Passed: Final answer 300.00: right.

      August 2026 orders (1007–1012):
      
      - 1007: 2 × 120 = 240
      - 1008: 1 × 310 = 310
      - 1009: 2 × 310 = 620
      - 1010: 5 × 45 = 225
      - 1011: 3 × 120 = 360
      - 1012: 1 × 45 = 45
      
      Total = 240 + 310 + 620 + 225 + 360 + 45 = 1800
      
      Average = 1800 / 6 = 300.00
      
      Final answer: 300.00

      825 tokens in, 170 out · 1.7 s · $0.0037 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·

    • Gemini 3.1 Pro

      Passed: Final answer 300.00: right.

      Here are the orders for August 2026 and their calculated revenue (quantity × price):
      
      *   Order 1007: 2 × $120.00 = $240.00
      *   Order 1008: 1 × $310.00 = $310.00
      *   Order 1009: 2 × $310.00 = $620.00
      *   Order 1010: 5 × $45.00 = $225.00
      *   Order 1011: 3 × $120.00 = $360.00
      *   Order 1012: 1 × $45.00 = $45.00
      
      Total revenue for August = $240 + $310 + $620 + $225 + $360 + $45 = $1800.00
      Number of orders in August = 6
      Average revenue per order = $1800.00 / 6 = $300.00
      
      Final answer: 300.00

      779 tokens in, 955 out (682 of them reasoning) · 8.6 s · $0.0130 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

  3. 3. Revenue change from July to August

    Everyday. Claude: 5 of 5 models passed, Claude Haiku 4.5 the cheapest at $0.0022 · Gemini: 2 of 2 models passed, Gemini 3.8 Flash the cheapest at $0.0037. The prompt in full.

    Claude Sonnet 5.5 and Gemini 3.1 Pro both passed this one. Claude Sonnet 5.5 answered sooner (2.3 s to 9.0 s), and Claude Sonnet 5.5's reply cost less ($0.0048 to $0.0135); Claude Sonnet 5.5 wrote 273 tokens of reply, Gemini 3.1 Pro 356.

    • Claude Sonnet 5.5

      Passed: Final answer 67.4%: right.

      **July:**
      - 1001: 2 × 45 = 90
      - 1002: 1 × 120 = 120
      - 1003: 1 × 310 = 310
      - 1004: 4 × 45 = 180
      - 1005: 3 × 45 = 135
      - 1006: 2 × 120 = 240
      
      July total = 1075
      
      **August:**
      - 1007: 2 × 120 = 240
      - 1008: 1 × 310 = 310
      - 1009: 2 × 310 = 620
      - 1010: 5 × 45 = 225
      - 1011: 3 × 120 = 360
      - 1012: 1 × 45 = 45
      
      August total = 1800
      
      **Change:** (1800 − 1075) / 1075 = 725 / 1075 ≈ 0.67442 → 67.4%
      
      Final answer: 67.4%

      833 tokens in, 273 out · 2.3 s · $0.0048 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·

    • Gemini 3.1 Pro

      Passed: Final answer 67.4%: right.

      **July 2026 Revenue:**
      * Order 1001: 2 x $45 = $90
      * Order 1002: 1 x $120 = $120
      * Order 1003: 1 x $310 = $310
      * Order 1004: 4 x $45 = $180
      * Order 1005: 3 x $45 = $135
      * Order 1006: 2 x $120 = $240
      * **July Total:** $1,075
      
      **August 2026 Revenue:**
      * Order 1007: 2 x $120 = $240
      * Order 1008: 1 x $310 = $310
      * Order 1009: 2 x $310 = $620
      * Order 1010: 5 x $45 = $225
      * Order 1011: 3 x $120 = $360
      * Order 1012: 1 x $45 = $45
      * **August Total:** $1,800
      
      **Percentage Change:**
      ((1800 - 1075) / 1075) * 100 = (725 / 1075) * 100 = 67.4418...%
      
      Final answer: 67.4%

      786 tokens in, 990 out (634 of them reasoning) · 9.0 s · $0.0135 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

  4. 4. A median, filtered two ways

    Hard. Claude: 5 of 5 models passed, Claude Haiku 4.5 the cheapest at $0.0016 · Gemini: 2 of 2 models passed, Gemini 3.8 Flash the cheapest at $0.0010. The prompt in full.

    Claude Sonnet 5.5 and Gemini 3.1 Pro both passed this one. Claude Sonnet 5.5 answered sooner (1.5 s to 7.1 s), and Claude Sonnet 5.5's reply cost less ($0.0030 to $0.0082); Claude Sonnet 5.5 wrote 127 tokens of reply, Gemini 3.1 Pro 185.

    • Claude Sonnet 5.5

      Passed: Final answer 6.25 hours: right.

      Billing + high tickets: T1 (5.5), T4 (2), T6 (9), T9 (7), T10 (26), T12 (4).
      
      Sorted: 2, 4, 5.5, 7, 9, 26
      
      Six values, so the median is the average of the 3rd and 4th: (5.5 + 7) / 2 = 6.25
      
      Final answer: 6.25 hours

      693 tokens in, 127 out · 1.5 s · $0.0030 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·

    • Gemini 3.1 Pro

      Passed: Final answer 6.25: right.

      To find the median `hours_to_resolve` for high-priority tickets handled by the Billing team, we first filter the data for those specific tickets:
      
      *   T1: 5.5
      *   T4: 2
      *   T6: 9
      *   T9: 7
      *   T10: 26
      *   T12: 4
      
      Next, we sort these values in ascending order:
      2, 4, 5.5, 7, 9, 26
      
      Since there are 6 values (an even number), the median is the average of the two middle numbers (the 3rd and 4th values):
      (5.5 + 7) / 2 = 12.5 / 2 = 6.25
      
      Final answer: 6.25

      551 tokens in, 590 out (405 of them reasoning) · 7.1 s · $0.0082 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

  5. 5. Correlation between ad spend and sign-ups

    Hard. Claude: 3 of 5 models passed, Claude Sonnet 5.5 the cheapest at $0.0070 · Gemini: 1 of 2 models passed, Gemini 3.1 Pro at $0.0310. The prompt in full.

    Claude Sonnet 5.5 and Gemini 3.1 Pro both passed this one. Claude Sonnet 5.5 answered sooner (4.2 s to 16.5 s), and Claude Sonnet 5.5's reply cost less ($0.0070 to $0.0310); Claude Sonnet 5.5 wrote 547 tokens of reply, Gemini 3.1 Pro 588.

    • Claude Sonnet 5.5

      Passed: Final answer 0.97: right.

      Data: x = ad_spend, y = signups (n = 8)
      
      Sums:
      - Σx = 500+800+650+1200+900+300+1100+700 = 6150 → mean x = 768.75
      - Σy = 42+55+49+71+60+35+64+58 = 434 → mean y = 54.25
      
      Deviations (dx, dy), products:
      - 500: dx=-268.75, dy=-12.25 → dxdy=3292.19; dx²=72226.56; dy²=150.06
      - 800: dx=31.25, dy=0.75 → 23.44; 976.56; 0.56
      - 650: dx=-118.75, dy=-5.25 → 623.44; 14101.56; 27.56
      - 1200: dx=431.25, dy=16.75 → 7223.44; 185976.56; 280.56
      - 900: dx=131.25, dy=5.75 → 754.69; 17226.56; 33.06
      - 300: dx=-468.75, dy=-19.25 → 9023.44; 219726.56; 370.56
      - 1100: dx=331.25, dy=9.75 → 3229.69; 109726.56; 95.06
      - 700: dx=-68.75, dy=3.75 → -257.81; 4726.56; 14.06
      
      Totals:
      - Sxy ≈ 23,912.5
      - Sxx ≈ 624,687.5
      - Syy ≈ 971.5
      
      r = 23,912.5 / √(624,687.5 × 971.5) = 23,912.5 / √606,884,000 ≈ 23,912.5 / 24,635.4 ≈ 0.97
      
      Final answer: 0.97

      609 tokens in, 547 out · 4.2 s · $0.0070 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·

    • Gemini 3.1 Pro

      Passed: Final answer 0.97: right.

      To calculate the Pearson correlation coefficient ($r$), we use the formula:
      $r = \frac{n(\sum xy) - (\sum x)(\sum y)}{\sqrt{[n\sum x^2 - (\sum x)^2][n\sum y^2 - (\sum y)^2]}}$
      
      Given the data for ad_spend ($x$) and signups ($y$):
      - $n = 8$
      - $\sum x = 500 + 800 + 650 + 1200 + 900 + 300 + 1100 + 700 = 6,150$
      - $\sum y = 42 + 55 + 49 + 71 + 60 + 35 + 64 + 58 = 434$
      - $\sum xy = (500 \times 42) + \dots + (700 \times 58) = 357,550$
      - $\sum x^2 = (500^2) + \dots + (700^2) = 5,352,500$
      - $\sum y^2 = (42^2) + \dots + (58^2) = 24,516$
      
      Plugging these into the formula:
      Numerator: $8(357,550) - (6,150)(434) = 2,860,400 - 2,669,100 = 191,300$
      Denominator $x$: $8(5,352,500) - (6,150)^2 = 42,820,000 - 37,822,500 = 4,997,500$
      Denominator $y$: $8(24,516) - (434)^2 = 196,128 - 188,356 = 7,772$
      
      $r = \frac{191,300}{\sqrt{4,997,500 \times 7,772}}$
      $r = \frac{191,300}{\sqrt{38,840,570,000}}$
      $r = \frac{191,300}{197,080.11} \approx 0.97067$
      
      Final answer: 0.97

      505 tokens in, 2,496 out (1,908 of them reasoning) · 16.5 s · $0.0310 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

Each Claude model against each Gemini model

Every Claude model against every Gemini model on the same 5 data analysis prompts: passes, the median wait and what the replies cost. A lead under 10% counts as close.

  • Claude Fable 5.1 vs Gemini 3.1 Pro: 5 of 5 each; Claude Fable 5.1 answered 1.8× sooner at the median and Gemini 3.1 Pro cost 1.8× less.

  • Claude Fable 5.1 vs Gemini 3.8 Flash: Claude Fable 5.1 5 of 5, Gemini 3.8 Flash 4; Claude Fable 5.1 answered 1.1× sooner at the median and Gemini 3.8 Flash cost 10.2× less.

  • Claude Opus 5.5 vs Gemini 3.1 Pro: 5 of 5 each; Claude Opus 5.5 answered 1.9× sooner at the median and cost 1.2× less. Claude Opus 5.5 vs Gemini 3.1 Pro (preview), on every job.

  • Claude Opus 5.5 vs Gemini 3.8 Flash: Claude Opus 5.5 5 of 5, Gemini 3.8 Flash 4; Claude Opus 5.5 answered 1.2× sooner at the median and Gemini 3.8 Flash cost 4.7× less.

  • Claude Sonnet 5.5 vs Gemini 3.1 Pro: 5 of 5 each; Claude Sonnet 5.5 answered 3.8× sooner at the median and cost 3.3× less. Claude Sonnet 5.5 vs Gemini 3.1 Pro (preview), on every job.

  • Claude Sonnet 5.5 vs Gemini 3.8 Flash: Claude Sonnet 5.5 5 of 5, Gemini 3.8 Flash 4; Claude Sonnet 5.5 answered 2.4× sooner at the median and Gemini 3.8 Flash cost 1.7× less.

  • Claude Sonnet 5 vs Gemini 3.1 Pro: Claude Sonnet 5 4 of 5, Gemini 3.1 Pro 5; Claude Sonnet 5 answered 1.6× sooner at the median and cost 2.2× less. Claude Sonnet 5 vs Gemini 3.1 Pro (preview), on every job.

  • Claude Sonnet 5 vs Gemini 3.8 Flash: 4 of 5 each; they took about as long and Gemini 3.8 Flash cost 2.5× less. Claude Sonnet 5 vs Gemini 3.8 Flash, on every job.

  • Claude Haiku 4.5 vs Gemini 3.1 Pro: Claude Haiku 4.5 3 of 5, Gemini 3.1 Pro 5; Claude Haiku 4.5 answered 3.2× sooner at the median and cost 6.2× less. Claude Haiku 4.5 vs Gemini 3.1 Pro (preview), on every job.

  • Claude Haiku 4.5 vs Gemini 3.8 Flash: Claude Haiku 4.5 3 of 5, Gemini 3.8 Flash 4; Claude Haiku 4.5 answered 2.0× sooner at the median and cost 1.1× less. Claude Haiku 4.5 vs Gemini 3.8 Flash, on every job.

How the data analysis replies are scored

Every Claude and Gemini reply above was checked the same way as every other model's, by the rules published with the prompts: how the data analysis prompts are scored, and each one in full.

The differences at a glance

What follows from each model's facts in our catalog.

  • The lineups

    Claude: 5 models, Claude Fable 5.1, Claude Opus 5.5, Claude Sonnet 5.5, Claude Sonnet 5, and Claude Haiku 4.5. Gemini: 2 models, Gemini 3.1 Pro (preview) and Gemini 3.8 Flash.

  • Price per message

    The least expensive Claude model is Claude Haiku 4.5 (250 messages a month on Pro); the least expensive Gemini model is Gemini 3.8 Flash (250 messages a month on Pro).

  • Context window

    Claude goes up to 1M tokens (Claude Fable 5.1); Gemini up to 1.05M tokens (Gemini 3.1 Pro (preview)).

  • Images and PDFs

    Every model here reads images. Every model here takes a PDF as a whole file.

  • On the Free plan

    Free's one-time trial of 5 messages covers Claude Sonnet 5.5, Claude Sonnet 5, Claude Haiku 4.5, Gemini 3.1 Pro (preview), and Gemini 3.8 Flash. Paid plans have every model, with messages every month.

Every Claude and Gemini model's context window, files and API price: Claude vs Gemini.

Where your messages go

In llmwise, a message to Claude or Gemini goes to the model's maker, or through OpenRouter when llmwise can't reach the maker directly. The Privacy Policy has the details.

Questions

Which is better, Claude or Gemini for data analysis?

In our data analysis test runs on September 28, 2026, Claude's 5 models passed 22 of 25; Claude Fable 5.1, Claude Opus 5.5, and Claude Sonnet 5.5 each passed 5 of 5. Gemini's 2 models passed 9 of 10; its best, Gemini 3.1 Pro, passed 5 of 5 ($0.0155 a reply). Claude Sonnet 5.5 and Gemini 3.1 Pro each passed 5 of the 5 prompts, so these data analysis prompts don't split Claude and Gemini; the replies on this page show how they differ.

Is Claude or Gemini cheaper?

In llmwise, the least expensive Claude model is Claude Haiku 4.5 (250 messages a month on Pro), and the least expensive Gemini model is Gemini 3.8 Flash (250 messages a month on Pro). At API list prices (September 2026), a typical message of 4,000 tokens in and 700 out costs $0.0075 on Claude Haiku 4.5 and $0.0056 on Gemini 3.8 Flash.

Can I use Claude and Gemini in the same chat?

Yes. Ask Claude Sonnet 5.5 a question, then switch the picker to Gemini 3.1 Pro and ask again: Gemini 3.1 Pro sees the whole conversation, Claude Sonnet 5.5's answer included.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.