Gemini 3.1 Pro · Data analysis
Gemini 3.1 Pro for data analysis: our test runs
Gemini 3.1 Pro ran our 5 data analysis prompts in llmwise, through the same pipeline your messages take. Here's every reply as it came, whether it passed, what it cost and what it counts as on Pro, next to the other models.
Based on 5 of our test runs on , through OpenRouter with the app's own prompt and settings.
Short answer
In our test runs on September 27, 2026, Gemini 3.1 Pro passed 5 of 5 of our data analysis prompts (both hard ones): tied 1st of 15 models. Each reply counted as 1 message on Pro, where it has 125 a month, and cost $0.0155 to run on average.
Prompt by prompt
Each of the 5 data analysis prompts, and how Gemini 3.1 Pro's reply did.
| Prompt | Result | Cost | Time |
|---|---|---|---|
| The region with the most revenueEveryday | Passed: Final answer North, 1085: right. | $0.0117 | 8.5 s |
| Average order value in AugustEveryday | Passed: Final answer 300.00: right. | $0.0130 | 8.6 s |
| Revenue change from July to AugustEveryday | Passed: Final answer 67.4%: right. | $0.0135 | 9.0 s |
| A median, filtered two waysHard | Passed: Final answer 6.25: right. | $0.0082 | 7.1 s |
| Correlation between ad spend and sign-upsHard | Passed: Final answer 0.97: right. | $0.0310 | 16.5 s |
Every data analysis reply from Gemini 3.1 Pro
The region with the most revenue
An everyday prompt: “The region with the most revenue”, in full.
Passed: Final answer North, 1085: right.
Here is the revenue calculation for each region (quantity × price): * **West**: * 2 × $45 = $90 * 1 × $310 = $310 * 2 × $120 = $240 * 5 × $45 = $225 * *Total* = $865 * **East**: * 1 × $120 = $120 * 3 × $45 = $135 * 1 × $310 = $310 * 3 × $120 = $360 * *Total* = $925 * **North**: * 4 × $45 = $180 * 2 × $120 = $240 * 2 × $310 = $620 * 1 × $45 = $45 * *Total* = $1085 Final answer: North, $1085770 tokens in, 849 out (594 of them reasoning) · 8.5 s · $0.0117 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
Average order value in August
An everyday prompt: “Average order value in August”, in full.
Passed: Final answer 300.00: right.
Here are the orders for August 2026 and their calculated revenue (quantity × price): * Order 1007: 2 × $120.00 = $240.00 * Order 1008: 1 × $310.00 = $310.00 * Order 1009: 2 × $310.00 = $620.00 * Order 1010: 5 × $45.00 = $225.00 * Order 1011: 3 × $120.00 = $360.00 * Order 1012: 1 × $45.00 = $45.00 Total revenue for August = $240 + $310 + $620 + $225 + $360 + $45 = $1800.00 Number of orders in August = 6 Average revenue per order = $1800.00 / 6 = $300.00 Final answer: 300.00
779 tokens in, 955 out (682 of them reasoning) · 8.6 s · $0.0130 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
Revenue change from July to August
An everyday prompt: “Revenue change from July to August”, in full.
Passed: Final answer 67.4%: right.
**July 2026 Revenue:** * Order 1001: 2 x $45 = $90 * Order 1002: 1 x $120 = $120 * Order 1003: 1 x $310 = $310 * Order 1004: 4 x $45 = $180 * Order 1005: 3 x $45 = $135 * Order 1006: 2 x $120 = $240 * **July Total:** $1,075 **August 2026 Revenue:** * Order 1007: 2 x $120 = $240 * Order 1008: 1 x $310 = $310 * Order 1009: 2 x $310 = $620 * Order 1010: 5 x $45 = $225 * Order 1011: 3 x $120 = $360 * Order 1012: 1 x $45 = $45 * **August Total:** $1,800 **Percentage Change:** ((1800 - 1075) / 1075) * 100 = (725 / 1075) * 100 = 67.4418...% Final answer: 67.4%
786 tokens in, 990 out (634 of them reasoning) · 9.0 s · $0.0135 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
A median, filtered two ways
A hard prompt: “A median, filtered two ways”, in full.
Passed: Final answer 6.25: right.
To find the median `hours_to_resolve` for high-priority tickets handled by the Billing team, we first filter the data for those specific tickets: * T1: 5.5 * T4: 2 * T6: 9 * T9: 7 * T10: 26 * T12: 4 Next, we sort these values in ascending order: 2, 4, 5.5, 7, 9, 26 Since there are 6 values (an even number), the median is the average of the two middle numbers (the 3rd and 4th values): (5.5 + 7) / 2 = 12.5 / 2 = 6.25 Final answer: 6.25
551 tokens in, 590 out (405 of them reasoning) · 7.1 s · $0.0082 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
Correlation between ad spend and sign-ups
A hard prompt: “Correlation between ad spend and sign-ups”, in full.
Passed: Final answer 0.97: right.
To calculate the Pearson correlation coefficient ($r$), we use the formula: $r = \frac{n(\sum xy) - (\sum x)(\sum y)}{\sqrt{[n\sum x^2 - (\sum x)^2][n\sum y^2 - (\sum y)^2]}}$ Given the data for ad_spend ($x$) and signups ($y$): - $n = 8$ - $\sum x = 500 + 800 + 650 + 1200 + 900 + 300 + 1100 + 700 = 6,150$ - $\sum y = 42 + 55 + 49 + 71 + 60 + 35 + 64 + 58 = 434$ - $\sum xy = (500 \times 42) + \dots + (700 \times 58) = 357,550$ - $\sum x^2 = (500^2) + \dots + (700^2) = 5,352,500$ - $\sum y^2 = (42^2) + \dots + (58^2) = 24,516$ Plugging these into the formula: Numerator: $8(357,550) - (6,150)(434) = 2,860,400 - 2,669,100 = 191,300$ Denominator $x$: $8(5,352,500) - (6,150)^2 = 42,820,000 - 37,822,500 = 4,997,500$ Denominator $y$: $8(24,516) - (434)^2 = 196,128 - 188,356 = 7,772$ $r = \frac{191,300}{\sqrt{4,997,500 \times 7,772}}$ $r = \frac{191,300}{\sqrt{38,840,570,000}}$ $r = \frac{191,300}{197,080.11} \approx 0.97067$ Final answer: 0.97505 tokens in, 2,496 out (1,908 of them reasoning) · 16.5 s · $0.0310 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
How it compares on data analysis
Gemini 3.1 Pro beside the job's picks among every model. 11 models share the hard-problems pick, Gemini 3.1 Pro among them.
| Model | Passed | On Pro | Cost per reply |
|---|---|---|---|
| DeepSeek V4 ProOur pick: hard problems (shared), best value | 5 of 5 | 250 a month on Pro | $0.0016 |
| GPT-6 LunaOur pick: hard problems (shared), everyday | 5 of 5 | 60 a day on Pro | $0.0002 |
| Gemini 3.1 ProOur pick: hard problems (shared) | 5 of 5 | 125 a month on Pro | $0.0155 |
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
How these runs were done
Final answer. Automatic. The reply's last “Final answer:” line must hold the right value.
Gemini 3.1 Pro, and data analysis, elsewhere
Questions
Is Gemini 3.1 Pro good for data analysis?
In our test runs it passed 5 of 5 data analysis prompts, tied 1st of the 15 models in llmwise. Every reply is on this page, so you can judge them yourself.
How many of my messages does a data analysis reply from Gemini 3.1 Pro use?
1 message each on Pro, where it has 125 a month on Pro. The price of a message is fixed and shown before you send it, however long the reply.
How were these runs done?
The same way for every model: each prompt sent through llmwise's own pipeline, each reply checked the same way. The methods page has every prompt and how each is scored.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.