Comparison · Data analysis
ChatGPT vs DeepSeek for data analysis
On llmwise Pro, GPT-6 Sol gets up to 125 messages a month and DeepSeek V4 Pro up to 250 messages a month. ChatGPT is OpenAI's own app for its GPT models; llmwise has GPT models in its own chat, not ChatGPT itself. We ran the same 5 data analysis prompts on all 3 GPT models and all 2 DeepSeek models and published every reply: the results, then GPT-6 Luna against DeepSeek V4.1 Flash prompt by prompt, then how GPT and DeepSeek compare on price per message, context and files.
Based on 25 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
Short answer
In our data analysis test runs on September 27, 2026, GPT's 3 models passed 15 of 15; GPT-6 Astra, GPT-6 Sol, and GPT-6 Luna each passed 5 of 5. DeepSeek's 2 models passed 10 of 10; DeepSeek V4 Pro and DeepSeek V4.1 Flash each passed 5 of 5. GPT-6 Luna and DeepSeek V4.1 Flash each passed 5 of the 5 prompts, so these data analysis prompts don't split GPT and DeepSeek; the replies on this page show how they differ.
GPT and DeepSeek on our data analysis test runs
Every GPT and DeepSeek model in llmwise on our 5 data analysis prompts: how many replies passed, what each counted as on Pro, and what it cost to run.
| Model | Passed | Hard ones | Messages used on Pro | Cost per reply | Time per reply |
|---|---|---|---|---|---|
| GPT-6 AstraOpenAI | 5 of 5 | 2 of 2 | 1 each, of 31 a month on Pro | $0.0154 | 4.0 s |
| GPT-6 SolOpenAI | 5 of 5 | 2 of 2 | 1 each, of 125 a month on Pro | $0.0034 | 3.5 s |
| GPT-6 LunaOpenAI | 5 of 5 | 2 of 2 | 1 each, of 60 a day on Pro | $0.0002 | 3.4 s |
| DeepSeek V4 ProDeepSeek | 5 of 5 | 2 of 2 | 1 each, of 250 a month on Pro | $0.0016 | 10.1 s |
| DeepSeek V4.1 FlashDeepSeek | 5 of 5 | 2 of 2 | 1 each, of 60 a day on Pro | $0.0005 | 4.3 s |
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
GPT vs DeepSeek, prompt by prompt
Each data analysis prompt with GPT-6 Luna and DeepSeek V4.1 Flash side by side, as they replied: for each side, the model that passed the most data analysis prompts (then the most hard ones, then the cheaper).
Data analysis: GPT-6 Luna passed 5 of 5 and DeepSeek V4.1 Flash 5 of 5. GPT-6 Luna answered 1.2× sooner at the median, 2.5 s against 3.0 s. GPT-6 Luna cost 2.9× less, $0.0009 against $0.0026 for the 5 replies. DeepSeek V4.1 Flash's replies ran 148% longer, in tokens of reply, thinking not counted.
1. The region with the most revenue
Everyday. GPT: 3 of 3 models passed, GPT-6 Luna the cheapest at $0.0001 · DeepSeek: 2 of 2 models passed, DeepSeek V4.1 Flash the cheapest at $0.0004. The prompt in full.
GPT-6 Luna and DeepSeek V4.1 Flash both passed this one. GPT-6 Luna answered sooner (2.5 s to 3.0 s), and GPT-6 Luna's reply cost less ($0.0001 to $0.0004); GPT-6 Luna wrote 70 tokens of reply, DeepSeek V4.1 Flash 93.
GPT-6 Luna
Passed: Final answer North — 1,085.00: right.
West: $90 + $310 + $240 + $225 = $865 East: $120 + $135 + $310 + $360 = $925 North: $180 + $240 + $620 + $45 = $1,085 Final answer: North — $1,085.00
640 tokens in, 135 out (65 of them reasoning) · 2.5 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·
DeepSeek V4.1 Flash
Passed: Final answer North, with 1,085.00 in revenue.: right.
**Working (revenue = quantity × price, summed by region):** - **West:** 90 + 310 + 240 + 225 = **865** - **East:** 120 + 135 + 310 + 360 = **925** - **North:** 180 + 240 + 620 + 45 = **1085** North leads. Final answer: North, with 1,085.00 in revenue.
683 tokens in, 196 out (103 of them reasoning) · 3.0 s · $0.0004 · 1 message on Pro · answered by deepseek/deepseek-v4.1-flash via Parasail ·
2. Average order value in August
Everyday. GPT: 3 of 3 models passed, GPT-6 Luna the cheapest at $0.0001 · DeepSeek: 2 of 2 models passed, DeepSeek V4.1 Flash the cheapest at $0.0003. The prompt in full.
GPT-6 Luna and DeepSeek V4.1 Flash both passed this one. DeepSeek V4.1 Flash answered sooner (2.0 s to 2.6 s), and GPT-6 Luna's reply cost less ($0.0001 to $0.0003); GPT-6 Luna wrote 64 tokens of reply, DeepSeek V4.1 Flash 156.
GPT-6 Luna
Passed: Final answer 300.00: right.
August revenue: $240 + $310 + $620 + $225 + $360 + $45 = $1,800. There were 6 orders, so the average is $1,800 ÷ 6 = $300.00. Final answer: $300.00
647 tokens in, 131 out (67 of them reasoning) · 2.6 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·
DeepSeek V4.1 Flash
Passed: Final answer 300.00: right.
August 2026 orders (1007–1012): - 1007: 2 × 120.00 = 240.00 - 1008: 1 × 310.00 = 310.00 - 1009: 2 × 310.00 = 620.00 - 1010: 5 × 45.00 = 225.00 - 1011: 3 × 120.00 = 360.00 - 1012: 1 × 45.00 = 45.00 Total = 1,800.00 across 6 orders → 1,800.00 / 6 = 300.00 Final answer: 300.00
690 tokens in, 267 out (111 of them reasoning) · 2.0 s · $0.0003 · 1 message on Pro · answered by deepseek/deepseek-v4.1-flash via StreamLake ·
3. Revenue change from July to August
Everyday. GPT: 3 of 3 models passed, GPT-6 Luna the cheapest at $0.0001 · DeepSeek: 2 of 2 models passed, DeepSeek V4.1 Flash the cheapest at $0.0003. The prompt in full.
GPT-6 Luna and DeepSeek V4.1 Flash both passed this one. GPT-6 Luna answered sooner (2.5 s to 10.2 s), and GPT-6 Luna's reply cost less ($0.0001 to $0.0003); GPT-6 Luna wrote 66 tokens of reply, DeepSeek V4.1 Flash 357.
GPT-6 Luna
Passed: Final answer 67.4% increase: right.
July revenue: $1,075.00. August revenue: $1,800.00. Percentage change = (($1,800 − $1,075) ÷ $1,075) × 100 = 67.4% increase. Final answer: 67.4% increase
652 tokens in, 135 out (69 of them reasoning) · 2.5 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·
DeepSeek V4.1 Flash
Passed: Final answer +67.4%: right.
**July 2026 revenue:** - 1001: 2×45 = 90 - 1002: 1×120 = 120 - 1003: 1×310 = 310 - 1004: 4×45 = 180 - 1005: 3×45 = 135 - 1006: 2×120 = 240 - **Total = 1,075** **August 2026 revenue:** - 1007: 2×120 = 240 - 1008: 1×310 = 310 - 1009: 2×310 = 620 - 1010: 5×45 = 225 - 1011: 3×120 = 360 - 1012: 1×45 = 45 - **Total = 1,800** **Change:** (1,800 − 1,075) / 1,075 = 725 / 1,075 ≈ 0.6744 → +67.4% Final answer: +67.4%
695 tokens in, 456 out (99 of them reasoning) · 10.2 s · $0.0003 · 1 message on Pro · answered by deepseek/deepseek-v4.1-flash via Relace ·
4. A median, filtered two ways
Hard. GPT: 3 of 3 models passed, GPT-6 Luna the cheapest at $0.0001 · DeepSeek: 2 of 2 models passed, DeepSeek V4.1 Flash the cheapest at $0.0002. The prompt in full.
GPT-6 Luna and DeepSeek V4.1 Flash both passed this one. DeepSeek V4.1 Flash answered sooner (1.6 s to 2.5 s), and GPT-6 Luna's reply cost less ($0.0001 to $0.0002); GPT-6 Luna wrote 75 tokens of reply, DeepSeek V4.1 Flash 107.
GPT-6 Luna
Passed: Final answer 6.25 hours: right.
Billing high-priority resolution times, sorted: 2, 4, 5.5, 7, 9, 26. With six values, the median is the average of the middle two: (5.5 + 7) / 2 = 6.25 hours. Final answer: 6.25 hours
512 tokens in, 128 out (53 of them reasoning) · 2.5 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·
DeepSeek V4.1 Flash
Passed: Final answer 6.25: right.
**Filter:** team = Billing AND priority = high | ticket | hours | |---|---| | T4 | 2 | | T12 | 4 | | T1 | 5.5 | | T9 | 7 | | T6 | 9 | | T10 | 26 | That's 6 values; median = mean of the 3rd and 4th = (5.5 + 7) / 2 = **6.25** Final answer: 6.25
554 tokens in, 209 out (102 of them reasoning) · 1.6 s · $0.0002 · 1 message on Pro · answered by deepseek/deepseek-v4.1-flash via Novita ·
5. Correlation between ad spend and sign-ups
Hard. GPT: 3 of 3 models passed, GPT-6 Luna the cheapest at $0.0004 · DeepSeek: 2 of 2 models passed, DeepSeek V4.1 Flash the cheapest at $0.0013. The prompt in full.
GPT-6 Luna and DeepSeek V4.1 Flash both passed this one. DeepSeek V4.1 Flash answered sooner (4.8 s to 6.7 s), and GPT-6 Luna's reply cost less ($0.0004 to $0.0013); GPT-6 Luna wrote 99 tokens of reply, DeepSeek V4.1 Flash 213.
GPT-6 Luna
Passed: Final answer 0.97: right.
Using \(r=\frac{\sum (x-\bar{x})(y-\bar{y})}{\sqrt{\sum(x-\bar{x})^2\sum(y-\bar{y})^2}}\), the centered sums are \(23{,}912.5\), \(624{,}687.5\), and \(971.5\). Thus \(r \approx 0.9707\). Final answer: 0.97451 tokens in, 651 out (552 of them reasoning) · 6.7 s · $0.0004 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·
DeepSeek V4.1 Flash
Passed: Final answer 0.97: right.
**Working** n = 8, Σx = 6150, Σy = 434, Σxy = 357,550, Σx² = 5,352,500, Σy² = 24,516 Numerator: n·Σxy − Σx·Σy = 8(357,550) − 6150(434) = 2,860,400 − 2,669,100 = 191,300 Denominator: - n·Σx² − (Σx)² = 42,820,000 − 37,822,500 = 4,997,500 - n·Σy² − (Σy)² = 196,128 − 188,356 = 7,772 - √(4,997,500 × 7,772) ≈ 197,080.1 r = 191,300 / 197,080.1 ≈ 0.9707 Final answer: 0.97
492 tokens in, 1,515 out (1,302 of them reasoning) · 4.8 s · $0.0013 · 1 message on Pro · answered by deepseek/deepseek-v4.1-flash via Novita ·
Each GPT model against each DeepSeek model
Every GPT model against every DeepSeek model on the same 5 data analysis prompts: passes, the median wait and what the replies cost. A lead under 10% counts as close.
GPT-6 Astra vs DeepSeek V4 Pro: 5 of 5 each; DeepSeek V4 Pro answered 1.3× sooner at the median and cost 9.5× less.
GPT-6 Astra vs DeepSeek V4.1 Flash: 5 of 5 each; DeepSeek V4.1 Flash answered 1.1× sooner at the median and cost 30.1× less.
GPT-6 Sol vs DeepSeek V4 Pro: 5 of 5 each; DeepSeek V4 Pro answered 1.1× sooner at the median and cost 2.1× less. GPT-6 Sol vs DeepSeek V4 Pro, on every job.
GPT-6 Sol vs DeepSeek V4.1 Flash: 5 of 5 each; they took about as long and DeepSeek V4.1 Flash cost 6.6× less.
GPT-6 Luna vs DeepSeek V4 Pro: 5 of 5 each; they took about as long and GPT-6 Luna cost 9.2× less. GPT-6 Luna vs DeepSeek V4 Pro, on every job.
GPT-6 Luna vs DeepSeek V4.1 Flash: 5 of 5 each; GPT-6 Luna answered 1.2× sooner at the median and cost 2.9× less. GPT-6 Luna vs DeepSeek V4.1 Flash, on every job.
How the data analysis replies are scored
Every GPT and DeepSeek reply above was checked the same way as every other model's, by the rules published with the prompts: how the data analysis prompts are scored, and each one in full.
The differences at a glance
What follows from each model's facts in our catalog.
The lineups
GPT: 3 models, GPT-6 Astra, GPT-6 Sol, and GPT-6 Luna. DeepSeek: 2 models, DeepSeek V4 Pro and DeepSeek V4.1 Flash.
Price per message
The least expensive GPT model is GPT-6 Luna (60 messages a day on Pro); the least expensive DeepSeek model is DeepSeek V4.1 Flash (60 messages a day on Pro).
Context window
GPT goes up to 1.05M tokens (GPT-6 Astra); DeepSeek up to 1.05M tokens (DeepSeek V4 Pro).
Images and PDFs
DeepSeek V4 Pro doesn't read images. DeepSeek V4 Pro and DeepSeek V4.1 Flash get a PDF's text rather than the file itself.
On the Free plan
Free's one-time trial of 5 messages covers GPT-6 Sol, GPT-6 Luna, DeepSeek V4 Pro, and DeepSeek V4.1 Flash. Paid plans have every model, with messages every month.
Every GPT and DeepSeek model's context window, files and API price: DeepSeek vs ChatGPT.
Where your messages go
In llmwise, a message to GPT goes to its maker, OpenAI, or through OpenRouter when llmwise can't reach the maker directly. DeepSeek models are served only through OpenRouter, by endpoints that don't store or train on prompts. The Privacy Policy has the details.
Questions
Which is better, ChatGPT or DeepSeek for data analysis?
In our data analysis test runs on September 27, 2026, GPT's 3 models passed 15 of 15; GPT-6 Astra, GPT-6 Sol, and GPT-6 Luna each passed 5 of 5. DeepSeek's 2 models passed 10 of 10; DeepSeek V4 Pro and DeepSeek V4.1 Flash each passed 5 of 5. GPT-6 Luna and DeepSeek V4.1 Flash each passed 5 of the 5 prompts, so these data analysis prompts don't split GPT and DeepSeek; the replies on this page show how they differ.
Is GPT or DeepSeek cheaper?
In llmwise, the least expensive GPT model is GPT-6 Luna (60 messages a day on Pro), and the least expensive DeepSeek model is DeepSeek V4.1 Flash (60 messages a day on Pro). At API list prices (September 2026), a typical message of 4,000 tokens in and 700 out costs $0.0008 on GPT-6 Luna and $0.0011 on DeepSeek V4.1 Flash.
Can I use GPT and DeepSeek in the same chat?
Yes. Ask GPT-6 Luna a question, then switch the picker to DeepSeek V4.1 Flash and ask again: DeepSeek V4.1 Flash sees the whole conversation, GPT-6 Luna's answer included.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.