Model vs model
GPT-6 Sol vs Grok 4.7
GPT-6 Sol and Grok 4.7 are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.
Model prices and specs checked against OpenRouter's GPT-6 Sol page, OpenRouter's Grok 4.7 page. Updated .
Short answer
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for GPT-6 Sol. Beyond that, they read the same files and neither is easier to try. In our test runs, GPT-6 Sol passed 45 of the 50 prompts both answered and Grok 4.7 47; 4 prompts split them, most on writing (4 to 3).
GPT-6 Sol vs Grok 4.7, prompt by prompt
Every prompt GPT-6 Sol and Grok 4.7 both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.
Of the 50 prompts both answered, both passed 44, only GPT-6 Sol passed 1, only Grok 4.7 passed 3, and neither passed 2. GPT-6 Sol answered sooner on 47 of the 50 and Grok 4.7 on 3; the rest were within 10% of each other. The 50 replies cost $0.1297 on GPT-6 Sol and $0.3163 on Grok 4.7: 2.4× less on GPT-6 Sol.
The 4 prompts only one of GPT-6 Sol and Grok 4.7 passed
Announce a second bakery shop on LinkedIn (writing): GPT-6 Sol passed and Grok 4.7 didn't. GPT-6 Sol: Graded 4.8 of 5 on average (lowest 4). Grok 4.7: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.
An article in three bullets (summarization): Grok 4.7 passed and GPT-6 Sol didn't. GPT-6 Sol: Graded 4.0 of 5 on average (lowest 3); but 66 words, over the 60 allowed. Grok 4.7: Graded 4.7 of 5 on average (lowest 4).
A late order (customer support): Grok 4.7 passed and GPT-6 Sol didn't. GPT-6 Sol: Graded 3.7 of 5 on average (lowest 1). Grok 4.7: Graded 4.3 of 5 on average (lowest 4).
Idioms into natural Japanese (translation): Grok 4.7 passed and GPT-6 Sol didn't. GPT-6 Sol: Back-translation chrF 0.40 (pass at 0.4); reads back too far from the original. Grok 4.7: Back-translation chrF 0.51 (pass at 0.4).
Job by job, the widest gaps first
Translation: GPT-6 Sol passed 4 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 4.3× sooner at the median, 3.2 s against 13.8 s. GPT-6 Sol cost 2.5× less, $0.0140 against $0.0347 for the 5 replies. GPT-6 Sol's replies ran 15% longer, in tokens of reply, thinking not counted.
Writing: GPT-6 Sol passed 4 of 5 and Grok 4.7 3 of 5. GPT-6 Sol answered 1.2× sooner at the median, 3.5 s against 4.0 s. GPT-6 Sol cost 1.9× less, $0.0113 against $0.0210 for the 5 replies. Their replies ran to about the same length.
Customer support: GPT-6 Sol passed 3 of 5 and Grok 4.7 4 of 5. GPT-6 Sol answered 2.4× sooner at the median, 3.2 s against 7.9 s. GPT-6 Sol cost 1.6× less, $0.0146 against $0.0239 for the 5 replies. Grok 4.7's replies ran 45% longer, in tokens of reply, thinking not counted.
Summarization: GPT-6 Sol passed 4 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 2.3× sooner at the median, 2.0 s against 4.7 s. GPT-6 Sol cost 1.5× less, $0.0110 against $0.0161 for the 5 replies. Their replies ran to about the same length.
Coding: GPT-6 Sol passed 5 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 5.3× sooner at the median, 4.8 s against 25.3 s. GPT-6 Sol cost 4.2× less, $0.0262 against $0.1099 for the 5 replies. Grok 4.7's replies ran 30% longer, in tokens of reply, thinking not counted.
Math: GPT-6 Sol passed 5 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 4.1× sooner at the median, 2.2 s against 9.1 s. GPT-6 Sol cost 2.8× less, $0.0090 against $0.0251 for the 5 replies. Grok 4.7's replies ran 62% longer, in tokens of reply, thinking not counted.
Data analysis: GPT-6 Sol passed 5 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 3.4× sooner at the median, 2.8 s against 9.4 s. GPT-6 Sol cost 2.5× less, $0.0169 against $0.0423 for the 5 replies. Grok 4.7's replies ran 48% longer, in tokens of reply, thinking not counted.
SQL: GPT-6 Sol passed 5 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 3.7× sooner at the median, 1.3 s against 4.9 s. GPT-6 Sol cost 1.8× less, $0.0104 against $0.0189 for the 5 replies. GPT-6 Sol's replies ran 15% longer, in tokens of reply, thinking not counted.
Agents and tool use: GPT-6 Sol passed 5 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 1.3× sooner at the median, 1.9 s against 2.5 s. GPT-6 Sol cost 1.6× less, $0.0081 against $0.0129 for the 5 replies. Their replies ran to about the same length.
RAG and answering from documents: GPT-6 Sol passed 5 of 5 and Grok 4.7 5 of 5. GPT-6 Sol answered 1.7× sooner at the median, 1.1 s against 1.9 s. GPT-6 Sol cost 1.4× less, $0.0081 against $0.0115 for the 5 replies. GPT-6 Sol's replies ran 25% longer, in tokens of reply, thinking not counted.
All 50 prompts: who passed, who answered sooner, who cost less
| Prompt | Result | Sooner | Cheaper |
|---|---|---|---|
| Turn a title into a URL slug | Both passed | GPT-6 Sol, 2.1×took 3.3 s and 6.7 s | GPT-6 Sol, 1.7×cost $0.0028 and $0.0048 |
| Parse a duration like “1h 30m” | Both passed | GPT-6 Sol, 5.3×took 4.8 s and 25.3 s | GPT-6 Sol, 3.6×cost $0.0042 and $0.0151 |
| Merge overlapping intervals | Both passed | GPT-6 Sol, 2.3×took 1.7 s and 3.8 s | GPT-6 Sol, 1.1×cost $0.0022 and $0.0025 |
| Evaluate an arithmetic expression, no eval | Both passed | GPT-6 Sol, 12.8×took 9.7 s and 124.9 s | GPT-6 Sol, 5.3×cost $0.0089 and $0.0472 |
| Parse CSV with quoted fields | Both passed | GPT-6 Sol, 11.3×took 8.9 s and 100.4 s | GPT-6 Sol, 5.0×cost $0.0081 and $0.0403 |
| Announce a second bakery shop on LinkedIn | Only GPT-6 Sol | GPT-6 Sol, 1.2×took 3.5 s and 4.0 s | Grok 4.7, 1.2×cost $0.0027 and $0.0022 |
| Rewrite corporate jargon in plain words | Neither passed | GPT-6 Sol, 3.9×took 5.0 s and 19.3 s | GPT-6 Sol, 2.7×cost $0.0033 and $0.0088 |
| Decline a meeting and offer two times | Both passed | GPT-6 Sol, 1.7×took 1.3 s and 2.3 s | GPT-6 Sol, 1.3×cost $0.0014 and $0.0019 |
| A product announcement with five rules | Both passed | GPT-6 Sol, 1.6×took 1.6 s and 2.6 s | GPT-6 Sol, 1.2×cost $0.0014 and $0.0017 |
| Argue both sides of free buses | Both passed | GPT-6 Sol, 2.9×took 4.2 s and 12.1 s | GPT-6 Sol, 2.6×cost $0.0025 and $0.0065 |
| A discount, then sales tax | Both passed | GPT-6 Sol, 2.3×took 1.7 s and 3.9 s | GPT-6 Sol, 2.2×cost $0.0013 and $0.0027 |
| Pens at 3 for $4 | Both passed | GPT-6 Sol, 12.3×took 3.3 s and 40.7 s | GPT-6 Sol, 3.5×cost $0.0024 and $0.0087 |
| Compound interest over three years | Both passed | GPT-6 Sol, 3.9×took 1.4 s and 5.5 s | GPT-6 Sol, 2.2×cost $0.0014 and $0.0031 |
| Four-digit numbers whose digits sum to 9 | Both passed | GPT-6 Sol, 4.5×took 2.5 s and 11.1 s | GPT-6 Sol, 2.9×cost $0.0019 and $0.0056 |
| The highest of three dice is a 5 | Both passed | GPT-6 Sol, 4.1×took 2.2 s and 9.1 s | GPT-6 Sol, 2.5×cost $0.0020 and $0.0050 |
| An article in three bullets | Only Grok 4.7 | GPT-6 Sol, 4.8×took 2.0 s and 9.8 s | GPT-6 Sol, 2.5×cost $0.0020 and $0.0051 |
| An email thread in one sentence | Both passed | Grok 4.7, 1.2×took 2.0 s and 1.7 s | GPT-6 Sol, 1.3×cost $0.0014 and $0.0018 |
| Decisions and action items from a meeting | Both passed | GPT-6 Sol, 1.4×took 3.5 s and 4.7 s | Grok 4.7, 1.2×cost $0.0034 and $0.0028 |
| A quarterly memo for the CEO | Both passed | GPT-6 Sol, 2.3×took 2.6 s and 6.1 s | GPT-6 Sol, 1.8×cost $0.0024 and $0.0044 |
| A study with a negative result | Both passed | GPT-6 Sol, 1.8×took 1.7 s and 3.0 s | GPT-6 Sol, 1.2×cost $0.0017 and $0.0020 |
| The region with the most revenue | Both passed | GPT-6 Sol, 3.3×took 2.8 s and 9.4 s | GPT-6 Sol, 2.2×cost $0.0026 and $0.0057 |
| Average order value in August | Both passed | GPT-6 Sol, 2.0×took 2.8 s and 5.5 s | GPT-6 Sol, 1.5×cost $0.0025 and $0.0037 |
| Revenue change from July to August | Both passed | GPT-6 Sol, 3.6×took 2.8 s and 10.1 s | GPT-6 Sol, 2.9×cost $0.0025 and $0.0071 |
| A median, filtered two ways | Both passed | GPT-6 Sol, 1.6×took 2.5 s and 4.2 s | GPT-6 Sol, 1.5×cost $0.0024 and $0.0036 |
| Correlation between ad spend and sign-ups | Both passed | GPT-6 Sol, 6.9×took 6.6 s and 45.4 s | GPT-6 Sol, 3.2×cost $0.0069 and $0.0222 |
| A late order | Only Grok 4.7 | GPT-6 Sol, 5.4×took 3.2 s and 17.2 s | GPT-6 Sol, 3.1×cost $0.0025 and $0.0076 |
| A return inside the window | Both passed | GPT-6 Sol, 3.2×took 1.5 s and 4.7 s | GPT-6 Sol, 1.9×cost $0.0017 and $0.0032 |
| A frustrated customer | Neither passed | GPT-6 Sol, 1.6×took 5.9 s and 9.4 s | GPT-6 Sol, 1.2×cost $0.0043 and $0.0050 |
| A refund request outside the window | Both passed | GPT-6 Sol, 2.5×took 3.0 s and 7.6 s | GPT-6 Sol, 1.7×cost $0.0025 and $0.0041 |
| A message with a planted instruction | Both passed | GPT-6 Sol, 1.8×took 4.4 s and 7.9 s | Closecost $0.0037 and $0.0040 |
| A delivery message into Spanish | Both passed | GPT-6 Sol, 4.3×took 3.2 s and 13.8 s | GPT-6 Sol, 2.5×cost $0.0026 and $0.0066 |
| A product description into French | Both passed | GPT-6 Sol, 10.8×took 2.1 s and 22.7 s | GPT-6 Sol, 5.2×cost $0.0019 and $0.0097 |
| A meeting note into German | Both passed | GPT-6 Sol, 2.5×took 5.0 s and 12.6 s | GPT-6 Sol, 1.7×cost $0.0035 and $0.0060 |
| Idioms into natural Japanese | Only Grok 4.7 | GPT-6 Sol, 4.2×took 2.4 s and 10.0 s | GPT-6 Sol, 2.3×cost $0.0021 and $0.0048 |
| A lease clause into Brazilian Portuguese | Both passed | GPT-6 Sol, 3.2×took 4.6 s and 15.0 s | GPT-6 Sol, 1.9×cost $0.0040 and $0.0076 |
| Customers in one country | Both passed | GPT-6 Sol, 1.7×took 1.1 s and 1.8 s | GPT-6 Sol, 1.5×cost $0.0012 and $0.0018 |
| Count orders by status | Both passed | GPT-6 Sol, 1.3×took 1.1 s and 1.4 s | GPT-6 Sol, 1.4×cost $0.0012 and $0.0017 |
| Revenue by category | Both passed | GPT-6 Sol, 3.7×took 1.3 s and 4.9 s | GPT-6 Sol, 1.7×cost $0.0018 and $0.0031 |
| Every customer, even those without orders | Both passed | GPT-6 Sol, 3.7×took 1.4 s and 5.3 s | GPT-6 Sol, 2.0×cost $0.0018 and $0.0037 |
| Monthly revenue with a running total | Both passed | GPT-6 Sol, 3.8×took 4.5 s and 16.9 s | GPT-6 Sol, 2.0×cost $0.0043 and $0.0086 |
| A fact from one section | Both passed | GPT-6 Sol, 2.4×took 1.1 s and 2.7 s | GPT-6 Sol, 1.7×cost $0.0015 and $0.0025 |
| Core hours and start times | Both passed | GPT-6 Sol, 1.7×took 1.1 s and 1.9 s | GPT-6 Sol, 1.4×cost $0.0016 and $0.0022 |
| Two sections in one answer | Both passed | GPT-6 Sol, 1.2×took 1.2 s and 1.5 s | Closecost $0.0016 and $0.0018 |
| A later amendment changes the answer | Both passed | GPT-6 Sol, 2.8×took 2.0 s and 5.5 s | GPT-6 Sol, 1.5×cost $0.0020 and $0.0030 |
| A question the handbook doesn't answer | Both passed | GPT-6 Sol, 1.9×took 0.9 s and 1.8 s | GPT-6 Sol, 1.5×cost $0.0014 and $0.0021 |
| Pick the tool and work out the date | Both passed | Grok 4.7, 1.4×took 2.2 s and 1.5 s | GPT-6 Sol, 1.2×cost $0.0016 and $0.0018 |
| Convert a currency | Both passed | Grok 4.7, 1.2×took 1.8 s and 1.5 s | GPT-6 Sol, 1.1×cost $0.0012 and $0.0013 |
| Book a meeting from a sentence | Both passed | GPT-6 Sol, 4.9×took 1.9 s and 9.1 s | GPT-6 Sol, 2.3×cost $0.0021 and $0.0048 |
| Search, but don't book | Both passed | GPT-6 Sol, 2.2×took 1.1 s and 2.5 s | GPT-6 Sol, 1.5×cost $0.0014 and $0.0020 |
| Two calls with a unit conversion | Both passed | GPT-6 Sol, 1.6×took 2.8 s and 4.5 s | GPT-6 Sol, 1.6×cost $0.0018 and $0.0029 |
GPT-6 Sol vs Grok 4.7 in our test runs
GPT-6 Sol and Grok 4.7 on the same prompts, job by job: how many replies passed their check.
Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
| Job | GPT-6 Sol | Grok 4.7 |
|---|---|---|
| Coding | 5 of 5 | 5 of 5 |
| Writing | 4 of 5 | 3 of 5 |
| Math | 5 of 5 | 5 of 5 |
| Summarization | 4 of 5 | 5 of 5 |
| Data analysis | 5 of 5 | 5 of 5 |
| Customer support | 3 of 5 | 4 of 5 |
| Translation | 4 of 5 | 5 of 5 |
| SQL | 5 of 5 | 5 of 5 |
| RAG and answering from documents | 5 of 5 | 5 of 5 |
| Agents and tool use | 5 of 5 | 5 of 5 |
| All jobs | 45 of 50 · 1 message each · $0.0026 a reply | 47 of 50 · 1 message each · $0.0063 a reply |
The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.
Where they differed most, writing: both replies to a prompt only one of them passed.
An everyday prompt: “Announce a second bakery shop on LinkedIn”, in full.
GPT-6 Sol
Passed: Graded 4.8 of 5 on average (lowest 4).
We’re happy to share that Crumb & Co. is opening a second shop in Leeds on 14 November. When we opened our first bakery, we hoped to make good bread and bakes for our neighbours. The support we’ve had since then has meant more than we can say. Opening another shop is a big step for our small team, and we’re grateful to everyone who has helped us get here. If you’re in Leeds on opening day, please stop by and say hello. Coffee is on us all day, and we’d love to welcome you into the new space. Thank you for making room for Crumb & Co. in your routine. We can’t wait to meet more of our neighbours.
389 tokens in, 189 out (39 of them reasoning) · 3.5 s · $0.0027 · 1 message on Pro · answered by openai/gpt-6-sol via OpenAI ·
Grok 4.7
Failed: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.
We're opening a second Crumb & Co. shop in Leeds on 14 November. The new bakery will offer the same bread, cakes and pastries you already know from our first shop, baked fresh each morning. If you've been asking for a Leeds location, this one is for you. On opening day, coffee is free with any purchase. Come in, say hello, and try something from the counter. We'd love to meet you and hear what you'd like to see on the shelves. Thank you for supporting a small bakery. We can't wait to welcome you.
1,624 tokens in, 207 out (93 of them reasoning) · 4.0 s · $0.0022 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·
GPT-6 Sol and Grok 4.7 on every plan
Whether the one-time free trial reaches each model, then each paid plan's messages on it.
| Plan | Price | GPT-6 Sol | Grok 4.7 |
|---|---|---|---|
| Free | $0 | In the one-time trial of 5 messages | In the one-time trial of 5 messages |
| Pro | $20 a month | Up to 125 a month | Up to 250 a month |
| Max | $50 a month | Up to 400 a month | Up to 800 a month |
| Ultra | $100 a month | Up to 900 a month | Up to 1,800 a month |
| Studio | $200 a month | Up to 2,000 a month | Up to 4,000 a month |
Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
What differs
Messages on Pro
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for GPT-6 Sol.
Context window
GPT-6 Sol takes up to 1.05M tokens; Grok 4.7 up to 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Images and PDFs
Both read images. Both take a PDF as the whole file, pages and all.
On Free
Both are in the free trial.
Where messages go
GPT-6 Sol: Sent to OpenAI directly. Grok 4.7: Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts.
Fact by fact
| Fact | GPT-6 Sol | Grok 4.7 |
|---|---|---|
| Context window | 1.05M tokens | 500K tokens |
| Reads images | Yes | Yes |
| PDFs | Whole file | Whole file |
| Reasoning | Yes | Yes |
| API price (September 2026) | $2.00 in / $10.00 out per million tokens | $1.60 in / $4.80 out per million tokens |
| A typical message at API prices (4,000 tokens in, 700 out) | $0.0150 | $0.0098 |
| A $10 top-up adds | 100 messages | 200 messages |
| Where a message goes | Sent to OpenAI directly. | Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. |
| If the provider fails | If OpenAI fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to GPT-6 Sol through OpenRouter instead. | xAI is its only host, so there's no other host to move to: if xAI fails, send the message again or pick another model. |
| Anthropic's safety fallback | Doesn't apply | Doesn't apply |
GPT-6 Sol or Grok 4.7?
From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.
Pick Grok 4.7: more messages for the money, up to 250 messages a month on Pro.
Each model's page, the families, and other pairs
GPT-6 Sol vs Grok 4.7 is one pair of models. The page below covers the whole families.
- GPT-6 Sol: price, limits and messages on every plan
- Grok 4.7: price, limits and messages on every plan
- Grok vs ChatGPT
- Claude Opus 5.5 vs GPT-6 Sol
- GPT-6 Astra vs GPT-6 Sol
- GPT-6 Sol vs GPT-6 Luna
- Claude Fable 5.1 vs Grok 4.7
- GPT-6 Astra vs Grok 4.7
- Claude Sonnet 5 vs Grok 4.7
- Every model-vs-model page
GPT-6 Sol is a GPT model; Grok 4.7 is a Grok model.
Questions
Is GPT-6 Sol or Grok 4.7 cheaper in llmwise?
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for GPT-6 Sol. Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.
Can I try GPT-6 Sol and Grok 4.7 for free?
Yes: both are in the free trial of 5 messages.
Which has the bigger context window, GPT-6 Sol or Grok 4.7?
GPT-6 Sol: 1.05M tokens, against 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Can I use GPT-6 Sol and Grok 4.7 in the same chat?
Yes. Pick GPT-6 Sol for one message and Grok 4.7 for the next; the second sees the whole chat, including the first one's answer.
Which did better in your test runs, GPT-6 Sol or Grok 4.7?
On the same 50 prompts, run on September 27, 2026, GPT-6 Sol passed 45 and Grok 4.7 passed 47. The table on this page has each job, and every reply is published.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.