Model vs model
Gemini 3.1 Pro (preview) vs Grok 4.7
Gemini 3.1 Pro (preview) and Grok 4.7 are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.
Model prices and specs checked against OpenRouter's Gemini 3.1 Pro page, OpenRouter's Grok 4.7 page. Updated .
Short answer
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Gemini 3.1 Pro (preview). Beyond that, they read the same files and neither is easier to try. In our test runs, Gemini 3.1 Pro (preview) passed 49 of the 50 prompts both answered and Grok 4.7 47; 4 prompts split them, most on writing (4 to 3).
Gemini 3.1 Pro (preview) vs Grok 4.7, prompt by prompt
Every prompt Gemini 3.1 Pro (preview) and Grok 4.7 both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.
Of the 50 prompts both answered, both passed 46, only Gemini 3.1 Pro (preview) passed 3, only Grok 4.7 passed 1, and neither passed 0. Gemini 3.1 Pro (preview) answered sooner on 18 of the 50 and Grok 4.7 on 26; the rest were within 10% of each other. The 50 replies cost $0.5334 on Gemini 3.1 Pro (preview) and $0.3163 on Grok 4.7: 1.7× less on Grok 4.7.
The 4 prompts only one of Gemini 3.1 Pro (preview) and Grok 4.7 passed
Announce a second bakery shop on LinkedIn (writing): Gemini 3.1 Pro (preview) passed and Grok 4.7 didn't. Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3). Grok 4.7: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.
Rewrite corporate jargon in plain words (writing): Gemini 3.1 Pro (preview) passed and Grok 4.7 didn't. Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3). Grok 4.7: Graded 3.7 of 5 on average (lowest 3).
Argue both sides of free buses (writing): Grok 4.7 passed and Gemini 3.1 Pro (preview) didn't. Gemini 3.1 Pro (preview): Graded 3.7 of 5 on average (lowest 3). Grok 4.7: Graded 4.3 of 5 on average (lowest 4).
A frustrated customer (customer support): Gemini 3.1 Pro (preview) passed and Grok 4.7 didn't. Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3). Grok 4.7: Graded 3.3 of 5 on average (lowest 3).
Job by job, the widest gaps first
Customer support: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 4 of 5. Grok 4.7 answered 1.1× sooner at the median, 8.7 s against 7.9 s. Grok 4.7 cost 2.2× less, $0.0529 against $0.0239 for the 5 replies. Their replies ran to about the same length.
Writing: Gemini 3.1 Pro (preview) passed 4 of 5 and Grok 4.7 3 of 5. Grok 4.7 answered 2.3× sooner at the median, 9.1 s against 4.0 s. Grok 4.7 cost 2.2× less, $0.0459 against $0.0210 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 38% longer, in tokens of reply, thinking not counted.
Summarization: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 1.7× sooner at the median, 8.0 s against 4.7 s. Grok 4.7 cost 2.7× less, $0.0441 against $0.0161 for the 5 replies. Grok 4.7's replies ran 11% longer, in tokens of reply, thinking not counted.
RAG and answering from documents: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 3.0× sooner at the median, 5.6 s against 1.9 s. Grok 4.7 cost 2.5× less, $0.0290 against $0.0115 for the 5 replies. Their replies ran to about the same length.
Agents and tool use: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 2.1× sooner at the median, 5.3 s against 2.5 s. Grok 4.7 cost 2.2× less, $0.0284 against $0.0129 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 77% longer, in tokens of reply, thinking not counted.
SQL: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 1.5× sooner at the median, 7.4 s against 4.9 s. Grok 4.7 cost 2.0× less, $0.0386 against $0.0189 for the 5 replies. Their replies ran to about the same length.
Translation: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Gemini 3.1 Pro (preview) answered 1.3× sooner at the median, 10.3 s against 13.8 s. Grok 4.7 cost 2.0× less, $0.0699 against $0.0347 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 13% longer, in tokens of reply, thinking not counted.
Math: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Their median waits were close, 8.3 s against 9.1 s. Grok 4.7 cost 2.0× less, $0.0501 against $0.0251 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 92% longer, in tokens of reply, thinking not counted.
Data analysis: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Their median waits were close, 8.6 s against 9.4 s. Grok 4.7 cost 1.8× less, $0.0773 against $0.0423 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 178% longer, in tokens of reply, thinking not counted.
Coding: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Gemini 3.1 Pro (preview) answered 2.5× sooner at the median, 10.3 s against 25.3 s. Gemini 3.1 Pro (preview) cost 1.1× less, $0.0973 against $0.1099 for the 5 replies. Their replies ran to about the same length.
All 50 prompts: who passed, who answered sooner, who cost less
| Prompt | Result | Sooner | Cheaper |
|---|---|---|---|
| Turn a title into a URL slug | Both passed | Gemini 3.1 Pro (preview), 1.1×took 6.0 s and 6.7 s | Grok 4.7, 1.2×cost $0.0056 and $0.0048 |
| Parse a duration like “1h 30m” | Both passed | Gemini 3.1 Pro (preview), 2.5×took 10.3 s and 25.3 s | Grok 4.7, 1.1×cost $0.0172 and $0.0151 |
| Merge overlapping intervals | Both passed | Grok 4.7, 2.2×took 8.6 s and 3.8 s | Grok 4.7, 4.2×cost $0.0103 and $0.0025 |
| Evaluate an arithmetic expression, no eval | Both passed | Gemini 3.1 Pro (preview), 6.1×took 20.6 s and 124.9 s | Gemini 3.1 Pro (preview), 1.3×cost $0.0351 and $0.0472 |
| Parse CSV with quoted fields | Both passed | Gemini 3.1 Pro (preview), 4.9×took 20.6 s and 100.4 s | Gemini 3.1 Pro (preview), 1.4×cost $0.0291 and $0.0403 |
| Announce a second bakery shop on LinkedIn | Only Gemini 3.1 Pro (preview) | Grok 4.7, 2.4×took 9.5 s and 4.0 s | Grok 4.7, 5.1×cost $0.0113 and $0.0022 |
| Rewrite corporate jargon in plain words | Only Gemini 3.1 Pro (preview) | Gemini 3.1 Pro (preview), 2.2×took 8.8 s and 19.3 s | Grok 4.7, 1.1×cost $0.0101 and $0.0088 |
| Decline a meeting and offer two times | Both passed | Grok 4.7, 2.7×took 6.2 s and 2.3 s | Grok 4.7, 3.1×cost $0.0058 and $0.0019 |
| A product announcement with five rules | Both passed | Grok 4.7, 3.5×took 9.2 s and 2.6 s | Grok 4.7, 5.7×cost $0.0095 and $0.0017 |
| Argue both sides of free buses | Only Grok 4.7 | Gemini 3.1 Pro (preview), 1.3×took 9.1 s and 12.1 s | Grok 4.7, 1.4×cost $0.0092 and $0.0065 |
| A discount, then sales tax | Both passed | Closetook 3.6 s and 3.9 s | Grok 4.7, 1.5×cost $0.0042 and $0.0027 |
| Pens at 3 for $4 | Both passed | Gemini 3.1 Pro (preview), 4.2×took 9.6 s and 40.7 s | Grok 4.7, 1.4×cost $0.0121 and $0.0087 |
| Compound interest over three years | Both passed | Grok 4.7, 1.1×took 6.2 s and 5.5 s | Grok 4.7, 3.2×cost $0.0099 and $0.0031 |
| Four-digit numbers whose digits sum to 9 | Both passed | Gemini 3.1 Pro (preview), 1.2×took 9.6 s and 11.1 s | Grok 4.7, 2.7×cost $0.0150 and $0.0056 |
| The highest of three dice is a 5 | Both passed | Closetook 8.3 s and 9.1 s | Grok 4.7, 1.8×cost $0.0090 and $0.0050 |
| An article in three bullets | Both passed | Gemini 3.1 Pro (preview), 1.2×took 8.0 s and 9.8 s | Grok 4.7, 1.7×cost $0.0088 and $0.0051 |
| An email thread in one sentence | Both passed | Grok 4.7, 3.6×took 6.3 s and 1.7 s | Grok 4.7, 3.2×cost $0.0058 and $0.0018 |
| Decisions and action items from a meeting | Both passed | Grok 4.7, 1.6×took 7.6 s and 4.7 s | Grok 4.7, 2.9×cost $0.0081 and $0.0028 |
| A quarterly memo for the CEO | Both passed | Grok 4.7, 1.8×took 11.2 s and 6.1 s | Grok 4.7, 3.2×cost $0.0139 and $0.0044 |
| A study with a negative result | Both passed | Grok 4.7, 2.7×took 8.2 s and 3.0 s | Grok 4.7, 3.7×cost $0.0075 and $0.0020 |
| The region with the most revenue | Both passed | Gemini 3.1 Pro (preview), 1.1×took 8.5 s and 9.4 s | Grok 4.7, 2.1×cost $0.0117 and $0.0057 |
| Average order value in August | Both passed | Grok 4.7, 1.6×took 8.6 s and 5.5 s | Grok 4.7, 3.5×cost $0.0130 and $0.0037 |
| Revenue change from July to August | Both passed | Gemini 3.1 Pro (preview), 1.1×took 9.0 s and 10.1 s | Grok 4.7, 1.9×cost $0.0135 and $0.0071 |
| A median, filtered two ways | Both passed | Grok 4.7, 1.7×took 7.1 s and 4.2 s | Grok 4.7, 2.3×cost $0.0082 and $0.0036 |
| Correlation between ad spend and sign-ups | Both passed | Gemini 3.1 Pro (preview), 2.7×took 16.5 s and 45.4 s | Grok 4.7, 1.4×cost $0.0310 and $0.0222 |
| A late order | Both passed | Gemini 3.1 Pro (preview), 2.0×took 8.7 s and 17.2 s | Grok 4.7, 1.6×cost $0.0125 and $0.0076 |
| A return inside the window | Both passed | Grok 4.7, 1.5×took 6.9 s and 4.7 s | Grok 4.7, 2.4×cost $0.0076 and $0.0032 |
| A frustrated customer | Only Gemini 3.1 Pro (preview) | Closetook 9.8 s and 9.4 s | Grok 4.7, 2.2×cost $0.0110 and $0.0050 |
| A refund request outside the window | Both passed | Closetook 7.4 s and 7.6 s | Grok 4.7, 2.5×cost $0.0103 and $0.0041 |
| A message with a planted instruction | Both passed | Grok 4.7, 1.4×took 10.6 s and 7.9 s | Grok 4.7, 2.9×cost $0.0116 and $0.0040 |
| A delivery message into Spanish | Both passed | Gemini 3.1 Pro (preview), 1.2×took 11.8 s and 13.8 s | Grok 4.7, 2.4×cost $0.0157 and $0.0066 |
| A product description into French | Both passed | Gemini 3.1 Pro (preview), 2.5×took 9.1 s and 22.7 s | Grok 4.7, 1.5×cost $0.0141 and $0.0097 |
| A meeting note into German | Both passed | Closetook 13.0 s and 12.6 s | Grok 4.7, 2.6×cost $0.0158 and $0.0060 |
| Idioms into natural Japanese | Both passed | Closetook 10.3 s and 10.0 s | Grok 4.7, 2.5×cost $0.0118 and $0.0048 |
| A lease clause into Brazilian Portuguese | Both passed | Gemini 3.1 Pro (preview), 1.9×took 8.1 s and 15.0 s | Grok 4.7, 1.6×cost $0.0125 and $0.0076 |
| Customers in one country | Both passed | Grok 4.7, 2.7×took 4.9 s and 1.8 s | Grok 4.7, 1.6×cost $0.0029 and $0.0018 |
| Count orders by status | Both passed | Grok 4.7, 3.5×took 5.0 s and 1.4 s | Grok 4.7, 2.3×cost $0.0039 and $0.0017 |
| Revenue by category | Both passed | Grok 4.7, 1.5×took 7.4 s and 4.9 s | Grok 4.7, 2.4×cost $0.0074 and $0.0031 |
| Every customer, even those without orders | Both passed | Grok 4.7, 1.5×took 7.9 s and 5.3 s | Grok 4.7, 2.4×cost $0.0089 and $0.0037 |
| Monthly revenue with a running total | Both passed | Gemini 3.1 Pro (preview), 1.5×took 11.1 s and 16.9 s | Grok 4.7, 1.8×cost $0.0155 and $0.0086 |
| A fact from one section | Both passed | Grok 4.7, 2.0×took 5.3 s and 2.7 s | Grok 4.7, 1.8×cost $0.0045 and $0.0025 |
| Core hours and start times | Both passed | Grok 4.7, 3.1×took 5.8 s and 1.9 s | Grok 4.7, 2.3×cost $0.0052 and $0.0022 |
| Two sections in one answer | Both passed | Grok 4.7, 3.8×took 5.6 s and 1.5 s | Grok 4.7, 2.7×cost $0.0047 and $0.0018 |
| A later amendment changes the answer | Both passed | Grok 4.7, 1.5×took 8.6 s and 5.5 s | Grok 4.7, 3.4×cost $0.0099 and $0.0030 |
| A question the handbook doesn't answer | Both passed | Grok 4.7, 3.1×took 5.6 s and 1.8 s | Grok 4.7, 2.2×cost $0.0046 and $0.0021 |
| Pick the tool and work out the date | Both passed | Grok 4.7, 2.8×took 4.3 s and 1.5 s | Grok 4.7, 2.7×cost $0.0048 and $0.0018 |
| Convert a currency | Both passed | Grok 4.7, 3.5×took 5.3 s and 1.5 s | Grok 4.7, 3.3×cost $0.0044 and $0.0013 |
| Book a meeting from a sentence | Both passed | Gemini 3.1 Pro (preview), 1.3×took 7.1 s and 9.1 s | Grok 4.7, 1.4×cost $0.0068 and $0.0048 |
| Search, but don't book | Both passed | Grok 4.7, 1.9×took 4.7 s and 2.5 s | Grok 4.7, 3.0×cost $0.0061 and $0.0020 |
| Two calls with a unit conversion | Both passed | Grok 4.7, 1.5×took 6.9 s and 4.5 s | Grok 4.7, 2.2×cost $0.0063 and $0.0029 |
Gemini 3.1 Pro (preview) vs Grok 4.7 in our test runs
Gemini 3.1 Pro (preview) and Grok 4.7 on the same prompts, job by job: how many replies passed their check.
Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
| Job | Gemini 3.1 Pro (preview) | Grok 4.7 |
|---|---|---|
| Coding | 5 of 5 | 5 of 5 |
| Writing | 4 of 5 | 3 of 5 |
| Math | 5 of 5 | 5 of 5 |
| Summarization | 5 of 5 | 5 of 5 |
| Data analysis | 5 of 5 | 5 of 5 |
| Customer support | 5 of 5 | 4 of 5 |
| Translation | 5 of 5 | 5 of 5 |
| SQL | 5 of 5 | 5 of 5 |
| RAG and answering from documents | 5 of 5 | 5 of 5 |
| Agents and tool use | 5 of 5 | 5 of 5 |
| All jobs | 49 of 50 · 1 message each · $0.0107 a reply | 47 of 50 · 1 message each · $0.0063 a reply |
The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.
Where they differed most, writing: both replies to a prompt only one of them passed.
An everyday prompt: “Announce a second bakery shop on LinkedIn”, in full.
Gemini 3.1 Pro
Passed: Graded 4.0 of 5 on average (lowest 3).
We are absolutely thrilled to share some wonderful news with you all. Our little bakery, Crumb & Co., is growing! Thanks to your incredible support over the years, we are opening the doors to our second shop in Leeds on the 14th of November. When we first started baking, we never imagined we would be able to share our pastries and breads with so many lovely people. This new chapter in Leeds means the world to us, and we cannot wait to become a part of the local community. To celebrate this exciting milestone, we would love for you to join us on our opening day. Pop in to say hello, check out the new space, and enjoy a free cup of coffee on us. We are so grateful for your continued support and look forward to welcoming you to our new home very soon.
417 tokens in, 876 out (707 of them reasoning) · 9.5 s · $0.0113 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
Grok 4.7
Failed: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.
We're opening a second Crumb & Co. shop in Leeds on 14 November. The new bakery will offer the same bread, cakes and pastries you already know from our first shop, baked fresh each morning. If you've been asking for a Leeds location, this one is for you. On opening day, coffee is free with any purchase. Come in, say hello, and try something from the counter. We'd love to meet you and hear what you'd like to see on the shelves. Thank you for supporting a small bakery. We can't wait to welcome you.
1,624 tokens in, 207 out (93 of them reasoning) · 4.0 s · $0.0022 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·
Gemini 3.1 Pro (preview) and Grok 4.7 on every plan
Whether the one-time free trial reaches each model, then each paid plan's messages on it.
| Plan | Price | Gemini 3.1 Pro (preview) | Grok 4.7 |
|---|---|---|---|
| Free | $0 | In the one-time trial of 5 messages | In the one-time trial of 5 messages |
| Pro | $20 a month | Up to 125 a month | Up to 250 a month |
| Max | $50 a month | Up to 400 a month | Up to 800 a month |
| Ultra | $100 a month | Up to 900 a month | Up to 1,800 a month |
| Studio | $200 a month | Up to 2,000 a month | Up to 4,000 a month |
Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
What differs
Messages on Pro
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Gemini 3.1 Pro (preview).
Context window
Gemini 3.1 Pro (preview) takes up to 1.05M tokens; Grok 4.7 up to 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Images and PDFs
Both read images. Both take a PDF as the whole file, pages and all.
On Free
Both are in the free trial.
Where messages go
Gemini 3.1 Pro (preview): Sent to Google directly. Grok 4.7: Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts.
Fact by fact
| Fact | Gemini 3.1 Pro (preview) | Grok 4.7 |
|---|---|---|
| Context window | 1.05M tokens | 500K tokens |
| Reads images | Yes | Yes |
| PDFs | Whole file | Whole file |
| Reasoning | Yes | Yes |
| API price (September 2026) | $2.00 in / $12.00 out per million tokens | $1.60 in / $4.80 out per million tokens |
| A typical message at API prices (4,000 tokens in, 700 out) | $0.0164 | $0.0098 |
| A $10 top-up adds | 100 messages | 200 messages |
| Where a message goes | Sent to Google directly. | Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. |
| If the provider fails | If Google fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to Gemini 3.1 Pro (preview) through OpenRouter instead. | xAI is its only host, so there's no other host to move to: if xAI fails, send the message again or pick another model. |
| Anthropic's safety fallback | Doesn't apply | Doesn't apply |
Gemini 3.1 Pro (preview) or Grok 4.7?
From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.
Pick Grok 4.7: more messages for the money, up to 250 messages a month on Pro.
Each model's page, the families, and other pairs
Gemini 3.1 Pro (preview) vs Grok 4.7 is one pair of models. The page below covers the whole families.
- Gemini 3.1 Pro: price, limits and messages on every plan
- Grok 4.7: price, limits and messages on every plan
- Gemini vs Grok
- Claude Sonnet 5.5 vs Gemini 3.1 Pro (preview)
- Claude Opus 5.5 vs Gemini 3.1 Pro (preview)
- Claude Sonnet 5 vs Gemini 3.1 Pro (preview)
- Claude Fable 5.1 vs Grok 4.7
- GPT-6 Astra vs Grok 4.7
- Claude Sonnet 5 vs Grok 4.7
- Every model-vs-model page
Gemini 3.1 Pro (preview) is a Gemini model; Grok 4.7 is a Grok model.
Questions
Is Gemini 3.1 Pro (preview) or Grok 4.7 cheaper in llmwise?
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Gemini 3.1 Pro (preview). Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.
Can I try Gemini 3.1 Pro (preview) and Grok 4.7 for free?
Yes: both are in the free trial of 5 messages.
Which has the bigger context window, Gemini 3.1 Pro (preview) or Grok 4.7?
Gemini 3.1 Pro (preview): 1.05M tokens, against 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Can I use Gemini 3.1 Pro (preview) and Grok 4.7 in the same chat?
Yes. Pick Gemini 3.1 Pro (preview) for one message and Grok 4.7 for the next; the second sees the whole chat, including the first one's answer.
Which did better in your test runs, Gemini 3.1 Pro (preview) or Grok 4.7?
On the same 50 prompts, run on September 27, 2026, Gemini 3.1 Pro (preview) passed 49 and Grok 4.7 passed 47. The table on this page has each job, and every reply is published.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.