Skip to content

Model vs model

Gemini 3.1 Pro (preview) vs Grok 4.7

Gemini 3.1 Pro (preview) and Grok 4.7 are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.

Model prices and specs checked against OpenRouter's Gemini 3.1 Pro page, OpenRouter's Grok 4.7 page. Updated .

Short answer

Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Gemini 3.1 Pro (preview). Beyond that, they read the same files and neither is easier to try. In our test runs, Gemini 3.1 Pro (preview) passed 49 of the 50 prompts both answered and Grok 4.7 47; 4 prompts split them, most on writing (4 to 3).

Gemini 3.1 Pro (preview) vs Grok 4.7, prompt by prompt

Every prompt Gemini 3.1 Pro (preview) and Grok 4.7 both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.

Of the 50 prompts both answered, both passed 46, only Gemini 3.1 Pro (preview) passed 3, only Grok 4.7 passed 1, and neither passed 0. Gemini 3.1 Pro (preview) answered sooner on 18 of the 50 and Grok 4.7 on 26; the rest were within 10% of each other. The 50 replies cost $0.5334 on Gemini 3.1 Pro (preview) and $0.3163 on Grok 4.7: 1.7× less on Grok 4.7.

The 4 prompts only one of Gemini 3.1 Pro (preview) and Grok 4.7 passed

  • Announce a second bakery shop on LinkedIn (writing): Gemini 3.1 Pro (preview) passed and Grok 4.7 didn't. Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3). Grok 4.7: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.

  • Rewrite corporate jargon in plain words (writing): Gemini 3.1 Pro (preview) passed and Grok 4.7 didn't. Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3). Grok 4.7: Graded 3.7 of 5 on average (lowest 3).

  • Argue both sides of free buses (writing): Grok 4.7 passed and Gemini 3.1 Pro (preview) didn't. Gemini 3.1 Pro (preview): Graded 3.7 of 5 on average (lowest 3). Grok 4.7: Graded 4.3 of 5 on average (lowest 4).

  • A frustrated customer (customer support): Gemini 3.1 Pro (preview) passed and Grok 4.7 didn't. Gemini 3.1 Pro (preview): Graded 4.0 of 5 on average (lowest 3). Grok 4.7: Graded 3.3 of 5 on average (lowest 3).

Job by job, the widest gaps first

  • Customer support: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 4 of 5. Grok 4.7 answered 1.1× sooner at the median, 8.7 s against 7.9 s. Grok 4.7 cost 2.2× less, $0.0529 against $0.0239 for the 5 replies. Their replies ran to about the same length.

  • Writing: Gemini 3.1 Pro (preview) passed 4 of 5 and Grok 4.7 3 of 5. Grok 4.7 answered 2.3× sooner at the median, 9.1 s against 4.0 s. Grok 4.7 cost 2.2× less, $0.0459 against $0.0210 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 38% longer, in tokens of reply, thinking not counted.

  • Summarization: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 1.7× sooner at the median, 8.0 s against 4.7 s. Grok 4.7 cost 2.7× less, $0.0441 against $0.0161 for the 5 replies. Grok 4.7's replies ran 11% longer, in tokens of reply, thinking not counted.

  • RAG and answering from documents: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 3.0× sooner at the median, 5.6 s against 1.9 s. Grok 4.7 cost 2.5× less, $0.0290 against $0.0115 for the 5 replies. Their replies ran to about the same length.

  • Agents and tool use: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 2.1× sooner at the median, 5.3 s against 2.5 s. Grok 4.7 cost 2.2× less, $0.0284 against $0.0129 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 77% longer, in tokens of reply, thinking not counted.

  • SQL: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Grok 4.7 answered 1.5× sooner at the median, 7.4 s against 4.9 s. Grok 4.7 cost 2.0× less, $0.0386 against $0.0189 for the 5 replies. Their replies ran to about the same length.

  • Translation: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Gemini 3.1 Pro (preview) answered 1.3× sooner at the median, 10.3 s against 13.8 s. Grok 4.7 cost 2.0× less, $0.0699 against $0.0347 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 13% longer, in tokens of reply, thinking not counted.

  • Math: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Their median waits were close, 8.3 s against 9.1 s. Grok 4.7 cost 2.0× less, $0.0501 against $0.0251 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 92% longer, in tokens of reply, thinking not counted.

  • Data analysis: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Their median waits were close, 8.6 s against 9.4 s. Grok 4.7 cost 1.8× less, $0.0773 against $0.0423 for the 5 replies. Gemini 3.1 Pro (preview)'s replies ran 178% longer, in tokens of reply, thinking not counted.

  • Coding: Gemini 3.1 Pro (preview) passed 5 of 5 and Grok 4.7 5 of 5. Gemini 3.1 Pro (preview) answered 2.5× sooner at the median, 10.3 s against 25.3 s. Gemini 3.1 Pro (preview) cost 1.1× less, $0.0973 against $0.1099 for the 5 replies. Their replies ran to about the same length.

All 50 prompts: who passed, who answered sooner, who cost less
Gemini 3.1 Pro (preview) and Grok 4.7 on each prompt of our test runs
PromptResultSoonerCheaper
Turn a title into a URL slugBoth passedGemini 3.1 Pro (preview), 1.1×took 6.0 s and 6.7 sGrok 4.7, 1.2×cost $0.0056 and $0.0048
Parse a duration like “1h 30m”Both passedGemini 3.1 Pro (preview), 2.5×took 10.3 s and 25.3 sGrok 4.7, 1.1×cost $0.0172 and $0.0151
Merge overlapping intervalsBoth passedGrok 4.7, 2.2×took 8.6 s and 3.8 sGrok 4.7, 4.2×cost $0.0103 and $0.0025
Evaluate an arithmetic expression, no evalBoth passedGemini 3.1 Pro (preview), 6.1×took 20.6 s and 124.9 sGemini 3.1 Pro (preview), 1.3×cost $0.0351 and $0.0472
Parse CSV with quoted fieldsBoth passedGemini 3.1 Pro (preview), 4.9×took 20.6 s and 100.4 sGemini 3.1 Pro (preview), 1.4×cost $0.0291 and $0.0403
Announce a second bakery shop on LinkedInOnly Gemini 3.1 Pro (preview)Grok 4.7, 2.4×took 9.5 s and 4.0 sGrok 4.7, 5.1×cost $0.0113 and $0.0022
Rewrite corporate jargon in plain wordsOnly Gemini 3.1 Pro (preview)Gemini 3.1 Pro (preview), 2.2×took 8.8 s and 19.3 sGrok 4.7, 1.1×cost $0.0101 and $0.0088
Decline a meeting and offer two timesBoth passedGrok 4.7, 2.7×took 6.2 s and 2.3 sGrok 4.7, 3.1×cost $0.0058 and $0.0019
A product announcement with five rulesBoth passedGrok 4.7, 3.5×took 9.2 s and 2.6 sGrok 4.7, 5.7×cost $0.0095 and $0.0017
Argue both sides of free busesOnly Grok 4.7Gemini 3.1 Pro (preview), 1.3×took 9.1 s and 12.1 sGrok 4.7, 1.4×cost $0.0092 and $0.0065
A discount, then sales taxBoth passedClosetook 3.6 s and 3.9 sGrok 4.7, 1.5×cost $0.0042 and $0.0027
Pens at 3 for $4Both passedGemini 3.1 Pro (preview), 4.2×took 9.6 s and 40.7 sGrok 4.7, 1.4×cost $0.0121 and $0.0087
Compound interest over three yearsBoth passedGrok 4.7, 1.1×took 6.2 s and 5.5 sGrok 4.7, 3.2×cost $0.0099 and $0.0031
Four-digit numbers whose digits sum to 9Both passedGemini 3.1 Pro (preview), 1.2×took 9.6 s and 11.1 sGrok 4.7, 2.7×cost $0.0150 and $0.0056
The highest of three dice is a 5Both passedClosetook 8.3 s and 9.1 sGrok 4.7, 1.8×cost $0.0090 and $0.0050
An article in three bulletsBoth passedGemini 3.1 Pro (preview), 1.2×took 8.0 s and 9.8 sGrok 4.7, 1.7×cost $0.0088 and $0.0051
An email thread in one sentenceBoth passedGrok 4.7, 3.6×took 6.3 s and 1.7 sGrok 4.7, 3.2×cost $0.0058 and $0.0018
Decisions and action items from a meetingBoth passedGrok 4.7, 1.6×took 7.6 s and 4.7 sGrok 4.7, 2.9×cost $0.0081 and $0.0028
A quarterly memo for the CEOBoth passedGrok 4.7, 1.8×took 11.2 s and 6.1 sGrok 4.7, 3.2×cost $0.0139 and $0.0044
A study with a negative resultBoth passedGrok 4.7, 2.7×took 8.2 s and 3.0 sGrok 4.7, 3.7×cost $0.0075 and $0.0020
The region with the most revenueBoth passedGemini 3.1 Pro (preview), 1.1×took 8.5 s and 9.4 sGrok 4.7, 2.1×cost $0.0117 and $0.0057
Average order value in AugustBoth passedGrok 4.7, 1.6×took 8.6 s and 5.5 sGrok 4.7, 3.5×cost $0.0130 and $0.0037
Revenue change from July to AugustBoth passedGemini 3.1 Pro (preview), 1.1×took 9.0 s and 10.1 sGrok 4.7, 1.9×cost $0.0135 and $0.0071
A median, filtered two waysBoth passedGrok 4.7, 1.7×took 7.1 s and 4.2 sGrok 4.7, 2.3×cost $0.0082 and $0.0036
Correlation between ad spend and sign-upsBoth passedGemini 3.1 Pro (preview), 2.7×took 16.5 s and 45.4 sGrok 4.7, 1.4×cost $0.0310 and $0.0222
A late orderBoth passedGemini 3.1 Pro (preview), 2.0×took 8.7 s and 17.2 sGrok 4.7, 1.6×cost $0.0125 and $0.0076
A return inside the windowBoth passedGrok 4.7, 1.5×took 6.9 s and 4.7 sGrok 4.7, 2.4×cost $0.0076 and $0.0032
A frustrated customerOnly Gemini 3.1 Pro (preview)Closetook 9.8 s and 9.4 sGrok 4.7, 2.2×cost $0.0110 and $0.0050
A refund request outside the windowBoth passedClosetook 7.4 s and 7.6 sGrok 4.7, 2.5×cost $0.0103 and $0.0041
A message with a planted instructionBoth passedGrok 4.7, 1.4×took 10.6 s and 7.9 sGrok 4.7, 2.9×cost $0.0116 and $0.0040
A delivery message into SpanishBoth passedGemini 3.1 Pro (preview), 1.2×took 11.8 s and 13.8 sGrok 4.7, 2.4×cost $0.0157 and $0.0066
A product description into FrenchBoth passedGemini 3.1 Pro (preview), 2.5×took 9.1 s and 22.7 sGrok 4.7, 1.5×cost $0.0141 and $0.0097
A meeting note into GermanBoth passedClosetook 13.0 s and 12.6 sGrok 4.7, 2.6×cost $0.0158 and $0.0060
Idioms into natural JapaneseBoth passedClosetook 10.3 s and 10.0 sGrok 4.7, 2.5×cost $0.0118 and $0.0048
A lease clause into Brazilian PortugueseBoth passedGemini 3.1 Pro (preview), 1.9×took 8.1 s and 15.0 sGrok 4.7, 1.6×cost $0.0125 and $0.0076
Customers in one countryBoth passedGrok 4.7, 2.7×took 4.9 s and 1.8 sGrok 4.7, 1.6×cost $0.0029 and $0.0018
Count orders by statusBoth passedGrok 4.7, 3.5×took 5.0 s and 1.4 sGrok 4.7, 2.3×cost $0.0039 and $0.0017
Revenue by categoryBoth passedGrok 4.7, 1.5×took 7.4 s and 4.9 sGrok 4.7, 2.4×cost $0.0074 and $0.0031
Every customer, even those without ordersBoth passedGrok 4.7, 1.5×took 7.9 s and 5.3 sGrok 4.7, 2.4×cost $0.0089 and $0.0037
Monthly revenue with a running totalBoth passedGemini 3.1 Pro (preview), 1.5×took 11.1 s and 16.9 sGrok 4.7, 1.8×cost $0.0155 and $0.0086
A fact from one sectionBoth passedGrok 4.7, 2.0×took 5.3 s and 2.7 sGrok 4.7, 1.8×cost $0.0045 and $0.0025
Core hours and start timesBoth passedGrok 4.7, 3.1×took 5.8 s and 1.9 sGrok 4.7, 2.3×cost $0.0052 and $0.0022
Two sections in one answerBoth passedGrok 4.7, 3.8×took 5.6 s and 1.5 sGrok 4.7, 2.7×cost $0.0047 and $0.0018
A later amendment changes the answerBoth passedGrok 4.7, 1.5×took 8.6 s and 5.5 sGrok 4.7, 3.4×cost $0.0099 and $0.0030
A question the handbook doesn't answerBoth passedGrok 4.7, 3.1×took 5.6 s and 1.8 sGrok 4.7, 2.2×cost $0.0046 and $0.0021
Pick the tool and work out the dateBoth passedGrok 4.7, 2.8×took 4.3 s and 1.5 sGrok 4.7, 2.7×cost $0.0048 and $0.0018
Convert a currencyBoth passedGrok 4.7, 3.5×took 5.3 s and 1.5 sGrok 4.7, 3.3×cost $0.0044 and $0.0013
Book a meeting from a sentenceBoth passedGemini 3.1 Pro (preview), 1.3×took 7.1 s and 9.1 sGrok 4.7, 1.4×cost $0.0068 and $0.0048
Search, but don't bookBoth passedGrok 4.7, 1.9×took 4.7 s and 2.5 sGrok 4.7, 3.0×cost $0.0061 and $0.0020
Two calls with a unit conversionBoth passedGrok 4.7, 1.5×took 6.9 s and 4.5 sGrok 4.7, 2.2×cost $0.0063 and $0.0029

Gemini 3.1 Pro (preview) vs Grok 4.7 in our test runs

Gemini 3.1 Pro (preview) and Grok 4.7 on the same prompts, job by job: how many replies passed their check.

Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Gemini 3.1 Pro (preview) and Grok 4.7 in our test runs, job by job
JobGemini 3.1 Pro (preview)Grok 4.7
Coding5 of 55 of 5
Writing4 of 53 of 5
Math5 of 55 of 5
Summarization5 of 55 of 5
Data analysis5 of 55 of 5
Customer support5 of 54 of 5
Translation5 of 55 of 5
SQL5 of 55 of 5
RAG and answering from documents5 of 55 of 5
Agents and tool use5 of 55 of 5
All jobs49 of 50 · 1 message each · $0.0107 a reply47 of 50 · 1 message each · $0.0063 a reply

The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.

Where they differed most, writing: both replies to a prompt only one of them passed.

An everyday prompt: “Announce a second bakery shop on LinkedIn”, in full.

  • Gemini 3.1 Pro

    Passed: Graded 4.0 of 5 on average (lowest 3).

    We are absolutely thrilled to share some wonderful news with you all. Our little bakery, Crumb & Co., is growing! Thanks to your incredible support over the years, we are opening the doors to our second shop in Leeds on the 14th of November. 
    
    When we first started baking, we never imagined we would be able to share our pastries and breads with so many lovely people. This new chapter in Leeds means the world to us, and we cannot wait to become a part of the local community. 
    
    To celebrate this exciting milestone, we would love for you to join us on our opening day. Pop in to say hello, check out the new space, and enjoy a free cup of coffee on us. We are so grateful for your continued support and look forward to welcoming you to our new home very soon.

    417 tokens in, 876 out (707 of them reasoning) · 9.5 s · $0.0113 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·

  • Grok 4.7

    Failed: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.

    We're opening a second Crumb & Co. shop in Leeds on 14 November.
    
    The new bakery will offer the same bread, cakes and pastries you already know from our first shop, baked fresh each morning. If you've been asking for a Leeds location, this one is for you.
    
    On opening day, coffee is free with any purchase. Come in, say hello, and try something from the counter. We'd love to meet you and hear what you'd like to see on the shelves.
    
    Thank you for supporting a small bakery. We can't wait to welcome you.

    1,624 tokens in, 207 out (93 of them reasoning) · 4.0 s · $0.0022 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·

Gemini 3.1 Pro (preview) and Grok 4.7 on every plan

Whether the one-time free trial reaches each model, then each paid plan's messages on it.

Messages on Gemini 3.1 Pro (preview) and Grok 4.7, plan by plan
PlanPriceGemini 3.1 Pro (preview)Grok 4.7
Free$0In the one-time trial of 5 messagesIn the one-time trial of 5 messages
Pro$20 a monthUp to 125 a monthUp to 250 a month
Max$50 a monthUp to 400 a monthUp to 800 a month
Ultra$100 a monthUp to 900 a monthUp to 1,800 a month
Studio$200 a monthUp to 2,000 a monthUp to 4,000 a month

Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

What differs

  • Messages on Pro

    Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Gemini 3.1 Pro (preview).

  • Context window

    Gemini 3.1 Pro (preview) takes up to 1.05M tokens; Grok 4.7 up to 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

  • Images and PDFs

    Both read images. Both take a PDF as the whole file, pages and all.

  • On Free

    Both are in the free trial.

  • Where messages go

    Gemini 3.1 Pro (preview): Sent to Google directly. Grok 4.7: Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts.

Fact by fact

Gemini 3.1 Pro (preview) and Grok 4.7, fact by fact
FactGemini 3.1 Pro (preview)Grok 4.7
Context window1.05M tokens500K tokens
Reads imagesYesYes
PDFsWhole fileWhole file
ReasoningYesYes
API price (September 2026)$2.00 in / $12.00 out per million tokens$1.60 in / $4.80 out per million tokens
A typical message at API prices (4,000 tokens in, 700 out)$0.0164$0.0098
A $10 top-up adds100 messages200 messages
Where a message goesSent to Google directly.Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts.
If the provider failsIf Google fails before the reply starts (an overload, a server error, a dropped connection), llmwise sends the same request to Gemini 3.1 Pro (preview) through OpenRouter instead.xAI is its only host, so there's no other host to move to: if xAI fails, send the message again or pick another model.
Anthropic's safety fallbackDoesn't applyDoesn't apply
API prices are what our model catalog lists (Gemini: Google's list price; Grok: xAI's price, served through OpenRouter). In llmwise you pay per message, not per token: the counts above are what you get.

Gemini 3.1 Pro (preview) or Grok 4.7?

From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.

  • Pick Grok 4.7: more messages for the money, up to 250 messages a month on Pro.

Each model's page, the families, and other pairs

Gemini 3.1 Pro (preview) vs Grok 4.7 is one pair of models. The page below covers the whole families.

Questions

Is Gemini 3.1 Pro (preview) or Grok 4.7 cheaper in llmwise?

Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Gemini 3.1 Pro (preview). Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.

Can I try Gemini 3.1 Pro (preview) and Grok 4.7 for free?

Yes: both are in the free trial of 5 messages.

Which has the bigger context window, Gemini 3.1 Pro (preview) or Grok 4.7?

Gemini 3.1 Pro (preview): 1.05M tokens, against 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

Can I use Gemini 3.1 Pro (preview) and Grok 4.7 in the same chat?

Yes. Pick Gemini 3.1 Pro (preview) for one message and Grok 4.7 for the next; the second sees the whole chat, including the first one's answer.

Which did better in your test runs, Gemini 3.1 Pro (preview) or Grok 4.7?

On the same 50 prompts, run on September 27, 2026, Gemini 3.1 Pro (preview) passed 49 and Grok 4.7 passed 47. The table on this page has each job, and every reply is published.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.