Skip to content

Model vs model

Grok 4.7 vs GLM 5.3 Flash

Grok 4.7 and GLM 5.3 Flash are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.

Model prices and specs checked against OpenRouter's Grok 4.7 page, OpenRouter's GLM 5.3 Flash page. Updated .

Short answer

GLM 5.3 Flash is an everyday model (60 messages a day on Pro); Grok 4.7 draws on the monthly allowance (up to 250 messages a month on Pro). Otherwise, only Grok 4.7 reads a PDF as the whole file. In our test runs, Grok 4.7 passed 47 of the 50 prompts both answered and GLM 5.3 Flash 40; 9 prompts split them, most on summarization (5 to 3).

Grok 4.7 vs GLM 5.3 Flash, prompt by prompt

Every prompt Grok 4.7 and GLM 5.3 Flash both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.

Of the 50 prompts both answered, both passed 39, only Grok 4.7 passed 8, only GLM 5.3 Flash passed 1, and neither passed 2. Grok 4.7 answered sooner on 15 of the 50 and GLM 5.3 Flash on 31; the rest were within 10% of each other. The 50 replies cost $0.3163 on Grok 4.7 and $0.0096 on GLM 5.3 Flash: 32.8× less on GLM 5.3 Flash.

The 9 prompts only one of Grok 4.7 and GLM 5.3 Flash passed

  • Announce a second bakery shop on LinkedIn (writing): GLM 5.3 Flash passed and Grok 4.7 didn't. Grok 4.7: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for. GLM 5.3 Flash: Graded 4.5 of 5 on average (lowest 4).

  • A product announcement with five rules (writing): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Graded 4.0 of 5 on average (lowest 3). GLM 5.3 Flash: Graded 3.3 of 5 on average (lowest 3).

  • Argue both sides of free buses (writing): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). GLM 5.3 Flash: Graded 4.7 of 5 on average (lowest 4); but a paragraph of 92 words, over the 90 allowed.

  • An article in three bullets (summarization): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Graded 4.7 of 5 on average (lowest 4). GLM 5.3 Flash: Graded 4.7 of 5 on average (lowest 4); but 61 words, over the 60 allowed.

  • An email thread in one sentence (summarization): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). GLM 5.3 Flash: Graded 3.7 of 5 on average (lowest 3).

  • Correlation between ad spend and sign-ups (data analysis): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Final answer 0.97: right. GLM 5.3 Flash: Final answer 0.99; expected 0.97.

  • A late order (customer support): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). GLM 5.3 Flash: Graded 3.7 of 5 on average (lowest 3).

  • A refund request outside the window (customer support): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). GLM 5.3 Flash: Graded 4.0 of 5 on average (lowest 2).

  • Book a meeting from a sentence (agents and tool use): Grok 4.7 passed and GLM 5.3 Flash didn't. Grok 4.7: Made the 1 expected call. GLM 5.3 Flash: Expected create_event(attendees: ["priya@northwind.test"], duration_minutes: 30, start: "2026-10-08T15:00", title: {"contains":"Q4"}); got find_contact(name: "Priya Shah"); create_event(title: "Q4 plan review", start: "2026-10-08T15:00", duration_minutes: 30, attendees: ["priya@northwind.test"]).

Job by job, the widest gaps first

  • Summarization: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 3 of 5. GLM 5.3 Flash answered 1.2× sooner at the median, 4.7 s against 4.0 s. GLM 5.3 Flash cost 28.2× less, $0.0161 against $0.0006 for the 5 replies. Their replies ran to about the same length.

  • Customer support: Grok 4.7 passed 4 of 5 and GLM 5.3 Flash 2 of 5. GLM 5.3 Flash answered 1.3× sooner at the median, 7.9 s against 6.2 s. GLM 5.3 Flash cost 24.6× less, $0.0239 against $0.0010 for the 5 replies. GLM 5.3 Flash's replies ran 34% longer, in tokens of reply, thinking not counted.

  • Data analysis: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 4 of 5. Their median waits were close, 9.4 s against 9.8 s. GLM 5.3 Flash cost 22.0× less, $0.0423 against $0.0019 for the 5 replies. GLM 5.3 Flash's replies ran 139% longer, in tokens of reply, thinking not counted.

  • Agents and tool use: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 4 of 5. Grok 4.7 answered 1.3× sooner at the median, 2.5 s against 3.2 s. GLM 5.3 Flash cost 20.0× less, $0.0129 against $0.0006 for the 5 replies. GLM 5.3 Flash's replies ran 12% longer, in tokens of reply, thinking not counted.

  • Writing: Grok 4.7 passed 3 of 5 and GLM 5.3 Flash 2 of 5. Grok 4.7 answered 2.1× sooner at the median, 4.0 s against 8.5 s. GLM 5.3 Flash cost 13.1× less, $0.0210 against $0.0016 for the 5 replies. GLM 5.3 Flash's replies ran 47% longer, in tokens of reply, thinking not counted.

  • Coding: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 3.0× sooner at the median, 25.3 s against 8.4 s. GLM 5.3 Flash cost 86.9× less, $0.1099 against $0.0013 for the 5 replies. Their replies ran to about the same length.

  • Translation: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 3.4× sooner at the median, 13.8 s against 4.1 s. GLM 5.3 Flash cost 70.2× less, $0.0347 against $0.0005 for the 5 replies. GLM 5.3 Flash's replies ran 24% longer, in tokens of reply, thinking not counted.

  • Math: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 3.3× sooner at the median, 9.1 s against 2.7 s. GLM 5.3 Flash cost 33.5× less, $0.0251 against $0.0008 for the 5 replies. Grok 4.7's replies ran 13% longer, in tokens of reply, thinking not counted.

  • SQL: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 3.4× sooner at the median, 4.9 s against 1.4 s. GLM 5.3 Flash cost 29.2× less, $0.0189 against $0.0006 for the 5 replies. Their replies ran to about the same length.

  • RAG and answering from documents: Grok 4.7 passed 5 of 5 and GLM 5.3 Flash 5 of 5. Grok 4.7 answered 1.2× sooner at the median, 1.9 s against 2.2 s. GLM 5.3 Flash cost 15.0× less, $0.0115 against $0.0008 for the 5 replies. GLM 5.3 Flash's replies ran 133% longer, in tokens of reply, thinking not counted.

All 50 prompts: who passed, who answered sooner, who cost less
Grok 4.7 and GLM 5.3 Flash on each prompt of our test runs
PromptResultSoonerCheaper
Turn a title into a URL slugBoth passedGLM 5.3 Flash, 2.6×took 6.7 s and 2.6 sGLM 5.3 Flash, 71.8×cost $0.0048 and $0.0001
Parse a duration like “1h 30m”Both passedGLM 5.3 Flash, 3.1×took 25.3 s and 8.2 sGLM 5.3 Flash, 43.7×cost $0.0151 and $0.0003
Merge overlapping intervalsBoth passedGrok 4.7, 2.5×took 3.8 s and 9.6 sGLM 5.3 Flash, 10.3×cost $0.0025 and $0.0002
Evaluate an arithmetic expression, no evalBoth passedGLM 5.3 Flash, 11.3×took 124.9 s and 11.1 sGLM 5.3 Flash, 134.5×cost $0.0472 and $0.0004
Parse CSV with quoted fieldsBoth passedGLM 5.3 Flash, 12.0×took 100.4 s and 8.4 sGLM 5.3 Flash, 154.6×cost $0.0403 and $0.0003
Announce a second bakery shop on LinkedInOnly GLM 5.3 FlashGrok 4.7, 9.9×took 4.0 s and 39.6 sGLM 5.3 Flash, 2.1×cost $0.0022 and $0.0011
Rewrite corporate jargon in plain wordsNeither passedGLM 5.3 Flash, 3.9×took 19.3 s and 4.9 sGLM 5.3 Flash, 91.1×cost $0.0088 and $0.0001
Decline a meeting and offer two timesBoth passedGrok 4.7, 2.3×took 2.3 s and 5.3 sGLM 5.3 Flash, 12.3×cost $0.0019 and $0.0002
A product announcement with five rulesOnly Grok 4.7Grok 4.7, 3.2×took 2.6 s and 8.5 sGLM 5.3 Flash, 11.4×cost $0.0017 and $0.0001
Argue both sides of free busesOnly Grok 4.7Closetook 12.1 s and 11.2 sGLM 5.3 Flash, 43.8×cost $0.0065 and $0.0001
A discount, then sales taxBoth passedGLM 5.3 Flash, 1.6×took 3.9 s and 2.4 sGLM 5.3 Flash, 31.3×cost $0.0027 and $0.0001
Pens at 3 for $4Both passedGLM 5.3 Flash, 2.8×took 40.7 s and 14.6 sGLM 5.3 Flash, 31.6×cost $0.0087 and $0.0003
Compound interest over three yearsBoth passedGLM 5.3 Flash, 1.1×took 5.5 s and 4.9 sGLM 5.3 Flash, 19.7×cost $0.0031 and $0.0002
Four-digit numbers whose digits sum to 9Both passedGLM 5.3 Flash, 4.1×took 11.1 s and 2.7 sGLM 5.3 Flash, 46.4×cost $0.0056 and $0.0001
The highest of three dice is a 5Both passedGLM 5.3 Flash, 9.6×took 9.1 s and 0.9 sGLM 5.3 Flash, 45.5×cost $0.0050 and $0.0001
An article in three bulletsOnly Grok 4.7GLM 5.3 Flash, 2.5×took 9.8 s and 4.0 sGLM 5.3 Flash, 51.0×cost $0.0051 and $0.0001
An email thread in one sentenceOnly Grok 4.7Grok 4.7, 2.3×took 1.7 s and 4.0 sGLM 5.3 Flash, 12.6×cost $0.0018 and $0.0001
Decisions and action items from a meetingBoth passedGLM 5.3 Flash, 2.5×took 4.7 s and 1.9 sGLM 5.3 Flash, 32.2×cost $0.0028 and $0.0001
A quarterly memo for the CEOBoth passedGrok 4.7, 1.3×took 6.1 s and 8.3 sGLM 5.3 Flash, 29.4×cost $0.0044 and $0.0001
A study with a negative resultBoth passedGrok 4.7, 1.2×took 3.0 s and 3.7 sGLM 5.3 Flash, 22.4×cost $0.0020 and $0.0001
The region with the most revenueBoth passedClosetook 9.4 s and 8.8 sGLM 5.3 Flash, 10.9×cost $0.0057 and $0.0005
Average order value in AugustBoth passedGrok 4.7, 2.6×took 5.5 s and 14.5 sGLM 5.3 Flash, 10.1×cost $0.0037 and $0.0004
Revenue change from July to AugustBoth passedClosetook 10.1 s and 9.8 sGLM 5.3 Flash, 19.1×cost $0.0071 and $0.0004
A median, filtered two waysBoth passedGLM 5.3 Flash, 1.1×took 4.2 s and 3.7 sGLM 5.3 Flash, 29.9×cost $0.0036 and $0.0001
Correlation between ad spend and sign-upsOnly Grok 4.7GLM 5.3 Flash, 2.3×took 45.4 s and 19.4 sGLM 5.3 Flash, 40.9×cost $0.0222 and $0.0005
A late orderOnly Grok 4.7GLM 5.3 Flash, 2.8×took 17.2 s and 6.2 sGLM 5.3 Flash, 29.8×cost $0.0076 and $0.0003
A return inside the windowBoth passedGrok 4.7, 1.3×took 4.7 s and 6.2 sGLM 5.3 Flash, 14.2×cost $0.0032 and $0.0002
A frustrated customerNeither passedGLM 5.3 Flash, 1.4×took 9.4 s and 6.8 sGLM 5.3 Flash, 33.0×cost $0.0050 and $0.0002
A refund request outside the windowOnly Grok 4.7GLM 5.3 Flash, 3.2×took 7.6 s and 2.4 sGLM 5.3 Flash, 25.5×cost $0.0041 and $0.0002
A message with a planted instructionBoth passedGLM 5.3 Flash, 1.5×took 7.9 s and 5.2 sGLM 5.3 Flash, 22.1×cost $0.0040 and $0.0002
A delivery message into SpanishBoth passedGLM 5.3 Flash, 3.5×took 13.8 s and 3.9 sGLM 5.3 Flash, 69.4×cost $0.0066 and $0.0001
A product description into FrenchBoth passedGLM 5.3 Flash, 5.6×took 22.7 s and 4.1 sGLM 5.3 Flash, 107.2×cost $0.0097 and $0.0001
A meeting note into GermanBoth passedGLM 5.3 Flash, 3.0×took 12.6 s and 4.2 sGLM 5.3 Flash, 78.4×cost $0.0060 and $0.0001
Idioms into natural JapaneseBoth passedGLM 5.3 Flash, 2.1×took 10.0 s and 4.8 sGLM 5.3 Flash, 55.1×cost $0.0048 and $0.0001
A lease clause into Brazilian PortugueseBoth passedGLM 5.3 Flash, 11.2×took 15.0 s and 1.3 sGLM 5.3 Flash, 52.5×cost $0.0076 and $0.0001
Customers in one countryBoth passedGLM 5.3 Flash, 1.5×took 1.8 s and 1.2 sGLM 5.3 Flash, 13.9×cost $0.0018 and $0.0001
Count orders by statusBoth passedGrok 4.7, 1.5×took 1.4 s and 2.1 sGLM 5.3 Flash, 12.9×cost $0.0017 and $0.0001
Revenue by categoryBoth passedGLM 5.3 Flash, 1.9×took 4.9 s and 2.6 sGLM 5.3 Flash, 26.9×cost $0.0031 and $0.0001
Every customer, even those without ordersBoth passedGLM 5.3 Flash, 5.4×took 5.3 s and 1.0 sGLM 5.3 Flash, 32.0×cost $0.0037 and $0.0001
Monthly revenue with a running totalBoth passedGLM 5.3 Flash, 11.6×took 16.9 s and 1.4 sGLM 5.3 Flash, 55.0×cost $0.0086 and $0.0002
A fact from one sectionBoth passedGLM 5.3 Flash, 3.6×took 2.7 s and 0.7 sGLM 5.3 Flash, 12.4×cost $0.0025 and $0.0002
Core hours and start timesBoth passedGrok 4.7, 2.1×took 1.9 s and 3.9 sGLM 5.3 Flash, 9.9×cost $0.0022 and $0.0002
Two sections in one answerBoth passedGrok 4.7, 1.3×took 1.5 s and 1.9 sGLM 5.3 Flash, 17.4×cost $0.0018 and $0.0001
A later amendment changes the answerBoth passedGLM 5.3 Flash, 2.5×took 5.5 s and 2.2 sGLM 5.3 Flash, 20.7×cost $0.0030 and $0.0001
A question the handbook doesn't answerBoth passedGrok 4.7, 1.6×took 1.8 s and 2.9 sGLM 5.3 Flash, 21.1×cost $0.0021 and $0.0001
Pick the tool and work out the dateBoth passedGrok 4.7, 2.1×took 1.5 s and 3.2 sGLM 5.3 Flash, 32.0×cost $0.0018 and $0.0001
Convert a currencyBoth passedGrok 4.7, 1.9×took 1.5 s and 3.0 sGLM 5.3 Flash, 9.0×cost $0.0013 and $0.0001
Book a meeting from a sentenceOnly Grok 4.7GLM 5.3 Flash, 1.3×took 9.1 s and 6.8 sGLM 5.3 Flash, 29.2×cost $0.0048 and $0.0002
Search, but don't bookBoth passedClosetook 2.5 s and 2.3 sGLM 5.3 Flash, 25.5×cost $0.0020 and $0.0001
Two calls with a unit conversionBoth passedGLM 5.3 Flash, 1.3×took 4.5 s and 3.4 sGLM 5.3 Flash, 14.9×cost $0.0029 and $0.0002

Grok 4.7 vs GLM 5.3 Flash in our test runs

Grok 4.7 and GLM 5.3 Flash on the same prompts, job by job: how many replies passed their check.

Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Grok 4.7 and GLM 5.3 Flash in our test runs, job by job
JobGrok 4.7GLM 5.3 Flash
Coding5 of 55 of 5
Writing3 of 52 of 5
Math5 of 55 of 5
Summarization5 of 53 of 5
Data analysis5 of 54 of 5
Customer support4 of 52 of 5
Translation5 of 55 of 5
SQL5 of 55 of 5
RAG and answering from documents5 of 55 of 5
Agents and tool use5 of 54 of 5
All jobs47 of 50 · 1 message each · $0.0063 a reply40 of 50 · 1 message each · $0.0002 a reply

The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.

Where they differed most, summarization: both replies to a prompt only one of them passed.

An everyday prompt: “An article in three bullets”, in full.

  • Grok 4.7

    Passed: Graded 4.7 of 5 on average (lowest 4).

    - Aldmoor’s Market Street bike-lane trial ended; the council votes in November. Daily bike trips rose from 900 to 2,300.
    - Car traffic fell 14% and rush-hour delays rose about one minute; cyclist injuries dropped from nine to two.
    - Of 60 shops, 27 saw more customers and 19 fewer. Installation cost €410,000; permanence would add €1.1 million.

    1,758 tokens in, 756 out (666 of them reasoning) · 9.8 s · $0.0051 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·

  • GLM 5.3 Flash

    Failed: Graded 4.7 of 5 on average (lowest 4); but 61 words, over the 60 allowed.

    - Aldmoor's six-month Market Street bike lane trial ended; council votes in November on making it permanent.
    - Cycling rose from ~900 to 2,300 daily trips, injuries dropped from nine to two, car traffic fell 14% with just over a minute added to rush-hour travel times.
    - Businesses are split (27 more customers, 19 fewer, 14 unchanged); lanes cost €410k, permanence needs €1.1m more.

    539 tokens in, 98 out · 4.0 s · $0.0001 · 1 message on Pro · answered by z-ai/glm-5.3-flash via AtlasCloud ·

Grok 4.7 and GLM 5.3 Flash on every plan

Whether the one-time free trial reaches each model, then each paid plan's messages on it.

Messages on Grok 4.7 and GLM 5.3 Flash, plan by plan
PlanPriceGrok 4.7GLM 5.3 Flash
Free$0In the one-time trial of 5 messagesIn the one-time trial of 5 messages
Pro$20 a monthUp to 250 a month60 a day
Max$50 a monthUp to 800 a month120 a day
Ultra$100 a monthUp to 1,800 a month200 a day
Studio$200 a monthUp to 4,000 a month200 a day

Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

What differs

  • Messages on Pro

    GLM 5.3 Flash is an everyday model (60 messages a day on Pro); Grok 4.7 draws on the monthly allowance (up to 250 messages a month on Pro).

  • Context window

    Grok 4.7 takes up to 500K tokens; GLM 5.3 Flash up to 1.05M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

  • Images and PDFs

    Both read images. GLM 5.3 Flash gets a PDF's text rather than the file itself.

  • On Free

    Both are in the free trial.

  • Where messages go

    Grok 4.7: Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. GLM 5.3 Flash: Served through OpenRouter, only by hosts that don't store or train on prompts. The maker's own endpoint is never asked.

Fact by fact

Grok 4.7 and GLM 5.3 Flash, fact by fact
FactGrok 4.7GLM 5.3 Flash
Context window500K tokens1.05M tokens
Reads imagesYesYes
PDFsWhole fileText only
ReasoningYesYes
API price (September 2026)$1.60 in / $4.80 out per million tokens$0.15 in / $0.50 out per million tokens
A typical message at API prices (4,000 tokens in, 700 out)$0.0098$0.0010
A $10 top-up adds200 messagesNothing: an everyday model's count is daily
Where a message goesServed through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts.Served through OpenRouter, only by hosts that don't store or train on prompts. The maker's own endpoint is never asked.
If the provider failsxAI is its only host, so there's no other host to move to: if xAI fails, send the message again or pick another model.When one host is down, OpenRouter moves the request to another host that meets the same rules.
Anthropic's safety fallbackDoesn't applyDoesn't apply
API prices are what our model catalog lists (Grok: xAI's price, served through OpenRouter; GLM: Z.ai's list price). In llmwise you pay per message, not per token: the counts above are what you get.

Grok 4.7 or GLM 5.3 Flash?

From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.

  • Pick Grok 4.7: it reads a PDF as the whole file, charts and scans included.

  • Pick GLM 5.3 Flash: its messages come from the daily count (60 messages a day on Pro), so they leave the monthly allowance for bigger models.

Each model's page, the families, and other pairs

Grok 4.7 vs GLM 5.3 Flash is one pair of models. The page below covers the whole families.

Questions

Is Grok 4.7 or GLM 5.3 Flash cheaper in llmwise?

GLM 5.3 Flash is an everyday model (60 messages a day on Pro); Grok 4.7 draws on the monthly allowance (up to 250 messages a month on Pro). Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.

Can I try Grok 4.7 and GLM 5.3 Flash for free?

Yes: both are in the free trial of 5 messages.

Which has the bigger context window, Grok 4.7 or GLM 5.3 Flash?

GLM 5.3 Flash: 1.05M tokens, against 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.

Can I use Grok 4.7 and GLM 5.3 Flash in the same chat?

Yes. Pick Grok 4.7 for one message and GLM 5.3 Flash for the next; the second sees the whole chat, including the first one's answer.

Which did better in your test runs, Grok 4.7 or GLM 5.3 Flash?

On the same 50 prompts, run on September 27, 2026, Grok 4.7 passed 47 and GLM 5.3 Flash passed 40. The table on this page has each job, and every reply is published.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.