Model vs model
Grok 4.7 vs GLM 5.3
Grok 4.7 and GLM 5.3 are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.
Model prices and specs checked against OpenRouter's Grok 4.7 page, OpenRouter's GLM 5.3 page. Updated .
Short answer
They cost the same in llmwise: up to 250 messages a month on Pro on either. Otherwise, only Grok 4.7 reads images and only Grok 4.7 reads a PDF as the whole file. In our test runs, Grok 4.7 passed 47 of the 50 prompts both answered and GLM 5.3 48; 5 prompts split them, most on writing (3 to 4).
Grok 4.7 vs GLM 5.3, prompt by prompt
Every prompt Grok 4.7 and GLM 5.3 both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.
Of the 50 prompts both answered, both passed 45, only Grok 4.7 passed 2, only GLM 5.3 passed 3, and neither passed 0. Grok 4.7 answered sooner on 0 of the 50 and GLM 5.3 on 50; the rest were within 10% of each other. The 50 replies cost $0.3163 on Grok 4.7 and $0.0364 on GLM 5.3: 8.7× less on GLM 5.3.
The 5 prompts only one of Grok 4.7 and GLM 5.3 passed
Announce a second bakery shop on LinkedIn (writing): GLM 5.3 passed and Grok 4.7 didn't. Grok 4.7: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for. GLM 5.3: Graded 4.3 of 5 on average (lowest 4).
Rewrite corporate jargon in plain words (writing): GLM 5.3 passed and Grok 4.7 didn't. Grok 4.7: Graded 3.7 of 5 on average (lowest 3). GLM 5.3: Graded 4.0 of 5 on average (lowest 3).
Argue both sides of free buses (writing): Grok 4.7 passed and GLM 5.3 didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). GLM 5.3: Graded 3.7 of 5 on average (lowest 3); but a paragraph of 92 words, over the 90 allowed.
An article in three bullets (summarization): Grok 4.7 passed and GLM 5.3 didn't. Grok 4.7: Graded 4.7 of 5 on average (lowest 4). GLM 5.3: Graded 4.3 of 5 on average (lowest 3); but 68 words, over the 60 allowed.
A frustrated customer (customer support): GLM 5.3 passed and Grok 4.7 didn't. Grok 4.7: Graded 3.3 of 5 on average (lowest 3). GLM 5.3: Graded 4.7 of 5 on average (lowest 4).
Job by job, the widest gaps first
Customer support: Grok 4.7 passed 4 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 3.4× sooner at the median, 7.9 s against 2.3 s. GLM 5.3 cost 10.3× less, $0.0239 against $0.0023 for the 5 replies. Their replies ran to about the same length.
Writing: Grok 4.7 passed 3 of 5 and GLM 5.3 4 of 5. GLM 5.3 answered 2.3× sooner at the median, 4.0 s against 1.8 s. GLM 5.3 cost 7.3× less, $0.0210 against $0.0029 for the 5 replies. GLM 5.3's replies ran 41% longer, in tokens of reply, thinking not counted.
Summarization: Grok 4.7 passed 5 of 5 and GLM 5.3 4 of 5. GLM 5.3 answered 4.3× sooner at the median, 4.7 s against 1.1 s. GLM 5.3 cost 5.9× less, $0.0161 against $0.0027 for the 5 replies. Their replies ran to about the same length.
Coding: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 24.0× sooner at the median, 25.3 s against 1.1 s. GLM 5.3 cost 15.0× less, $0.1099 against $0.0073 for the 5 replies. Grok 4.7's replies ran 10% longer, in tokens of reply, thinking not counted.
Math: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 9.6× sooner at the median, 9.1 s against 0.9 s. GLM 5.3 cost 14.2× less, $0.0251 against $0.0018 for the 5 replies. Grok 4.7's replies ran 39% longer, in tokens of reply, thinking not counted.
SQL: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 4.3× sooner at the median, 4.9 s against 1.2 s. GLM 5.3 cost 12.0× less, $0.0189 against $0.0016 for the 5 replies. Their replies ran to about the same length.
Translation: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 8.7× sooner at the median, 13.8 s against 1.6 s. GLM 5.3 cost 9.9× less, $0.0347 against $0.0035 for the 5 replies. GLM 5.3's replies ran 76% longer, in tokens of reply, thinking not counted.
Agents and tool use: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 2.8× sooner at the median, 2.5 s against 0.9 s. GLM 5.3 cost 5.8× less, $0.0129 against $0.0022 for the 5 replies. Their replies ran to about the same length.
RAG and answering from documents: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 2.9× sooner at the median, 1.9 s against 0.6 s. GLM 5.3 cost 4.6× less, $0.0115 against $0.0025 for the 5 replies. GLM 5.3's replies ran 34% longer, in tokens of reply, thinking not counted.
Data analysis: Grok 4.7 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 4.8× sooner at the median, 9.4 s against 1.9 s. GLM 5.3 cost 4.4× less, $0.0423 against $0.0096 for the 5 replies. GLM 5.3's replies ran 24% longer, in tokens of reply, thinking not counted.
All 50 prompts: who passed, who answered sooner, who cost less
| Prompt | Result | Sooner | Cheaper |
|---|---|---|---|
| Turn a title into a URL slug | Both passed | GLM 5.3, 15.4×took 6.7 s and 0.4 s | GLM 5.3, 4.7×cost $0.0048 and $0.0010 |
| Parse a duration like “1h 30m” | Both passed | GLM 5.3, 24.0×took 25.3 s and 1.1 s | GLM 5.3, 10.7×cost $0.0151 and $0.0014 |
| Merge overlapping intervals | Both passed | GLM 5.3, 4.6×took 3.8 s and 0.8 s | GLM 5.3, 3.0×cost $0.0025 and $0.0008 |
| Evaluate an arithmetic expression, no eval | Both passed | GLM 5.3, 10.0×took 124.9 s and 12.5 s | GLM 5.3, 27.8×cost $0.0472 and $0.0017 |
| Parse CSV with quoted fields | Both passed | GLM 5.3, 38.0×took 100.4 s and 2.6 s | GLM 5.3, 16.8×cost $0.0403 and $0.0024 |
| Announce a second bakery shop on LinkedIn | Only GLM 5.3 | GLM 5.3, 1.4×took 4.0 s and 2.9 s | GLM 5.3, 5.9×cost $0.0022 and $0.0004 |
| Rewrite corporate jargon in plain words | Only GLM 5.3 | GLM 5.3, 10.9×took 19.3 s and 1.8 s | GLM 5.3, 43.6×cost $0.0088 and $0.0002 |
| Decline a meeting and offer two times | Both passed | GLM 5.3, 1.6×took 2.3 s and 1.4 s | GLM 5.3, 7.0×cost $0.0019 and $0.0003 |
| A product announcement with five rules | Both passed | GLM 5.3, 3.1×took 2.6 s and 0.8 s | GLM 5.3, 2.2×cost $0.0017 and $0.0008 |
| Argue both sides of free buses | Only Grok 4.7 | GLM 5.3, 6.8×took 12.1 s and 1.8 s | GLM 5.3, 5.1×cost $0.0065 and $0.0013 |
| A discount, then sales tax | Both passed | GLM 5.3, 4.1×took 3.9 s and 0.9 s | GLM 5.3, 13.5×cost $0.0027 and $0.0002 |
| Pens at 3 for $4 | Both passed | GLM 5.3, 27.5×took 40.7 s and 1.5 s | GLM 5.3, 48.8×cost $0.0087 and $0.0002 |
| Compound interest over three years | Both passed | GLM 5.3, 9.5×took 5.5 s and 0.6 s | GLM 5.3, 5.7×cost $0.0031 and $0.0005 |
| Four-digit numbers whose digits sum to 9 | Both passed | GLM 5.3, 6.0×took 11.1 s and 1.9 s | GLM 5.3, 22.3×cost $0.0056 and $0.0003 |
| The highest of three dice is a 5 | Both passed | GLM 5.3, 10.5×took 9.1 s and 0.9 s | GLM 5.3, 8.5×cost $0.0050 and $0.0006 |
| An article in three bullets | Only Grok 4.7 | GLM 5.3, 3.6×took 9.8 s and 2.8 s | GLM 5.3, 11.4×cost $0.0051 and $0.0004 |
| An email thread in one sentence | Both passed | GLM 5.3, 1.7×took 1.7 s and 1.0 s | GLM 5.3, 7.7×cost $0.0018 and $0.0002 |
| Decisions and action items from a meeting | Both passed | GLM 5.3, 3.1×took 4.7 s and 1.5 s | GLM 5.3, 12.0×cost $0.0028 and $0.0002 |
| A quarterly memo for the CEO | Both passed | GLM 5.3, 5.6×took 6.1 s and 1.1 s | GLM 5.3, 4.4×cost $0.0044 and $0.0010 |
| A study with a negative result | Both passed | GLM 5.3, 3.6×took 3.0 s and 0.8 s | GLM 5.3, 2.5×cost $0.0020 and $0.0008 |
| The region with the most revenue | Both passed | GLM 5.3, 4.5×took 9.4 s and 2.1 s | GLM 5.3, 14.5×cost $0.0057 and $0.0004 |
| Average order value in August | Both passed | GLM 5.3, 2.8×took 5.5 s and 1.9 s | GLM 5.3, 10.1×cost $0.0037 and $0.0004 |
| Revenue change from July to August | Both passed | GLM 5.3, 7.3×took 10.1 s and 1.4 s | GLM 5.3, 4.5×cost $0.0071 and $0.0016 |
| A median, filtered two ways | Both passed | GLM 5.3, 4.4×took 4.2 s and 1.0 s | GLM 5.3, 3.4×cost $0.0036 and $0.0011 |
| Correlation between ad spend and sign-ups | Both passed | GLM 5.3, 8.1×took 45.4 s and 5.6 s | GLM 5.3, 3.6×cost $0.0222 and $0.0062 |
| A late order | Both passed | GLM 5.3, 6.6×took 17.2 s and 2.6 s | GLM 5.3, 19.3×cost $0.0076 and $0.0004 |
| A return inside the window | Both passed | GLM 5.3, 3.0×took 4.7 s and 1.6 s | GLM 5.3, 13.7×cost $0.0032 and $0.0002 |
| A frustrated customer | Only GLM 5.3 | GLM 5.3, 4.1×took 9.4 s and 2.3 s | GLM 5.3, 16.2×cost $0.0050 and $0.0003 |
| A refund request outside the window | Both passed | GLM 5.3, 6.7×took 7.6 s and 1.1 s | GLM 5.3, 4.5×cost $0.0041 and $0.0009 |
| A message with a planted instruction | Both passed | GLM 5.3, 1.5×took 7.9 s and 5.3 s | GLM 5.3, 8.3×cost $0.0040 and $0.0005 |
| A delivery message into Spanish | Both passed | GLM 5.3, 9.3×took 13.8 s and 1.5 s | GLM 5.3, 28.7×cost $0.0066 and $0.0002 |
| A product description into French | Both passed | GLM 5.3, 15.4×took 22.7 s and 1.5 s | GLM 5.3, 32.7×cost $0.0097 and $0.0003 |
| A meeting note into German | Both passed | GLM 5.3, 8.0×took 12.6 s and 1.6 s | GLM 5.3, 32.4×cost $0.0060 and $0.0002 |
| Idioms into natural Japanese | Both passed | GLM 5.3, 5.8×took 10.0 s and 1.7 s | GLM 5.3, 4.1×cost $0.0048 and $0.0012 |
| A lease clause into Brazilian Portuguese | Both passed | GLM 5.3, 8.8×took 15.0 s and 1.7 s | GLM 5.3, 4.7×cost $0.0076 and $0.0016 |
| Customers in one country | Both passed | GLM 5.3, 2.3×took 1.8 s and 0.8 s | GLM 5.3, 13.7×cost $0.0018 and $0.0001 |
| Count orders by status | Both passed | GLM 5.3, 1.2×took 1.4 s and 1.2 s | GLM 5.3, 12.9×cost $0.0017 and $0.0001 |
| Revenue by category | Both passed | GLM 5.3, 4.0×took 4.9 s and 1.2 s | GLM 5.3, 14.8×cost $0.0031 and $0.0002 |
| Every customer, even those without orders | Both passed | GLM 5.3, 5.5×took 5.3 s and 1.0 s | GLM 5.3, 5.1×cost $0.0037 and $0.0007 |
| Monthly revenue with a running total | Both passed | GLM 5.3, 5.2×took 16.9 s and 3.2 s | GLM 5.3, 22.3×cost $0.0086 and $0.0004 |
| A fact from one section | Both passed | GLM 5.3, 3.0×took 2.7 s and 0.9 s | GLM 5.3, 9.3×cost $0.0025 and $0.0003 |
| Core hours and start times | Both passed | GLM 5.3, 3.7×took 1.9 s and 0.5 s | GLM 5.3, 3.1×cost $0.0022 and $0.0007 |
| Two sections in one answer | Both passed | GLM 5.3, 2.9×took 1.5 s and 0.5 s | GLM 5.3, 3.9×cost $0.0018 and $0.0005 |
| A later amendment changes the answer | Both passed | GLM 5.3, 4.4×took 5.5 s and 1.3 s | GLM 5.3, 5.0×cost $0.0030 and $0.0006 |
| A question the handbook doesn't answer | Both passed | GLM 5.3, 2.8×took 1.8 s and 0.6 s | GLM 5.3, 4.5×cost $0.0021 and $0.0005 |
| Pick the tool and work out the date | Both passed | GLM 5.3, 1.7×took 1.5 s and 0.9 s | GLM 5.3, 8.3×cost $0.0018 and $0.0002 |
| Convert a currency | Both passed | GLM 5.3, 1.4×took 1.5 s and 1.1 s | GLM 5.3, 9.7×cost $0.0013 and $0.0001 |
| Book a meeting from a sentence | Both passed | GLM 5.3, 5.3×took 9.1 s and 1.7 s | GLM 5.3, 8.6×cost $0.0048 and $0.0006 |
| Search, but don't book | Both passed | GLM 5.3, 4.6×took 2.5 s and 0.5 s | GLM 5.3, 3.7×cost $0.0020 and $0.0005 |
| Two calls with a unit conversion | Both passed | GLM 5.3, 6.4×took 4.5 s and 0.7 s | GLM 5.3, 3.9×cost $0.0029 and $0.0008 |
Grok 4.7 vs GLM 5.3 in our test runs
Grok 4.7 and GLM 5.3 on the same prompts, job by job: how many replies passed their check.
Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
| Job | Grok 4.7 | GLM 5.3 |
|---|---|---|
| Coding | 5 of 5 | 5 of 5 |
| Writing | 3 of 5 | 4 of 5 |
| Math | 5 of 5 | 5 of 5 |
| Summarization | 5 of 5 | 4 of 5 |
| Data analysis | 5 of 5 | 5 of 5 |
| Customer support | 4 of 5 | 5 of 5 |
| Translation | 5 of 5 | 5 of 5 |
| SQL | 5 of 5 | 5 of 5 |
| RAG and answering from documents | 5 of 5 | 5 of 5 |
| Agents and tool use | 5 of 5 | 5 of 5 |
| All jobs | 47 of 50 · 1 message each · $0.0063 a reply | 48 of 50 · 1 message each · $0.0007 a reply |
The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.
Where they differed most, writing: both replies to a prompt only one of them passed.
An everyday prompt: “Announce a second bakery shop on LinkedIn”, in full.
Grok 4.7
Failed: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for.
We're opening a second Crumb & Co. shop in Leeds on 14 November. The new bakery will offer the same bread, cakes and pastries you already know from our first shop, baked fresh each morning. If you've been asking for a Leeds location, this one is for you. On opening day, coffee is free with any purchase. Come in, say hello, and try something from the counter. We'd love to meet you and hear what you'd like to see on the shelves. Thank you for supporting a small bakery. We can't wait to welcome you.
1,624 tokens in, 207 out (93 of them reasoning) · 4.0 s · $0.0022 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·
GLM 5.3
Passed: Graded 4.3 of 5 on average (lowest 4).
Some news we've been itching to share: Crumb & Co. is opening a second shop! On 14 November, we'll be firing up the ovens in Leeds, bringing the same breads, pastries and cakes you've come to know us for to a brand-new corner of the world. It's a big step for a small bakery, and honestly, we couldn't be more excited (and a little nervous, in the best way). The new shop will have everything you love from our original spot — slow-fermented sourdough, buttery croissants, and weekend treats that disappear fast — plus a few surprises we're still perfecting. To celebrate opening day, every customer gets a free coffee with their purchase. So come along, say hello, and grab a cup on us. Leeds, we can't wait to meet you. See you on 14 November — come hungry.
402 tokens in, 188 out · 2.9 s · $0.0004 · 1 message on Pro · answered by z-ai/glm-5.3 via Baidu ·
Grok 4.7 and GLM 5.3 on every plan
Whether the one-time free trial reaches each model, then each paid plan's messages on it.
| Plan | Price | Grok 4.7 | GLM 5.3 |
|---|---|---|---|
| Free | $0 | In the one-time trial of 5 messages | In the one-time trial of 5 messages |
| Pro | $20 a month | Up to 250 a month | Up to 250 a month |
| Max | $50 a month | Up to 800 a month | Up to 800 a month |
| Ultra | $100 a month | Up to 1,800 a month | Up to 1,800 a month |
| Studio | $200 a month | Up to 4,000 a month | Up to 4,000 a month |
Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
What differs
Messages on Pro
They cost the same in llmwise: up to 250 messages a month on Pro on either.
Context window
Grok 4.7 takes up to 500K tokens; GLM 5.3 up to 1.05M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Images and PDFs
GLM 5.3 doesn't read images. GLM 5.3 gets a PDF's text rather than the file itself.
On Free
Both are in the free trial.
Where messages go
Grok 4.7: Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. GLM 5.3: Served through OpenRouter, only by hosts that don't store or train on prompts. The maker's own endpoint is never asked.
Fact by fact
| Fact | Grok 4.7 | GLM 5.3 |
|---|---|---|
| Context window | 500K tokens | 1.05M tokens |
| Reads images | Yes | No |
| PDFs | Whole file | Text only |
| Reasoning | Yes | Yes |
| API price (September 2026) | $1.60 in / $4.80 out per million tokens | $1.40 in / $4.40 out per million tokens |
| A typical message at API prices (4,000 tokens in, 700 out) | $0.0098 | $0.0087 |
| A $10 top-up adds | 200 messages | 200 messages |
| Where a message goes | Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. | Served through OpenRouter, only by hosts that don't store or train on prompts. The maker's own endpoint is never asked. |
| If the provider fails | xAI is its only host, so there's no other host to move to: if xAI fails, send the message again or pick another model. | When one host is down, OpenRouter moves the request to another host that meets the same rules. |
| Anthropic's safety fallback | Doesn't apply | Doesn't apply |
Grok 4.7 or GLM 5.3?
From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.
Pick Grok 4.7: it reads images; it reads a PDF as the whole file, charts and scans included.
Pick GLM 5.3: it costs its maker less to run ($0.0087 a typical message at API prices), though in llmwise the count is the same.
Each model's page, the families, and other pairs
Grok 4.7 vs GLM 5.3 is one pair of models. The page below covers the whole families.
- Grok 4.7: price, limits and messages on every plan
- GLM 5.3: price, limits and messages on every plan
- Grok vs GLM
- Claude Fable 5.1 vs Grok 4.7
- GPT-6 Astra vs Grok 4.7
- Claude Sonnet 5 vs Grok 4.7
- GLM 5.3 vs GLM 5.3 Flash
- Kimi K3 vs GLM 5.3
- DeepSeek V4 Pro vs GLM 5.3
- Every model-vs-model page
Grok 4.7 is a Grok model; GLM 5.3 is a GLM model.
Questions
Is Grok 4.7 or GLM 5.3 cheaper in llmwise?
They cost the same in llmwise: up to 250 messages a month on Pro on either. Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.
Can I try Grok 4.7 and GLM 5.3 for free?
Yes: both are in the free trial of 5 messages.
Which has the bigger context window, Grok 4.7 or GLM 5.3?
GLM 5.3: 1.05M tokens, against 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Can I use Grok 4.7 and GLM 5.3 in the same chat?
Yes. Pick Grok 4.7 for one message and GLM 5.3 for the next; the second sees the whole chat, including the first one's answer.
Which did better in your test runs, Grok 4.7 or GLM 5.3?
On the same 50 prompts, run on September 27, 2026, Grok 4.7 passed 47 and GLM 5.3 passed 48. The table on this page has each job, and every reply is published.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.