Model vs model
Grok 4.7 vs Kimi K3
Grok 4.7 and Kimi K3 are both in llmwise. What each gets on every plan, what it reads, how it's served and what happens when its provider fails, from the catalog and the code that runs them.
Model prices and specs checked against OpenRouter's Grok 4.7 page, OpenRouter's Kimi K3 page. Updated .
Short answer
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Kimi K3. Otherwise, only Grok 4.7 reads a PDF as the whole file. In our test runs, Grok 4.7 passed 47 of the 50 prompts both answered and Kimi K3 46; 7 prompts split them, most on summarization (5 to 3).
Grok 4.7 vs Kimi K3, prompt by prompt
Every prompt Grok 4.7 and Kimi K3 both answered, compared directly, their biggest differences first. One run each, through OpenRouter: a wait depends on the provider and the load that day, so a lead under 10% counts as close.
Of the 50 prompts both answered, both passed 43, only Grok 4.7 passed 4, only Kimi K3 passed 3, and neither passed 0. Grok 4.7 answered sooner on 9 of the 50 and Kimi K3 on 37; the rest were within 10% of each other. The 50 replies cost $0.3163 on Grok 4.7 and $0.1938 on Kimi K3: 1.6× less on Kimi K3.
The 7 prompts only one of Grok 4.7 and Kimi K3 passed
Announce a second bakery shop on LinkedIn (writing): Kimi K3 passed and Grok 4.7 didn't. Grok 4.7: Graded 4.5 of 5 on average (lowest 4); but 94 words, under the 100 asked for. Kimi K3: Graded 4.5 of 5 on average (lowest 4).
Rewrite corporate jargon in plain words (writing): Kimi K3 passed and Grok 4.7 didn't. Grok 4.7: Graded 3.7 of 5 on average (lowest 3). Kimi K3: Graded 4.3 of 5 on average (lowest 4).
A product announcement with five rules (writing): Grok 4.7 passed and Kimi K3 didn't. Grok 4.7: Graded 4.0 of 5 on average (lowest 3). Kimi K3: Graded 4.3 of 5 on average (lowest 4); but doesn't end with a question.
An article in three bullets (summarization): Grok 4.7 passed and Kimi K3 didn't. Grok 4.7: Graded 4.7 of 5 on average (lowest 4). Kimi K3: Graded 4.0 of 5 on average (lowest 3); but 62 words, over the 60 allowed.
An email thread in one sentence (summarization): Grok 4.7 passed and Kimi K3 didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). Kimi K3: Graded 5.0 of 5 on average (lowest 5); but 32 words, over the 30 allowed.
A frustrated customer (customer support): Kimi K3 passed and Grok 4.7 didn't. Grok 4.7: Graded 3.3 of 5 on average (lowest 3). Kimi K3: Graded 4.0 of 5 on average (lowest 3).
A refund request outside the window (customer support): Grok 4.7 passed and Kimi K3 didn't. Grok 4.7: Graded 4.3 of 5 on average (lowest 4). Kimi K3: Graded 4.0 of 5 on average (lowest 2).
Job by job, the widest gaps first
Summarization: Grok 4.7 passed 5 of 5 and Kimi K3 3 of 5. Kimi K3 answered 1.4× sooner at the median, 4.7 s against 3.4 s. They cost about the same, $0.0161 against $0.0172 for the 5 replies. Kimi K3's replies ran 23% longer, in tokens of reply, thinking not counted.
Writing: Grok 4.7 passed 3 of 5 and Kimi K3 4 of 5. Kimi K3 answered 1.3× sooner at the median, 4.0 s against 3.0 s. They cost about the same, $0.0210 against $0.0192 for the 5 replies. Kimi K3's replies ran 54% longer, in tokens of reply, thinking not counted.
Coding: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Kimi K3 answered 6.1× sooner at the median, 25.3 s against 4.1 s. Kimi K3 cost 3.5× less, $0.1099 against $0.0316 for the 5 replies. Their replies ran to about the same length.
Translation: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Kimi K3 answered 2.6× sooner at the median, 13.8 s against 5.2 s. Kimi K3 cost 1.9× less, $0.0347 against $0.0179 for the 5 replies. Kimi K3's replies ran 135% longer, in tokens of reply, thinking not counted.
SQL: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Kimi K3 answered 3.1× sooner at the median, 4.9 s against 1.6 s. Kimi K3 cost 1.6× less, $0.0189 against $0.0118 for the 5 replies. Kimi K3's replies ran 14% longer, in tokens of reply, thinking not counted.
RAG and answering from documents: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Grok 4.7 answered 1.4× sooner at the median, 1.9 s against 2.6 s. Kimi K3 cost 1.6× less, $0.0115 against $0.0073 for the 5 replies. Kimi K3's replies ran 97% longer, in tokens of reply, thinking not counted.
Math: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Kimi K3 answered 1.7× sooner at the median, 9.1 s against 5.4 s. Kimi K3 cost 1.3× less, $0.0251 against $0.0187 for the 5 replies. Grok 4.7's replies ran 37% longer, in tokens of reply, thinking not counted.
Data analysis: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Kimi K3 answered 2.7× sooner at the median, 9.4 s against 3.5 s. Kimi K3 cost 1.2× less, $0.0423 against $0.0339 for the 5 replies. Kimi K3's replies ran 43% longer, in tokens of reply, thinking not counted.
Customer support: Grok 4.7 passed 4 of 5 and Kimi K3 4 of 5. Kimi K3 answered 1.8× sooner at the median, 7.9 s against 4.3 s. They cost about the same, $0.0239 against $0.0228 for the 5 replies. Kimi K3's replies ran 24% longer, in tokens of reply, thinking not counted.
Agents and tool use: Grok 4.7 passed 5 of 5 and Kimi K3 5 of 5. Kimi K3 answered 1.2× sooner at the median, 2.5 s against 2.0 s. They cost about the same, $0.0129 against $0.0135 for the 5 replies. Kimi K3's replies ran 25% longer, in tokens of reply, thinking not counted.
All 50 prompts: who passed, who answered sooner, who cost less
| Prompt | Result | Sooner | Cheaper |
|---|---|---|---|
| Turn a title into a URL slug | Both passed | Kimi K3, 2.6×took 6.7 s and 2.6 s | Kimi K3, 1.4×cost $0.0048 and $0.0035 |
| Parse a duration like “1h 30m” | Both passed | Kimi K3, 3.6×took 25.3 s and 7.1 s | Kimi K3, 2.4×cost $0.0151 and $0.0063 |
| Merge overlapping intervals | Both passed | Kimi K3, 1.9×took 3.8 s and 2.0 s | Grok 4.7, 1.5×cost $0.0025 and $0.0036 |
| Evaluate an arithmetic expression, no eval | Both passed | Kimi K3, 12.6×took 124.9 s and 9.9 s | Kimi K3, 4.2×cost $0.0472 and $0.0113 |
| Parse CSV with quoted fields | Both passed | Kimi K3, 24.3×took 100.4 s and 4.1 s | Kimi K3, 5.8×cost $0.0403 and $0.0069 |
| Announce a second bakery shop on LinkedIn | Only Kimi K3 | Closetook 4.0 s and 3.9 s | Grok 4.7, 2.1×cost $0.0022 and $0.0046 |
| Rewrite corporate jargon in plain words | Only Kimi K3 | Kimi K3, 9.7×took 19.3 s and 2.0 s | Kimi K3, 2.9×cost $0.0088 and $0.0030 |
| Decline a meeting and offer two times | Both passed | Grok 4.7, 1.3×took 2.3 s and 3.0 s | Grok 4.7, 1.8×cost $0.0019 and $0.0034 |
| A product announcement with five rules | Only Grok 4.7 | Kimi K3, 1.1×took 2.6 s and 2.3 s | Grok 4.7, 2.0×cost $0.0017 and $0.0033 |
| Argue both sides of free buses | Both passed | Kimi K3, 3.4×took 12.1 s and 3.6 s | Kimi K3, 1.3×cost $0.0065 and $0.0049 |
| A discount, then sales tax | Both passed | Grok 4.7, 1.3×took 3.9 s and 5.0 s | Grok 4.7, 1.4×cost $0.0027 and $0.0037 |
| Pens at 3 for $4 | Both passed | Kimi K3, 7.6×took 40.7 s and 5.4 s | Kimi K3, 1.9×cost $0.0087 and $0.0046 |
| Compound interest over three years | Both passed | Closetook 5.5 s and 5.9 s | Grok 4.7, 1.4×cost $0.0031 and $0.0044 |
| Four-digit numbers whose digits sum to 9 | Both passed | Kimi K3, 1.9×took 11.1 s and 6.0 s | Kimi K3, 1.7×cost $0.0056 and $0.0032 |
| The highest of three dice is a 5 | Both passed | Kimi K3, 3.8×took 9.1 s and 2.4 s | Kimi K3, 1.8×cost $0.0050 and $0.0027 |
| An article in three bullets | Only Grok 4.7 | Kimi K3, 3.5×took 9.8 s and 2.8 s | Kimi K3, 1.8×cost $0.0051 and $0.0028 |
| An email thread in one sentence | Only Grok 4.7 | Grok 4.7, 3.7×took 1.7 s and 6.4 s | Grok 4.7, 1.6×cost $0.0018 and $0.0030 |
| Decisions and action items from a meeting | Both passed | Kimi K3, 1.4×took 4.7 s and 3.4 s | Grok 4.7, 1.5×cost $0.0028 and $0.0044 |
| A quarterly memo for the CEO | Both passed | Grok 4.7, 2.5×took 6.1 s and 15.7 s | Grok 4.7, 1.1×cost $0.0044 and $0.0049 |
| A study with a negative result | Both passed | Closetook 3.0 s and 3.2 s | Closecost $0.0020 and $0.0021 |
| The region with the most revenue | Both passed | Kimi K3, 4.2×took 9.4 s and 2.2 s | Kimi K3, 1.2×cost $0.0057 and $0.0049 |
| Average order value in August | Both passed | Kimi K3, 2.1×took 5.5 s and 2.7 s | Grok 4.7, 1.4×cost $0.0037 and $0.0054 |
| Revenue change from July to August | Both passed | Kimi K3, 1.6×took 10.1 s and 6.4 s | Closecost $0.0071 and $0.0064 |
| A median, filtered two ways | Both passed | Kimi K3, 1.2×took 4.2 s and 3.5 s | Grok 4.7, 1.2×cost $0.0036 and $0.0042 |
| Correlation between ad spend and sign-ups | Both passed | Kimi K3, 2.2×took 45.4 s and 20.2 s | Kimi K3, 1.7×cost $0.0222 and $0.0131 |
| A late order | Both passed | Kimi K3, 6.9×took 17.2 s and 2.5 s | Kimi K3, 1.7×cost $0.0076 and $0.0045 |
| A return inside the window | Both passed | Kimi K3, 1.1×took 4.7 s and 4.3 s | Closecost $0.0032 and $0.0033 |
| A frustrated customer | Only Kimi K3 | Kimi K3, 1.7×took 9.4 s and 5.6 s | Grok 4.7, 1.2×cost $0.0050 and $0.0063 |
| A refund request outside the window | Only Grok 4.7 | Kimi K3, 3.0×took 7.6 s and 2.6 s | Closecost $0.0041 and $0.0045 |
| A message with a planted instruction | Both passed | Grok 4.7, 1.5×took 7.9 s and 11.5 s | Closecost $0.0040 and $0.0041 |
| A delivery message into Spanish | Both passed | Kimi K3, 4.7×took 13.8 s and 2.9 s | Kimi K3, 2.1×cost $0.0066 and $0.0031 |
| A product description into French | Both passed | Kimi K3, 4.2×took 22.7 s and 5.4 s | Kimi K3, 2.6×cost $0.0097 and $0.0037 |
| A meeting note into German | Both passed | Kimi K3, 4.2×took 12.6 s and 3.0 s | Kimi K3, 4.3×cost $0.0060 and $0.0014 |
| Idioms into natural Japanese | Both passed | Kimi K3, 1.9×took 10.0 s and 5.2 s | Grok 4.7, 1.1×cost $0.0048 and $0.0053 |
| A lease clause into Brazilian Portuguese | Both passed | Kimi K3, 1.5×took 15.0 s and 10.1 s | Kimi K3, 1.7×cost $0.0076 and $0.0044 |
| Customers in one country | Both passed | Kimi K3, 1.6×took 1.8 s and 1.1 s | Closecost $0.0018 and $0.0018 |
| Count orders by status | Both passed | Closetook 1.4 s and 1.5 s | Kimi K3, 2.0×cost $0.0017 and $0.0009 |
| Revenue by category | Both passed | Kimi K3, 3.1×took 4.9 s and 1.6 s | Kimi K3, 1.6×cost $0.0031 and $0.0019 |
| Every customer, even those without orders | Both passed | Kimi K3, 2.1×took 5.3 s and 2.5 s | Kimi K3, 1.2×cost $0.0037 and $0.0031 |
| Monthly revenue with a running total | Both passed | Kimi K3, 6.8×took 16.9 s and 2.5 s | Kimi K3, 2.0×cost $0.0086 and $0.0042 |
| A fact from one section | Both passed | Grok 4.7, 2.4×took 2.7 s and 6.5 s | Kimi K3, 2.4×cost $0.0025 and $0.0010 |
| Core hours and start times | Both passed | Kimi K3, 1.5×took 1.9 s and 1.2 s | Kimi K3, 2.0×cost $0.0022 and $0.0011 |
| Two sections in one answer | Both passed | Grok 4.7, 1.8×took 1.5 s and 2.6 s | Grok 4.7, 1.1×cost $0.0018 and $0.0020 |
| A later amendment changes the answer | Both passed | Kimi K3, 1.2×took 5.5 s and 4.8 s | Kimi K3, 1.3×cost $0.0030 and $0.0022 |
| A question the handbook doesn't answer | Both passed | Kimi K3, 1.9×took 1.8 s and 0.9 s | Kimi K3, 2.1×cost $0.0021 and $0.0010 |
| Pick the tool and work out the date | Both passed | Grok 4.7, 2.4×took 1.5 s and 3.7 s | Grok 4.7, 1.8×cost $0.0018 and $0.0032 |
| Convert a currency | Both passed | Grok 4.7, 1.9×took 1.5 s and 2.9 s | Grok 4.7, 1.9×cost $0.0013 and $0.0025 |
| Book a meeting from a sentence | Both passed | Kimi K3, 5.8×took 9.1 s and 1.6 s | Kimi K3, 1.5×cost $0.0048 and $0.0031 |
| Search, but don't book | Both passed | Kimi K3, 1.2×took 2.5 s and 2.0 s | Grok 4.7, 1.7×cost $0.0020 and $0.0034 |
| Two calls with a unit conversion | Both passed | Kimi K3, 2.6×took 4.5 s and 1.7 s | Kimi K3, 2.4×cost $0.0029 and $0.0012 |
Grok 4.7 vs Kimi K3 in our test runs
Grok 4.7 and Kimi K3 on the same prompts, job by job: how many replies passed their check.
Based on 100 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
| Job | Grok 4.7 | Kimi K3 |
|---|---|---|
| Coding | 5 of 5 | 5 of 5 |
| Writing | 3 of 5 | 4 of 5 |
| Math | 5 of 5 | 5 of 5 |
| Summarization | 5 of 5 | 3 of 5 |
| Data analysis | 5 of 5 | 5 of 5 |
| Customer support | 4 of 5 | 4 of 5 |
| Translation | 5 of 5 | 5 of 5 |
| SQL | 5 of 5 | 5 of 5 |
| RAG and answering from documents | 5 of 5 | 5 of 5 |
| Agents and tool use | 5 of 5 | 5 of 5 |
| All jobs | 47 of 50 · 1 message each · $0.0063 a reply | 46 of 50 · 1 message each · $0.0039 a reply |
The same prompts, sent to both through llmwise's own pipeline and checked the same way: five for each job. A bold count is the one that passed more. Messages: what each reply counted as on Pro. Cost: what OpenRouter charged us per reply, on average. Every prompt, and how it's scored.
Where they differed most, summarization: both replies to a prompt only one of them passed.
An everyday prompt: “An article in three bullets”, in full.
Grok 4.7
Passed: Graded 4.7 of 5 on average (lowest 4).
- Aldmoor’s Market Street bike-lane trial ended; the council votes in November. Daily bike trips rose from 900 to 2,300. - Car traffic fell 14% and rush-hour delays rose about one minute; cyclist injuries dropped from nine to two. - Of 60 shops, 27 saw more customers and 19 fewer. Installation cost €410,000; permanence would add €1.1 million.
1,758 tokens in, 756 out (666 of them reasoning) · 9.8 s · $0.0051 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·
Kimi K3
Failed: Graded 4.0 of 5 on average (lowest 3); but 62 words, over the 60 allowed.
- Aldmoor's six-month protected bike lane trial on Market Street ended; the council votes in November on making them permanent. - Bike trips rose from 900 to 2,300 daily, cyclist injuries dropped from nine to two, while car traffic fell 14% with only slightly longer rush-hour drive times. - Businesses were split on customer impact; installation cost €410,000, with permanent lanes requiring €1.1 million more.
608 tokens in, 109 out (10 of them reasoning) · 2.8 s · $0.0028 · 1 message on Pro · answered by moonshotai/kimi-k3 via Phala ·
Grok 4.7 and Kimi K3 on every plan
Whether the one-time free trial reaches each model, then each paid plan's messages on it.
| Plan | Price | Grok 4.7 | Kimi K3 |
|---|---|---|---|
| Free | $0 | In the one-time trial of 5 messages | In the one-time trial of 5 messages |
| Pro | $20 a month | Up to 250 a month | Up to 125 a month |
| Max | $50 a month | Up to 800 a month | Up to 400 a month |
| Ultra | $100 a month | Up to 1,800 a month | Up to 900 a month |
| Studio | $200 a month | Up to 4,000 a month | Up to 2,000 a month |
Prices don't include tax, which is added where it applies and shown before you pay. A paid plan's month is one allowance shared by every model, so each monthly count is the most you get if all of it goes to that model. It renews each billing period; everyday models refill daily at 00:00 UTC. Long chats count more per reply. How pricing works.
Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.
What differs
Messages on Pro
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Kimi K3.
Context window
Grok 4.7 takes up to 500K tokens; Kimi K3 up to 1.05M tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Images and PDFs
Both read images. Kimi K3 gets a PDF's text rather than the file itself.
On Free
Both are in the free trial.
Where messages go
Grok 4.7: Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. Kimi K3: Served through OpenRouter, only by hosts that don't store or train on prompts. The maker's own endpoint is never asked.
Fact by fact
| Fact | Grok 4.7 | Kimi K3 |
|---|---|---|
| Context window | 500K tokens | 1.05M tokens |
| Reads images | Yes | Yes |
| PDFs | Whole file | Text only |
| Reasoning | Yes | Yes |
| API price (September 2026) | $1.60 in / $4.80 out per million tokens | $3.00 in / $15.00 out per million tokens |
| A typical message at API prices (4,000 tokens in, 700 out) | $0.0098 | $0.0225 |
| A $10 top-up adds | 200 messages | 100 messages |
| Where a message goes | Served through OpenRouter by xAI alone, on an endpoint that doesn't store or train on prompts. | Served through OpenRouter, only by hosts that don't store or train on prompts. The maker's own endpoint is never asked. |
| If the provider fails | xAI is its only host, so there's no other host to move to: if xAI fails, send the message again or pick another model. | When one host is down, OpenRouter moves the request to another host that meets the same rules. |
| Anthropic's safety fallback | Doesn't apply | Doesn't apply |
Grok 4.7 or Kimi K3?
From the facts above and our test runs: the rest is how their answers suit your work, which one chat can show you.
Pick Grok 4.7: more messages for the money, up to 250 messages a month on Pro; it reads a PDF as the whole file, charts and scans included.
Each model's page, the families, and other pairs
Grok 4.7 vs Kimi K3 is one pair of models. The page below covers the whole families.
- Grok 4.7: price, limits and messages on every plan
- Kimi K3: price, limits and messages on every plan
- Grok vs Kimi
- Claude Fable 5.1 vs Grok 4.7
- GPT-6 Astra vs Grok 4.7
- Claude Sonnet 5 vs Grok 4.7
- Claude Fable 5.1 vs Kimi K3
- Kimi K3 vs GLM 5.3
- DeepSeek V4 Pro vs Kimi K3
- Every model-vs-model page
Grok 4.7 is a Grok model; Kimi K3 is a Kimi model.
Questions
Is Grok 4.7 or Kimi K3 cheaper in llmwise?
Grok 4.7 gets twice as many messages: up to 250 messages a month on Pro, against up to 125 messages a month for Kimi K3. Every paid plan's monthly allowance is shared by all models, so each count is the most you get if it all goes to that model.
Can I try Grok 4.7 and Kimi K3 for free?
Yes: both are in the free trial of 5 messages.
Which has the bigger context window, Grok 4.7 or Kimi K3?
Kimi K3: 1.05M tokens, against 500K tokens. A chat in llmwise holds up to 200k tokens, which fits in either, so the difference shows only through each maker's own API.
Can I use Grok 4.7 and Kimi K3 in the same chat?
Yes. Pick Grok 4.7 for one message and Kimi K3 for the next; the second sees the whole chat, including the first one's answer.
Which did better in your test runs, Grok 4.7 or Kimi K3?
On the same 50 prompts, run on September 27, 2026, Grok 4.7 passed 47 and Kimi K3 passed 46. The table on this page has each job, and every reply is published.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.