Skip to content

Comparison

Kimi vs GLM

On llmwise Pro, Kimi K3 up to 125 messages a month and GLM 5.3 up to 250. We ran the same 50 prompts across 10 jobs on every Kimi and GLM model and published every reply: each family's pick against the other's job by job, the prompts where they split, then their plans and how the lineups differ.

Based on 150 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .

Short answer

In our test runs on September 27, 2026, the same 50 prompts across 10 jobs: Kimi's one model passed 46 of 50 replies and GLM's 2 models passed 88 of 100 replies. Job by job, both families' best models shared the top result on 8 of the 10 jobs; GLM's alone had it on summarization and customer support.

Kimi vs GLM, job by job

On each job, Kimi's pick against GLM's: the model of each family that passed the most of the job's 5 prompts (then the most hard ones, then the cheaper). The jobs where they differ most come first.

  • Summarization: Kimi K3 passed 3 of 5 and GLM 5.3 4 of 5. GLM 5.3 answered 3.0× sooner at the median, 3.4 s against 1.1 s. GLM 5.3 cost 6.3× less, $0.0172 against $0.0027 for the 5 replies. Kimi K3's replies ran 14% longer, in tokens of reply, thinking not counted.

  • Customer support: Kimi K3 passed 4 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 1.9× sooner at the median, 4.3 s against 2.3 s. GLM 5.3 cost 9.8× less, $0.0228 against $0.0023 for the 5 replies. Kimi K3's replies ran 16% longer, in tokens of reply, thinking not counted.

  • Coding: Kimi K3 passed 5 of 5 and GLM 5.3 Flash 5 of 5. Kimi K3 answered 2.0× sooner at the median, 4.1 s against 8.4 s. GLM 5.3 Flash cost 25.0× less, $0.0316 against $0.0013 for the 5 replies. Kimi K3's replies ran 15% longer, in tokens of reply, thinking not counted.

  • Writing: Kimi K3 passed 4 of 5 and GLM 5.3 4 of 5. GLM 5.3 answered 1.7× sooner at the median, 3.0 s against 1.8 s. GLM 5.3 cost 6.7× less, $0.0192 against $0.0029 for the 5 replies. Their replies ran to about the same length.

  • Math: Kimi K3 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 2.0× sooner at the median, 5.4 s against 2.7 s. GLM 5.3 Flash cost 24.9× less, $0.0187 against $0.0008 for the 5 replies. GLM 5.3 Flash's replies ran 21% longer, in tokens of reply, thinking not counted.

  • Data analysis: Kimi K3 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 1.8× sooner at the median, 3.5 s against 1.9 s. GLM 5.3 cost 3.5× less, $0.0339 against $0.0096 for the 5 replies. Kimi K3's replies ran 15% longer, in tokens of reply, thinking not counted.

  • Translation: Kimi K3 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 1.3× sooner at the median, 5.2 s against 4.1 s. GLM 5.3 Flash cost 36.2× less, $0.0179 against $0.0005 for the 5 replies. Kimi K3's replies ran 89% longer, in tokens of reply, thinking not counted.

  • SQL: Kimi K3 passed 5 of 5 and GLM 5.3 Flash 5 of 5. Their median waits were close, 1.6 s against 1.4 s. GLM 5.3 Flash cost 18.3× less, $0.0118 against $0.0006 for the 5 replies. Kimi K3's replies ran 20% longer, in tokens of reply, thinking not counted.

  • RAG and answering from documents: Kimi K3 passed 5 of 5 and GLM 5.3 Flash 5 of 5. GLM 5.3 Flash answered 1.2× sooner at the median, 2.6 s against 2.2 s. GLM 5.3 Flash cost 9.5× less, $0.0073 against $0.0008 for the 5 replies. GLM 5.3 Flash's replies ran 18% longer, in tokens of reply, thinking not counted.

  • Agents and tool use: Kimi K3 passed 5 of 5 and GLM 5.3 5 of 5. GLM 5.3 answered 2.2× sooner at the median, 2.0 s against 0.9 s. GLM 5.3 cost 6.1× less, $0.0135 against $0.0022 for the 5 replies. Kimi K3's replies ran 25% longer, in tokens of reply, thinking not counted.

The 4 prompts only one of Kimi and GLM passed

Where one family's pick passed a prompt and the other's didn't, in each check's own words.

  • An email thread in one sentence (summarization): GLM 5.3 passed and Kimi K3 didn't. Kimi K3: Graded 5.0 of 5 on average (lowest 5); but 32 words, over the 30 allowed. GLM 5.3: Graded 4.3 of 5 on average (lowest 3).

  • A refund request outside the window (customer support): GLM 5.3 passed and Kimi K3 didn't. Kimi K3: Graded 4.0 of 5 on average (lowest 2). GLM 5.3: Graded 4.7 of 5 on average (lowest 4).

  • A product announcement with five rules (writing): GLM 5.3 passed and Kimi K3 didn't. Kimi K3: Graded 4.3 of 5 on average (lowest 4); but doesn't end with a question. GLM 5.3: Graded 4.3 of 5 on average (lowest 4).

  • Argue both sides of free buses (writing): Kimi K3 passed and GLM 5.3 didn't. Kimi K3: Graded 4.7 of 5 on average (lowest 4). GLM 5.3: Graded 3.7 of 5 on average (lowest 3); but a paragraph of 92 words, over the 90 allowed.

Each Kimi model against each GLM model

Every Kimi model against every GLM model on the same 50 prompts: each one's passes, how many prompts split them, and who was cheaper and quicker.

  • Kimi K3 vs GLM 5.3: 46 and 48 of 50; 4 prompts split them; GLM 5.3's replies cost 5.3× less in all, and GLM 5.3 answered sooner on 44, Kimi K3 on 2. Kimi K3 vs GLM 5.3.

  • Kimi K3 vs GLM 5.3 Flash: 46 and 40 of 50; 6 prompts split them; GLM 5.3 Flash's replies cost 20.1× less in all, and Kimi K3 answered sooner on 28, GLM 5.3 Flash on 16.

One summarization prompt, both replies

Kimi K3 and GLM 5.3 on a summarization prompt, the job where they differed most: both replies as they came.

An everyday prompt: “An email thread in one sentence”, in full.

  • Kimi K3

    Failed: Graded 5.0 of 5 on average (lowest 5); but 32 words, over the 30 allowed.

    Due to a paper shortage delaying brochure shipping to the 18th, Rosa and Idris arranged to pick them up themselves, booking a van for the 18th ahead of the 20th trade fair.

    574 tokens in, 65 out (9 of them reasoning) · 6.4 s · $0.0030 · 1 message on Pro · answered by moonshotai/kimi-k3 via Fireworks ·

  • GLM 5.3

    Passed: Graded 4.3 of 5 on average (lowest 3).

    The brochure shipment was delayed to the 18th due to a paper shortage, so Rosa will pick them up herself, with Idris booking a van for that morning.

    503 tokens in, 37 out · 1.0 s · $0.0002 · 1 message on Pro · answered by z-ai/glm-5.3 via Baidu ·

Every model, every job

All 3 Kimi and GLM models in llmwise across the 50 prompts, with what a reply cost to run and each model's count on Pro.

Every Kimi and GLM model across our test runs
ModelPassedHard onesCost per replyOn Pro
Kimi K3Moonshot46 of 5018 of 20$0.0039Up to 125 a month
GLM 5.3Z.ai48 of 5019 of 20$0.00073Up to 250 a month
GLM 5.3 FlashZ.ai40 of 5016 of 20$0.00019Up to 60 a day

Passed: replies that passed their check, of those scored. Cost: what OpenRouter charged us per reply, on average; in llmwise you pay per message, not per token. Every prompt and how it's scored.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

The lineups at a glance

What follows from each model's facts in our catalog.

  • The lineups

    Kimi: one model, Kimi K3. GLM: 2 models, GLM 5.3 and GLM 5.3 Flash.

  • Price per message

    Kimi's one model is Kimi K3 (125 messages a month on Pro); the least expensive GLM model is GLM 5.3 Flash (60 messages a day on Pro).

  • Context window

    Both go up to 1.05M tokens of context (Kimi K3 and GLM 5.3).

  • Images and PDFs

    GLM 5.3 doesn't read images. Kimi K3, GLM 5.3, and GLM 5.3 Flash get a PDF's text rather than the file itself.

  • On the Free plan

    Free's one-time trial of 5 messages covers Kimi K3, GLM 5.3, and GLM 5.3 Flash. Paid plans have every model, with messages every month.

Model by model

Two named models side by side, prompt by prompt, each with its messages on every plan.

More head-to-heads

Each of Kimi and GLM against the other families, every job from the same test runs.

Where your messages go

Kimi and GLM models are served only through OpenRouter, by endpoints that don't store or train on prompts. The Privacy Policy has the details.

Questions

Which is better, Kimi or GLM?

In our test runs on September 27, 2026, the same 50 prompts across 10 jobs: Kimi's one model passed 46 of 50 replies and GLM's 2 models passed 88 of 100 replies. Job by job, both families' best models shared the top result on 8 of the 10 jobs; GLM's alone had it on summarization and customer support.

Which is cheaper, Kimi or GLM?

In llmwise, the least expensive Kimi model is Kimi K3 (125 messages a month on Pro), and the least expensive GLM model is GLM 5.3 Flash (60 messages a day on Pro). At API list prices (September 2026), a typical message of 4,000 tokens in and 700 out costs $0.0225 on Kimi K3 and $0.0010 on GLM 5.3 Flash.

Can I use Kimi and GLM in the same chat?

Yes. Pick a model for each message; when you switch, the next model sees the whole conversation, including the other one's answers.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.