Tested prompt · RAG and answering from documents
Two sections in one answer: every AI model's reply, tested
We sent this everyday RAG and answering from documents prompt to all 16 models in llmwise, the same way the app sends a message, and checked every reply the same way. Here's each one as it came, with whether it passed, what it cost and how long it took.
Based on 16 of our test runs on , through OpenRouter with the app's own prompt and settings. Updated .
Short answer
All 16 models passed this RAG and answering from documents prompt's check (key facts). The cheapest reply that passed was GPT-6 Luna's, at $0.000079; the fastest, GLM 5.3's in 0.5 s. The dearest reply, Claude Fable 5.1's, cost 168 times as much ($0.0133).
The prompt, as sent, and its check
Checked by key facts, the same way for every model.
Two sections in one answer (everyday)
[9 lines every RAG and answering from documents prompt of ours shares, word for word: the whole prompt, on the methods page] Question: How many unused holiday days can I carry over, and can I carry over unused learning budget too?
The answer must state “5”, “§5”, “§6”.
Exactly what this prompt's replies are checked against, with every other prompt of our test runs.
Every model's result
All 16 models on this prompt, in catalog order.
| Model | Result | Cost | Time | Reply |
|---|---|---|---|---|
| Claude Fable 5.1Anthropic | Passed: Stated all 3 facts with the section. | $0.0133 | 6.3 s | 42 tokens |
| Claude Opus 5.5Anthropic | Passed: Stated all 3 facts with the section. | $0.0053 | 2.9 s | 43 tokens |
| Claude Sonnet 5.5Anthropic | Passed: Stated all 3 facts with the section. | $0.0027 | 1.3 s | 44 tokens |
| Claude Sonnet 5Anthropic | Passed: Stated all 3 facts with the section. | $0.0022 | 1.8 s | 42 tokens |
| Claude Haiku 4.5Anthropic | Passed: Stated all 3 facts with the section. | $0.0011 | 1.3 s | 71 tokens |
| GPT-6 AstraOpenAI | Passed: Stated all 3 facts with the section. | $0.0082 | 1.7 s | 35 tokens |
| GPT-6 SolOpenAI | Passed: Stated all 3 facts with the section. | $0.0016 | 1.2 s | 33 tokens |
| GPT-6 LunaOpenAI | Passed: Stated all 3 facts with the section. | $0.000079 | 1.4 s | 29 tokens |
| Gemini 3.1 Pro (preview)Google | Passed: Stated all 3 facts with the section. | $0.0047 | 5.6 s | 24 tokens |
| Gemini 3.8 FlashGoogle | Passed: Stated all 3 facts with the section. | $0.00065 | 3.1 s | 32 tokens |
| DeepSeek V4.1 FlashDeepSeek | Passed: Stated all 3 facts with the section. | $0.00018 | 1.3 s | 38 tokens |
| DeepSeek V4 ProDeepSeek | Passed: Stated all 3 facts with the section. | $0.00029 | 1.3 s | 30 tokens |
| Grok 4.7xAI | Passed: Stated all 3 facts with the section. | $0.0018 | 1.5 s | 27 tokens |
| Kimi K3Moonshot | Passed: Stated all 3 facts with the section. | $0.0020 | 2.6 s | 42 tokens |
| GLM 5.3Z.ai | Passed: Stated all 3 facts with the section. | $0.00046 | 0.5 s | 33 tokens |
| GLM 5.3 FlashZ.ai | Passed: Stated all 3 facts with the section. | $0.00010 | 1.9 s | 37 tokens |
Cost: what OpenRouter charged us for the reply. Time: from sending to the whole reply. Reply: its length in tokens, thinking not counted. In llmwise you pay per message, not per token: each of these replies counted as one message on Pro.
Every reply
Every reply passed: here they are from the cheapest up.
GPT-6 Luna
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days (§5). Unused learning budget does not carry over (§6).
647 tokens in, 29 out · 1.4 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·
GLM 5.3 Flash
Passed: Stated all 3 facts with the section.
You can carry over up to **5 unused holiday days** to the next year (§5). Unused learning budget, however, **does not carry over** (§6).
667 tokens in, 93 out (56 of them reasoning) · 1.9 s · $0.0001 · 1 message on Pro · answered by z-ai/glm-5.3-flash via Wafer ·
DeepSeek V4.1 Flash
Passed: Stated all 3 facts with the section.
**Holiday days:** Up to 5 unused days carry over to the next year (§5). **Learning budget:** No — unused learning budget does not carry over (§6).
694 tokens in, 63 out (25 of them reasoning) · 1.3 s · $0.0002 · 1 message on Pro · answered by deepseek/deepseek-v4.1-flash via Novita ·
DeepSeek V4 Pro
Passed: Stated all 3 facts with the section.
Up to 5 unused holiday days carry over to the next year (§5). Unused learning budget does **not** carry over (§6).
666 tokens in, 84 out (54 of them reasoning) · 1.3 s · $0.0003 · 1 message on Pro · answered by deepseek/deepseek-v4-pro-0813 via Ionstream ·
GLM 5.3
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days to the next year (§5). No, unused learning budget does not carry over (§6).
690 tokens in, 34 out (1 of them reasoning) · 0.5 s · $0.0005 · 1 message on Pro · answered by z-ai/glm-5.3 via Wafer ·
Gemini 3.8 Flash
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days to the next year (§5). No, unused learning budget does not carry over (§6).
706 tokens in, 32 out · 3.1 s · $0.0006 · 1 message on Pro · answered by google/gemini-3.8-flash via Google ·
Claude Haiku 4.5
Passed: Stated all 3 facts with the section.
According to the handbook: **Holiday days:** You can carry over up to 5 unused days to the next year (§5). **Learning budget:** Unused learning budget does **not** carry over. Your €1,200 annual learning budget is use-it-or-lose-it (§6).
717 tokens in, 71 out · 1.3 s · $0.0011 · 1 message on Pro · answered by anthropic/claude-haiku-4.5 via Amazon Bedrock ·
GPT-6 Sol
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days to the next year (§5). Unused learning budget does not carry over (§6).
647 tokens in, 33 out · 1.2 s · $0.0016 · 1 message on Pro · answered by openai/gpt-6-sol via OpenAI ·
Grok 4.7
Passed: Stated all 3 facts with the section.
Up to 5 unused holiday days can carry over to the next year (§5). Unused learning budget does not carry over (§6).
1,881 tokens in, 97 out (70 of them reasoning) · 1.5 s · $0.0018 · 1 message on Pro · answered by x-ai/grok-4.7 via xAI ·
Kimi K3
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days to the next year (§5). No — unused learning budget doesn’t carry over (§6).
737 tokens in, 162 out (120 of them reasoning) · 2.6 s · $0.0020 · 1 message on Pro · answered by moonshotai/kimi-k3 via Wafer ·
Claude Sonnet 5
Passed: Stated all 3 facts with the section.
Up to 5 unused holiday days can carry over to the next year, but unused learning budget does not carry over (§5, §6).
905 tokens in, 42 out · 1.8 s · $0.0022 · 1 message on Pro · answered by anthropic/claude-sonnet-5 via Claude Platform on AWS ·
Claude Sonnet 5.5
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days into the next year (§5). Unused learning budget doesn't carry over (§6).
909 tokens in, 44 out · 1.3 s · $0.0027 · 1 message on Pro · answered by anthropic/claude-sonnet-5.5 via Anthropic ·
Gemini 3.1 Pro
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days (§5). You cannot carry over unused learning budget (§6).
706 tokens in, 278 out (254 of them reasoning) · 5.6 s · $0.0047 · 1 message on Pro · answered by google/gemini-3.1-pro-preview via Google ·
Claude Opus 5.5
Passed: Stated all 3 facts with the section.
You can carry over up to 5 unused holiday days to the next year (§5). Unused learning budget can't be carried over (§6).
907 tokens in, 43 out · 2.9 s · $0.0053 · 1 message on Pro · answered by anthropic/claude-opus-5.5 via Claude Platform on AWS ·
GPT-6 Astra
Passed: Stated all 3 facts with the section.
You can carry over up to **5 unused holiday days** to the next year (§5). **Unused learning budget cannot be carried over** (§6).
647 tokens in, 35 out · 1.7 s · $0.0082 · 1 message on Pro · answered by openai/gpt-6-astra via OpenAI ·
Claude Fable 5.1
Passed: Stated all 3 facts with the section.
Up to 5 unused holiday days can carry over to the next year (§5). Unused learning budget does not carry over (§6).
907 tokens in, 42 out · 6.3 s · $0.0133 · 1 message on Pro · answered by anthropic/claude-fable-5.1 via Anthropic ·
More RAG and answering from documents prompts
The other RAG and answering from documents prompts, each with every model's reply, and the results across all five.
Questions
Which AI does best on “Two sections in one answer”?
All 16 models passed this RAG and answering from documents prompt's check (key facts). The cheapest reply that passed was GPT-6 Luna's, at $0.000079; the fastest, GLM 5.3's in 0.5 s. The dearest reply, Claude Fable 5.1's, cost 168 times as much ($0.0133).
What does a reply to “Two sections in one answer” cost?
Through the models' APIs, what OpenRouter charged us ran from $0.000079 (GPT-6 Luna) to $0.0133 (Claude Fable 5.1) for this prompt. In llmwise you don't pay by the token: a reply like these counts as one message on Pro, whichever model answers.
Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.
See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.