Skip to content

GPT-6 Luna · Writing

GPT-6 Luna for writing: our test runs

GPT-6 Luna ran our 5 writing prompts in llmwise, through the same pipeline your messages take. Here's every reply as it came, whether it passed, what it cost and what it counts as on Pro, next to the other models.

Based on 5 of our test runs on , through OpenRouter with the app's own prompt and settings.

Short answer

In our test runs on September 27, 2026, GPT-6 Luna passed 5 of 5 of our writing prompts (both hard ones): tied 1st of 15 models. Each reply counted as 1 message on Pro, where it has 60 a day, and cost $0.0001 to run on average.

Prompt by prompt

Each of the 5 writing prompts, and how GPT-6 Luna's reply did.

GPT-6 Luna on each writing prompt
PromptResultCostTime
Announce a second bakery shop on LinkedInEverydayPassed: Graded 4.5 of 5 on average (lowest 4).$0.00012.9 s
Rewrite corporate jargon in plain wordsEverydayPassed: Graded 4.7 of 5 on average (lowest 4).$0.00012.4 s
Decline a meeting and offer two timesEverydayPassed: Graded 5.0 of 5 on average (lowest 5).$0.00011.3 s
A product announcement with five rulesHardPassed: Graded 4.0 of 5 on average (lowest 3).$0.00011.5 s
Argue both sides of free busesHardPassed: Graded 4.3 of 5 on average (lowest 4).$0.00012.5 s

Every writing reply from GPT-6 Luna

  1. Announce a second bakery shop on LinkedIn

    An everyday prompt: “Announce a second bakery shop on LinkedIn”, in full.

    Passed: Graded 4.5 of 5 on average (lowest 4).

    We’re delighted to share that Crumb & Co. is opening a second shop in Leeds on 14 November!
    
    What began as one small bakery has grown thanks to the lovely people who’ve stopped by, shared a loaf, and made us feel part of the community. We’re excited to bring our freshly baked bread, pastries, and warm welcomes to a new neighbourhood.
    
    To celebrate opening day, we’ll be giving away free coffee. Pop in, say hello, and enjoy a cup with us—we’d love to meet you.
    
    Thank you to everyone who has supported Crumb & Co. so far. We can’t wait to open the doors and start this next chapter together. See you in Leeds on 14 November!

    389 tokens in, 149 out · 2.9 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·

  2. Rewrite corporate jargon in plain words

    An everyday prompt: “Rewrite corporate jargon in plain words”, in full.

    Passed: Graded 4.7 of 5 on average (lowest 4).

    In Q3, teams across the company worked together to focus more on customers. Retention measures improved by 12%, and the rate of customer loss fell by 0.3 percentage points in the enterprise segment. The company also adjusted its customer contact channels to make better use of available capacity.

    417 tokens in, 124 out (59 of them reasoning) · 2.4 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·

  3. Decline a meeting and offer two times

    An everyday prompt: “Decline a meeting and offer two times”, in full.

    Passed: Graded 5.0 of 5 on average (lowest 5).

    Subject: Vendor Demo
    
    Hi Tom,
    
    Thanks for the invitation. I’m presenting to the board on Wednesday at 11:00, so I can’t make the demo then. Would Thursday at 10:00 or Friday at 14:00 work instead?
    
    Best,  
    [Your Name]

    395 tokens in, 64 out · 1.3 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·

  4. A product announcement with five rules

    A hard prompt: “A product announcement with five rules”, in full.

    Passed: Graded 4.0 of 5 on average (lowest 3).

    Meet the Nordvik 2, a desk lamp designed to bring flexible light to your workspace. Choose from five brightness levels, charge it via USB-C, and turn its arm up to 270 degrees to direct light where you need it. The Nordvik 2 costs €79 and ships on 1 December. Ready to brighten your desk?

    406 tokens in, 73 out · 1.5 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·

  5. Argue both sides of free buses

    A hard prompt: “Argue both sides of free buses”, in full.

    Passed: Graded 4.3 of 5 on average (lowest 4).

    Making buses free would give every resident reliable access to jobs, school, health care, and daily necessities, regardless of income. It would also speed boarding, reduce fare-enforcement costs, and encourage people to leave cars at home—easing congestion, pollution, and parking pressure. A city that treats public transit as essential infrastructure, rather than a service people must qualify to afford, can build a more connected and equitable community.
    
    But “free” buses still have to be paid for, and the bill could divert money from more urgent improvements: frequent service, safer stops, cleaner vehicles, and routes to underserved neighborhoods. If buses remain slow, crowded, or unreliable, removing fares may do little to persuade drivers to switch, while reducing funds that could make transit genuinely useful. Targeted fare assistance can protect riders who need it most without subsidizing every trip, including those by people who can readily afford to pay.

    384 tokens in, 186 out · 2.5 s · $0.0001 · 1 message on Pro · answered by openai/gpt-6-luna via OpenAI ·

How it compares on writing

GPT-6 Luna beside the job's picks among every model. 2 models share the hard-problems pick, GPT-6 Luna among them.

GPT-6 Luna beside other models on writing
ModelPassedOn ProCost per reply
GLM 5.3Our pick: best value4 of 5250 a month on Pro$0.0006
GPT-6 LunaOur pick: hard problems (shared), everyday5 of 560 a day on Pro$0.0001
On Pro: Pro's count on each model. Cost: what OpenRouter charged us per reply, on average.

Every limit is published. Paid plans also have a monthly fair-use limit on AI cost: Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster. Every limit, explained.

How these runs were done

Rubric (graded). The grader model scores the reply from 1 to 5 on each published criterion. It passes with an average of 4 or more and no criterion under 3, and only if it also meets the prompt's automatic rules (length, words it must or mustn't use).

The grader is Claude Opus 5.5 at low reasoning effort; its own replies are graded by GPT-6 Astra, so no model grades itself. Its prompt and every rubric are on the methods page.

How the runs were done, and every writing prompt.

GPT-6 Luna, and writing, elsewhere

Questions

Is GPT-6 Luna good for writing?

In our test runs it passed 5 of 5 writing prompts, tied 1st of the 15 models in llmwise. Every reply is on this page, so you can judge them yourself.

How many of my messages does a writing reply from GPT-6 Luna use?

1 message each on Pro, where it has 60 a day on Pro. The price of a message is fixed and shown before you send it, however long the reply.

How were these runs done?

The same way for every model: each prompt sent through llmwise's own pipeline, each reply checked the same way. The methods page has every prompt and how each is scored.

Claude, GPT, Gemini, DeepSeek, Grok, Kimi, and GLM, in one chat.

See what a message costs before you send it. Free is 5 messages to try; sign in with an email link, no password or card.