The short version
- You always see the price before you send. Next to the send button, llmwise shows how many messages you have left on the model you picked. For example: "Claude Opus 5 · 62 messages left this month".
- Every message has a fixed price. A reply costs the same whether it's one line or three pages. What you see before sending is exactly what's used.
- Fast models refill every day on paid plans. You get a fresh daily number at 00:00 UTC, up to your plan's fair-use limit.
- Free is a one-time trial: 5 messages on fast models and premium models up to size 2.
- Every limit is published on this page. There are no hidden caps and no "5x more usage" without a number.
Plans at a glance
Prices are in US dollars per month. Tax is added at checkout, depending on where you live.
| Free | Pro | Max | Ultra | Studio | |
|---|---|---|---|---|---|
| Price | $0 | $20 | $50 | $100 | $200 |
| Fast models: GPT-6 Luna, DeepSeek V4.1 Flash, GLM 5.3 Flash | In the 5 | 60 a day | 120 a day | 200 a day | 200 a day |
| Claude Haiku 4.5, Gemini 3.8 Flash, Grok 4.7, DeepSeek V4 Pro, GLM 5.3 | In the 5 | 250 a month | 800 | 1,800 | 4,000 |
| Claude Sonnet 5, Gemini 3.1 Pro, Kimi K3 | In the 5 | 125 a month | 400 | 900 | 2,000 |
| Claude Opus 5 | — | 62 a month | 200 | 450 | 1,000 |
| Claude Fable 5.1, GPT-6 Astra | — | 31 a month | 100 | 225 | 500 |
| Longest reply | ~6,000 words | ~6,000 words | ~12,000 words | ~12,000 words | ~12,000 words |
| Research reports, code runs, apps, images, connectors | — | ✓ | ✓ | ✓ | ✓ |
| Top-ups | — | ✓ | ✓ | ✓ | ✓ |
Free is 5 messages to try llmwise, one time. They work on the fast models and on the premium models up to size 2; there's no daily refill and no monthly allowance.
The premium rows share one allowance. The numbers above are how many messages you'd get if you used only that model. Mix freely: using a bigger model simply uses your allowance faster. The next section explains exactly how.
How your monthly allowance works
Each paid plan comes with one monthly allowance for premium models:
| Plan | Monthly allowance |
|---|---|
| Free | 5 messages, one time |
| Pro | 250 |
| Max | 800 |
| Ultra | 1,800 |
| Studio | 4,000 |
Every premium model has a size, meaning how much of the allowance one message uses. The size reflects what the model costs to run (see "How we set these numbers" below).
| Size | Models | One message uses |
|---|---|---|
| 1 | Claude Haiku 4.5, Gemini 3.8 Flash, Grok 4.7, DeepSeek V4 Pro, GLM 5.3 | 1 |
| 2 | Claude Sonnet 5, Gemini 3.1 Pro, Kimi K3 | 2 |
| 4 | Claude Opus 5 | 4 |
| 8 | Claude Fable 5.1, GPT-6 Astra | 8 |
You never have to do this math. llmwise does it for you and shows the result as plain messages:
Example: Pro, start of the month (allowance 250)
The model picker shows: Fable 31 left · Opus 62 left · Sonnet 125 left · Haiku 250 left.You send one message to Claude Fable (size 8). Your allowance goes from 250 to 242, so every count updates together:
Fable 30 left · Opus 60 left · Sonnet 121 left · Haiku 242 left.
The count next to the send button always shows the model you have selected. Switch models and it updates instantly.
Your allowance renews on the same day each month as your subscription (your billing date), not on the 1st.
Fast models: a daily number that refills
GPT-6 Luna, DeepSeek V4.1 Flash and GLM 5.3 Flash are our fast models. They're quick, capable and cheap to run, so they don't touch your monthly allowance. Instead every paid plan gets a daily number:
| Pro | Max | Ultra | Studio |
|---|---|---|---|
| 60 a day | 120 a day | 200 a day | 200 a day |
The daily number refills at 00:00 UTC every day. Unused fast messages don't carry over to the next day, but you get a fresh set every day: use them all and you're out only until the day ends, up to the fair-use limit below.
What counts as one message
- One message = one reply from the model. The price is fixed per message and shown before you send, however long the reply is (up to your plan's longest reply).
- Regenerating a reply, or editing your message and sending it again, is a new message.
- Stopping a reply partway still counts, because the model has already done the work.
- Web search: when the model searches the web, or reads a link you give it, each search counts as one more message of the model you're using. On Free, web search works inside your 5 trial messages. You control searching with the Search button next to the message box:
- Auto (default): the model decides when a search helps.
- Always: it searches before every answer.
- Never: no web access at all, so no search charges.
- Tool steps are included. When the model uses a connector (Gmail, Notion and others) or one of your MCP servers, those steps are part of the message price. Actions that send or change something always ask you first.
Long chats
Every time you send a message, the model re-reads the whole conversation so far. A very long chat therefore costs much more to answer than a short one. To keep prices fixed and fair, very long chats count a little more, and you always see this before you send:
| Conversation size (what the model reads) | Each message counts |
|---|---|
| Up to 64,000 tokens (roughly 48,000 words) | Normal price |
| Over 64,000 tokens | 2× |
| Over 128,000 tokens | 4× |
| Over 200,000 tokens | The chat is full. Start a new chat to keep going. |
Tip: starting a new chat for a new topic keeps every message at its normal price. The count next to the send button shows when a chat has grown long.
Tools and what they cost
Some tools do much more work than a normal reply. Each has a fixed price from your monthly allowance, shown before it runs. Anything that takes a while asks for your approval first.
| Tool | Uses from your allowance | For comparison, on Pro (250 a month) |
|---|---|---|
| Research report · Standard: a cited report from about 10–20 searches | 15 | about 16 Standard reports if you used only these |
| Research report · Deep: a longer, deeper cited report | 75 | about 3 Deep reports |
| Generate an image | 4 | 62 images |
| Run code (Python or Node, in a secure sandbox) | 1 per run | 250 runs |
| Backend app (a small live app with a preview) | 5 per started 10 minutes | 50 ten-minute blocks |
| Documents, spreadsheets, slide decks, downloads (PDF, Word, PowerPoint, Excel) | Free. Only the message that writes them counts. | — |
| Connectors (Gmail, Notion, Slack and more) and MCP servers | Included in the message price | — |
Research reports, code runs, backend apps, images, connectors and MCP servers are available on paid plans.
Top-ups
If you run low before your allowance renews, you can buy a top-up on any paid plan:
$10 (plus tax) adds 200 to your allowance. That's 25 Fable or GPT-6 Astra messages, 50 Opus, 100 Sonnet, Gemini Pro or Kimi, or 200 Haiku-class messages, or any mix.
- Top-ups are used after your plan's own monthly allowance.
- They last 12 months from purchase.
- If you cancel, you can keep using your top-ups for 30 days after your plan ends, for messages only (fast models count 1; web search needs a plan), within your last billing period's fair-use limit. Anything left then waits until 12 months after purchase, in case you come back.
- Top-ups are a one-time purchase, never a subscription.
Limits we publish, so you never hit a surprise
| Limit | Every plan |
|---|---|
| Replies per hour | 100 per rolling hour. This protects the service from runaway scripts; normal use never gets near it. |
| Longest reply | ~6,000 words on Free and Pro, ~12,000 words on Max, Ultra and Studio |
| Chat size | 200,000 tokens per chat (see "Long chats") |
| Free plan | Free: 5 messages, one time, on fast models plus premium models up to size 2 (Haiku-class, Sonnet, Gemini Pro, Kimi) |
| Free web search | Inside your 5 trial messages, up to 3 searches a day |
| Fair-use limit | Pro $7.50 · Max $20 · Ultra $42 · Studio $85 of AI cost per billing period (a top-up adds $4) |
If we ever need to change a limit, we update this page first and say so in the app. A message's price never changes after you've seen it.
Fair-use limit
Every paid plan includes up to a fixed amount of AI cost each billing period, measured at the list prices in "How we set these numbers": Pro $7.50, Max $20, Ultra $42, Studio $85. Using every message on your plan at typical sizes stays under it; very large messages and heavy research use it faster.
We tell you when you've used 80%. If you reach it, every message pauses, the fast models' too, until your allowance renews or you buy a top-up, which adds $4 of room. We never pause you without saying why.
After you cancel, your top-ups' 30 days count toward your last billing period's fair-use limit. If you reach it then, they pause for good: nothing renews, and top-ups can't be bought without a plan. Anything left waits until they expire, in case you come back.
How we set these numbers
We pay the model makers per token: a token is roughly ¾ of a word. The model reads your conversation (input) and writes a reply (output). Output costs more than input. Here are the public list prices we pay, per million tokens, and what a typical message costs us. A typical message is about 4,000 tokens read and 700 written.
| Model | Input ($ / 1M tokens) | Output ($ / 1M tokens) | Typical message costs us | Size |
|---|---|---|---|---|
| GPT-6 Luna | $0.10 | $0.50 | $0.0008 | Fast |
| GLM 5.3 Flash | $0.15 | $0.50 | $0.0010 | Fast |
| DeepSeek V4.1 Flash | $0.17 | $0.60 | $0.0011 | Fast |
| DeepSeek V4 Pro | $0.44 | $2.90 | $0.0038 | 1 |
| Gemini 3.8 Flash | $0.75 | $3.75 | $0.0056 | 1 |
| Claude Haiku 4.5 | $1.00 | $5.00 | $0.0075 | 1 |
| GLM 5.3 | $1.40 | $4.40 | $0.0087 | 1 |
| Grok 4.7 | $1.60 | $4.80 | $0.0098 | 1 |
| Claude Sonnet 5 | $2.00 | $10.00 | $0.0150 | 2 |
| Gemini 3.1 Pro | $2.00 | $12.00 | $0.0164 | 2 |
| Kimi K3 | $3.00 | $15.00 | $0.0225 | 2 * |
| Claude Opus 5 | $5.00 | $25.00 | $0.0375 | 4 |
| Claude Fable 5.1 | $10.00 | $50.00 | $0.0750 | 8 |
| GPT-6 Astra | $10.00 | $50.00 | $0.0750 | 8 |
The rule: one unit of your allowance covers about $0.01 of a typical message. A model's size is its typical cost rounded up to the next of 1, 2, 4 or 8 units. That's why Claude Fable, which costs about ten times more to run than Claude Haiku, uses 8 where Haiku uses 1.
* Kimi K3's typical cost would round up to size 4. We price it at 2 on purpose, to make it easy to try.
Why a price per message instead of per token? Per-token billing means you only find out the cost after the reply. We'd rather you know before you send. Some messages cost us more than their price and some less; we take that risk, not you.
Prices can change. Model makers sometimes change their prices. If a change moves a model to a different size, we'll update this page and tell you in the app before it takes effect. A message you've already been shown a price for is never re-priced.
Upgrading, downgrading and cancelling
- Upgrade anytime. You're charged the difference for the rest of your billing month, plus tax. We show the exact amount before you confirm, whenever we can calculate it. Your new allowance starts right away.
- Downgrade anytime. The change happens at the end of your billing month. You keep your current plan until then.
- Cancel anytime in Settings → Billing, in two clicks. You keep your plan until the end of the month you've paid for, and nothing in your account is deleted. We'll ask why you're leaving, but you don't have to answer.
- Your chats, files, memory and projects stay yours on every plan, including Free.
Where to see your usage
- Next to the send button: messages left on the model you picked.
- In the model picker: messages left on every model.
- Settings → Billing: your plan, allowance left, top-ups, and your renewal date.