Plan
There’s a single subscription:
Manage your subscription at rumus.ai → Account, or in the app at Settings → Account.
How billing works
When you call a built-in model, your account is charged for the actual provider cost of that request plus a small Rumus token fee on top. The token fee covers routing, billing, support, and the infrastructure that keeps the built-in models working without you managing keys.Supported models
The full list, kept in sync with what’s exposed at Settings → AI → Models. Prices are per 1M tokens, in USD.Anthropic
OpenAI
Z.ai
DeepSeek
Moonshot AI
Privacy column: a check means the upstream provider has committed to not retaining or training on your data. Models without a check still send only the request itself — Rumus never logs the contents of your prompts or completions.
How to read the price columns
- Input / Output — what you send and what the model generates.
- Cache read — input tokens served from the prompt cache, billed cheaper than a fresh input. Long stable system prompts (skills, rules, big files) benefit the most.
- Cache write — input tokens being cached for future calls. Anthropic models price cache writes higher than fresh inputs; other providers cache automatically with no separate write price.
- Reasoning tokens — for models that “think” before answering, billed at the Output rate.
How much usage do I need?
A rough guide to help you choose between staying on the included credit and topping up:
These are observations, not limits. Your actual cost depends on which models you pick and how long the conversations are.
Monthly credits and overage
Pro includes $5 of AI credits each month. They’re spent first, automatically. Once they’re used up, requests continue against your account balance at the per-token prices above (plus the Rumus token fee). Unused credits don’t roll over — the $5 resets at the start of each billing cycle. There’s no automatic alerting when you cross a threshold. If you want to keep an eye on usage, check the dashboard regularly (see below).Top up your balance
If you expect to use more than the included credits, pre-load your account:1
Open Account settings
In the app, go to Settings → Account.
2
Click Top up
Next to your Balance, click the Top up button.
3
Pick a preset amount and confirm
Pick one of the preset top-up amounts and complete the payment. Once it clears, the credit lands in your balance and is immediately usable. You’ll get a receipt by email.
View your balance, usage, and bills
Two places to check what you’ve spent:Dashboard
Your current balance, monthly credit remaining, and usage charts (by day, by model).
Billing
The detailed per-request log: timestamp, model, token counts, base provider cost, the Rumus token fee, and total.
Privacy of built-in model traffic
Requests to built-in models are routed through Rumus’s API on the way to the upstream provider. We:- Do log token counts, model IDs, and timing for billing.
- Do not log message contents or completion text.
- Highlight models from providers with no-retention / no-training commitments using the privacy column above.
FAQ
What's the Rumus token fee — how much extra am I paying?
What's the Rumus token fee — how much extra am I paying?
A small percentage on top of the base provider cost. The exact rate and the amount per request are shown in the in-app pricing UI and on each line of your bill, so there’s no guessing — every charge is itemized.
Can I disable built-in models entirely and only use my own keys?
Can I disable built-in models entirely and only use my own keys?
Yes. Stay on Free, or stay on Pro for sync without using the built-in models — just add your own provider in Settings → AI → Models. Built-in models simply won’t be picked unless you choose them.
Do unused monthly credits roll over?
Do unused monthly credits roll over?
No. The $5 monthly credit resets at the start of each billing cycle.
A model on the list shows '–' for cache. Does it cache at all?
A model on the list shows '–' for cache. Does it cache at all?
Some upstream providers cache transparently with no separate cache-read price (you pay full input rate either way), while a few don’t cache at all. The dash means there’s no separate cache pricing to show — not that the model is broken.
Billing or model question we didn’t cover? Ask in the Rumus community.
Next steps
Bring your own key
Use Anthropic, OpenAI, Google, Z.AI, DeepSeek, Kimi, Ollama, or any OpenAI-compatible endpoint.
AI assistant
What you can actually do with the agent now that a model is set up.