Pricing
One subscription, every model
Plans include monthly token capacity that works across every model in your tier. Payments powered by Waffo Pancake.
How token billing works
Every request consumes credit from your plan based on three token types. Cache hits are dramatically cheaper — agent-style workloads that reuse context get the most out of every plan.
Monthly capacity per plan
Millions of tokens per month. Ranges depend on which model in your tier you use — flash models go furthest.
| Plan | Input only | Cache only | Output only | Typical usage* |
|---|---|---|---|---|
| Lite | 20M | 70–161M | ±6M | 24–28M |
| Basic | 16–40M | 140–804M | 4–12M | 23–55M |
| Starter | 32–80M | 280M–1.6B | 8–24M | 45–111M |
| Pro | 16–235M | 96M–4.7B | 5–70M | 29–324M |
| Max | 24–603M | 244M–12.1B | 5–179M | 28–832M |
Input/cache/output-only columns are theoretical extremes — real usage is always a mix. *Typical usage assumes an agent-style mix: every 1M input tokens comes with ~0.7M cache reads and ~0.1M output tokens. Capacity resets monthly while your subscription is active.
