Plans that grow with you
You are charged exactly the price shown. Each plan includes premium and Hanzo model usage each month.
Free
Meet Hanzo. Chat and build with Hanzo's free models.
No card needed
No card needed · Instant access
- Hanzo's free models: enso-free and zen-free
- Chat on free models, with limited usage
- The app's core features
- Upgrade any time from inside the app
Pro
For one person: every room in the hanzo.ai app, with premium and Hanzo model usage included each month.
Billed monthly
No commitment · Cancel anytime
- Everything in Free and:
- Every room in the hanzo.ai app
- Includes premium and Hanzo model usage each month
- Top up prepaid credit anytime for more
- Limited free usage when you reach your plan's included usage
- MCP, the SDK and the CLI included
Max
More included usage than Pro, with Enso orchestration, managed agents and a resident bot.
Billed monthly
No commitment · Cancel anytime
- Everything in Pro, plus:
- More included usage than Pro
- Includes premium and Hanzo model usage each month
- Top up prepaid credit anytime for more
- Limited free usage when you reach your plan's included usage
- Enso orchestration — dedicated routing through Hanzo's foundational framework model
- Managed agents and a resident bot, scheduled in the background
- Priority GPU inference queues
Each plan includes premium and Hanzo model usage each month. When it runs out you keep chatting with limited free usage, or choose to continue with prepaid credit, which you can top up anytime.
What a bill is made of
Four components, each on your spend breakdown under its own name. Two are quoted before they run; two are metered from what they used.
Model inference
per token, five tiersComputer
per second, while runningWeb tools
per callMedia generation
what the job costThe rates
Every rate here is the price you pay — the number the ledger books and the number your spend breakdown reports.
| Rate | ||
|---|---|---|
| Computer · vCPU | $0.0504/hr$0.000014/second | Charged by the second, while running. |
| Computer · memory | $0.0162/GiB-hr$0.0000045/GiB-second | Charged by the second, on the memory the machine was given. |
| Computer · paused | $0 | No vCPU, no memory, no storage. |
| Computer · created | $0 | There is no creation fee. |
| web_search | $0.002158/call | One query, up to ten results. |
| web_fetch | $0.001079/call | One page read as text. |
| MCP, SDK and CLI | $0 | The interfaces are not metered. |
| Seats | $0 | No per-seat fee, however many people use the organization. |
The completion window is the biggest lever
The same model at a different window is a different price, chosen per request. Pay the interactive tariff only when a person is waiting.
Every model, at the rate the ledger books
Model inference is metered per token in five disjoint tiers, at the tariff for the completion window the request asked for.
Measured, not modelled.
Every figure is corrected against vendor invoice brackets — list-price rate cards overstate real cost by up to 4.2x. The method, and the cost per completed task, are on the bench.
What the same stack costs to run yourself: the comparison and the calculator. Starting a company on it: the startup programme.