Try Hanzo
Transparent value · No hidden fees

Plans that grow with you

You are charged exactly the price shown. Each plan includes premium and Hanzo model usage each month.

Free

Meet Hanzo. Chat and build with Hanzo's free models.

Free

No card needed

Try Hanzo

No card needed · Instant access

  • Hanzo's free models: enso-free and zen-free
  • Chat on free models, with limited usage
  • The app's core features
  • Upgrade any time from inside the app

Pro

For one person: every room in the hanzo.ai app, with premium and Hanzo model usage included each month.

$20 USD/month

Billed monthly

Try Hanzo

No commitment · Cancel anytime

  • Everything in Free and:
  • Every room in the hanzo.ai app
  • Includes premium and Hanzo model usage each month
  • Top up prepaid credit anytime for more
  • Limited free usage when you reach your plan's included usage
  • MCP, the SDK and the CLI included

Max

More included usage than Pro, with Enso orchestration, managed agents and a resident bot.

From $100 USD/month

Billed monthly

Try Hanzo

No commitment · Cancel anytime

  • Everything in Pro, plus:
  • More included usage than Pro
  • Includes premium and Hanzo model usage each month
  • Top up prepaid credit anytime for more
  • Limited free usage when you reach your plan's included usage
  • Enso orchestration — dedicated routing through Hanzo's foundational framework model
  • Managed agents and a resident bot, scheduled in the background
  • Priority GPU inference queues

Each plan includes premium and Hanzo model usage each month. When it runs out you keep chatting with limited free usage, or choose to continue with prepaid credit, which you can top up anytime.

What a bill is made of

Four components, each on your spend breakdown under its own name. Two are quoted before they run; two are metered from what they used.

Model inference

per token, five tiers
Quoted before it runs
Input, cached reads, cache writes, output and reasoning are metered as five disjoint tiers, at the tariff for the completion window the request asked for.
inputcache_readcache_writeoutputreasoning

Computer

per second, while running
Quoted before it runs
vCPU and memory on what the machine was given. Nothing to create one, nothing while it sleeps, and the boot disk is inside the allowance — so a computer waiting between turns meters nothing.

Web tools

per call
Metered from what it used
Priced from what the provider charged, so a web call is booked after it runs. A fetch the domain policy refuses is stopped before the request and costs nothing.

Media generation

what the job cost
Metered from what it used
Billed once per finished job at its real cost; a render cannot be quoted before it runs, so it is admitted against the headroom under the agent's per-task cap instead. A job that fails is not billed.

The rates

Every rate here is the price you pay — the number the ledger books and the number your spend breakdown reports.

Rate
Computer · vCPU
$0.0504/hr$0.000014/second
Charged by the second, while running.
Computer · memory
$0.0162/GiB-hr$0.0000045/GiB-second
Charged by the second, on the memory the machine was given.
Computer · paused$0No vCPU, no memory, no storage.
Computer · created$0There is no creation fee.
web_search$0.002158/callOne query, up to ten results.
web_fetch$0.001079/callOne page read as text.
MCP, SDK and CLI$0The interfaces are not metered.
Seats$0No per-seat fee, however many people use the organization.

The completion window is the biggest lever

The same model at a different window is a different price, chosen per request. Pay the interactive tariff only when a person is waiting.

immediateAnswers now, at the highest tariff.
priorityAnswers soon, at a lower one.
looseAnswers eventually, at the lowest.

Every model, at the rate the ledger books

Model inference is metered per token in five disjoint tiers, at the tariff for the completion window the request asked for.

Measured, not modelled.

Every figure is corrected against vendor invoice brackets — list-price rate cards overstate real cost by up to 4.2x. The method, and the cost per completed task, are on the bench.

What the same stack costs to run yourself: the comparison and the calculator. Starting a company on it: the startup programme.