Hanzo Enso · available on Hanzo Cloud

One Model to Command Them All

Frontier-level performance without single-vendor lock-in. Enso dynamically orchestrates the world’s best models to tackle complex, multi-step tasks — plug collective intelligence into your workflows through a single API.

Enso is proprietary and available only via Hanzo Cloud. The open-weights Zen family stays free to self-host.

Available throughOpenAI + Anthropic APIClaude Code / Codex via Hanzo CLIHanzo SDKsMCP
What is Hanzo Enso

A multi-agent system, delivered as one model

Instead of hand-designing team roles and workflows, Enso learns to assemble agents from a pool and coordinate them through efficient, non-obvious collaboration patterns — automatically, per task.

01

One API — every model, every modality

Text, code, vision, documents, images, audio, and video through one OpenAI- and Anthropic-compatible endpoint. Enso unifies the frontier and open models across every modality — the first hyper-modal interface. You write one integration, not ten.

02

Superior on complex, multi-step work

Built for coding, reasoning, research, and other quality-critical workflows. Enso delivers stronger, more reliable results on hard, multi-step tasks than any single model.

03

You control the agent pool

Opt specific providers or models out of Enso’s pool to meet data, privacy, compliance, or org requirements — with a full audit trail of which models ran, on your organization’s cloud.

Enso vs Zen

Proprietary orchestration, on open foundations

Enso

Proprietary · Hanzo Cloud only

  • Learned orchestration over the best models
  • Flash · Pro · Ultra presets
  • One OpenAI + Anthropic endpoint
  • Managed, metered, audited on Hanzo Cloud
Explore Enso benchmarks

Zen

Open weights · run anywhere

  • Open-weight frontier models
  • Chat, code, and agents
  • Self-host on your own hardware
  • Free — or managed on Hanzo Cloud
Explore Zen models
The technology

Research-driven coordination for multi-agent intelligence

Enso is grounded in Hanzo’s research on learned model orchestration (HIP-0510): how a system can learn to assemble, route, and coordinate expert agents for each task instead of relying on hand-designed workflows.

Learned router

Microsecond routing

A lightweight coordinator scores every request and dispatches it to the right model in microseconds — routing overhead you can ignore, applied to every call.

Learned coordinator

Roles, turns, and verification

Enso assigns Thinker / Worker / Verifier roles and adaptively delegates across coding, math, reasoning, and knowledge tasks — coordinating diverse model pools to outperform any single worker.

How to use

Three presets — price × performance

Ultra, Pro, and Flash are distinct cost/quality contracts, monotonic in quality (98.0 > 96.0 > 92.9 GPQA-Diamond). Pick the one that fits your workload, or switch without changing your integration — one endpoint, OpenAI- and Anthropic-compatible.

Enso Ultra

Flagship

enso-ultra

98.0%GPQA-Diamond$5 → $20 /MTok

Maximum quality

Top-tier accuracy for hard, high-stakes problems — research reproduction, security analysis, and long-running autonomous work. Reaches 98.0% GPQA-Diamond at a price below premium single models like Opus and fable-5, which score far lower.

  • Research & paper reproduction
  • Security assessment
  • Deep, long-running tasks

Enso Pro

Default

enso · the default

96.0%GPQA-Diamond$3 → $12 /MTok

Balanced — the everyday default

Strong 96.0% GPQA-Diamond with sensible latency — the ideal default for real work: coding, code review, and responsive agents. Priced for everyday scale. Opt out of specific providers to meet data and compliance constraints.

  • Coding & code review
  • Responsive agents
  • Provider opt-out controls

Enso Flash

enso-flash

92.9%GPQA-Diamond$2 → $6 /MTok

Fastest, most economical

The high-volume default — lowest latency and cost for everyday chat, classification, extraction, and simple agent steps, at a strong 92.9% GPQA-Diamond.

  • High-volume, low latency
  • Cheapest per request
  • Great default for chat & tools
Quantitative results

Frontier capability, measured — without single-vendor risk

Enso reaches frontier-level results by routing each request to the right model in microseconds. Real, measured numbers — not a fabricated benchmark table.

98.0%
GPQA-Diamond
enso-ultra
<15µs
Routing overhead
per request
400+
Models available
frontier + open Zen
1 API
OpenAI + Anthropic
drop-in, or via Hanzo CLI

Enso delivers frontier capability without the risk of single-vendor export controls or lock-in — the router always dispatches to a currently-available model in its pool.

Efficiency & savings

The savings are the product

Enso delivers frontier accuracy and bills a fraction of what always calling a top model costs. enso-ultra reaches 98.0% GPQA-Diamond at a price below premium single models that score far lower.

98.0%
GPQA-Diamond · enso-ultra
< Opus
enso-ultra price, higher accuracy
89%
saved vs always calling a top model
3 tiers
Flash 92.9 · Pro 96.0 · Ultra 98.0 GPQA

Measured, stated plainly: enso-ultra 98.0%, enso-pro 96.0%, enso-flash 92.9% on GPQA-Diamond — each a distinct price/quality tier, all through one API.

01

Accuracy at cost

The goal is the top-left: high accuracy, low cost. enso-ultra sits there — 98.0% GPQA-Diamond at a price below premium single models like Opus and fable-5, which score far lower. Pro and Flash trade accuracy for even lower cost.

768084889296100$0.5$2$8$30$120GPQA %cheap ← output $/MTok → expensivegpt-5.5 — 93.6% GPQA-Diamond · $8.25/MTok (vendor-reported)gpt-5.5 93.6%gpt-5.2-pro — 93.2% GPQA-Diamond · $138.6/MTok (vendor-reported)gpt-5.2-pro 93.2%gpt-5.6-sol — 92.9% GPQA-Diamond · $25/MTokgpt-5.6-sol 92.9%kimi-k2.6 — 89.1% GPQA-Diamond · $2.71/MTok (vendor-reported)kimi-k2.6 89.1%qwen3.5-397b-a17b — 88.4% GPQA-Diamond · $2.04/MTok (vendor-reported)qwen3.5-397b-a17b 88.4%opus-4.8 — 87.4% GPQA-Diamond · $21/MTokopus-4.8 87.4%glm-5.2 — 85.6% GPQA-Diamond · $3.73/MTok (vendor-reported)glm-5.2 85.6%gemma-4-31b — 84.3% GPQA-Diamond · $0.44/MTok (vendor-reported)gemma-4-31b 84.3%fable-5 — 81.3% GPQA-Diamond · $42/MTokfable-5 81.3%enso-ultra — 98% GPQA-Diamond · $20/MTokenso-ultra 98%enso — 96% GPQA-Diamond · $12/MTokenso 96%enso-flash — 92.9% GPQA-Diamond · $6/MTokenso-flash 92.9%

Solid dots are Hanzo-measured; hollow dots are vendor-reported. enso-ultra (98.0%) leads on accuracy while pricing below premium single models — the win is accuracy-per-dollar across the three tiers.

02

Frontier accuracy without frontier prices

Top-tier results without paying a top-tier rate on every request. Output price per million tokens across models in the field:

fable-5not used to coordinate
$42
glm-5.2
$3.73
kimi-k2.6
$2.71
deepseek-v4-pro
$2.5
gemma-4-31b
$0.44
~95×cheaper coordination — for competitive quality

fable-5 costs $42/MTok and scores 81.3% solo on our harness — a premium coordinator that is expensive and worse. The cheap models Enso coordinates run $0.44$3.73/MTok: up to ~95× cheaper for the same coordination job.

03

Pay for what each request needs

Simple requests cost little; only the hardest work costs more. You pick the tier, and Enso keeps every request inside that price/quality contract.

simple request
~$0.09
the common case
harder request
~$0.43
more compute, only when needed
always-max baseline
~$0.82
paying top rate every time

Everyday requests cost a fraction of always paying the top rate — extra spend goes only to the hard fraction that needs it. Quality and cost are a property of the tier you choose, not a caller parameter. Per-request costs modeled from published token prices.

04

What it saves you

Estimate the monthly bill for your volume: always calling a top model, versus Enso routing most requests to a cheap one and escalating only the hard fraction.

Always top model
$35,000/mo
Enso (routed + adaptive)
$10,800/mo
You save
69%
$24,200/mo

Model: a top model bills the premium rate on every request; Enso serves the easy majority cheaply (~$0.002/req) and only the hard fraction costs more (~$0.43/req). Illustrative at published token prices for a typical short-answer request — your mix sets the exact number.

Built for

What teams build with Enso

Coding & code review

Enso finds the bugs a single model misses — comprehensive reviews that surface twenty issues where others flag three. Drop it into your existing coding tools unchanged.

Research & autonomy

Point Enso at a paper or a patent landscape and it works autonomously — reading, implementing, training, evaluating, and connecting sources across dozens of documents in hours, not days.

Security assessment

From one scoped instruction, Enso drives an end-to-end assessment — recon, injection and auth checks, and a clean report with evidence and retest steps — staying strictly inside scope.

Orchestration at scale

Frontier-level output with unusually strong persona and identity stability across long sessions — the property that matters most for production agent products.

Pricing

Pay for intelligence, not integrations

Usage-based, per-organization billing on Hanzo Cloud. When one agent handles a task you pay the standard rate for that model; when Enso coordinates several, you’re charged a single rate based on the top-tier model involved — never stacked fees.

Pay-as-you-go

For production workloads that need maximum reliability. Consumption-based tokens, served at higher priority, with transparent per-request cost you can predict and export.

  • Single rate — no stacked model fees
  • Per-request orchestration trace
  • Per-org usage & cost export
Subscription

For casual, everyday hands-on use. Every tier includes Flash, Pro, and Ultra — upgrade when you need longer, heavier, or more frequent sessions.

  • All three presets on every tier
  • Standard · Pro · Max usage tiers
  • Upgrade or downgrade anytime
FAQ

Questions, answered

Where can I use Hanzo Enso?

Enso is proprietary and available ONLY through Hanzo Cloud — a single endpoint that speaks both the OpenAI and Anthropic API styles natively. Point your existing OpenAI or Anthropic client at the Hanzo base URL and call an `enso-*` model id — or use it in Claude Code or Codex via the Hanzo CLI. (The open-weights Zen family, by contrast, is free to run on Hanzo Cloud or self-host anywhere.)

What are Flash, Pro, and Ultra?

The three default Enso presets: Flash for fast, high-volume work; Pro as the balanced everyday default for coding and agents; Ultra for maximum quality on hard, high-stakes problems. All behind one API — switch by changing the model id. Zen and other models remain available too.

How is Enso different from the Zen models?

Zen is the family of OPEN-WEIGHT models co-designed by Hanzo AI and the Zoo Labs Foundation (our nonprofit) that you can self-host. Enso is Hanzo’s PROPRIETARY orchestration layer on top — a learned router that assembles and coordinates the best available models (Zen and frontier) per task. Enso runs only on Hanzo Cloud; Zen runs anywhere.

Can I control which models or providers Enso uses?

Yes. Opt specific providers or models out of Enso’s pool to satisfy data-residency, privacy, or compliance requirements. Every request records which models actually ran.

Will my data be used to train models? Can I opt out?

No customer data is used to train models. Enso runs inside your Hanzo Cloud organization with a full audit trail; opt-out and data controls are first-class.

Can I see which underlying models Enso used for each query?

Yes. Each response carries the orchestration trace — the models selected, the roles they played, and the routing decisions — visible in the console and via the API.

Is Enso generally available?

Yes. Enso is available now on Hanzo Cloud and is the default for new chats and API requests — every default request routes through the Enso router, which selects the right tier (Flash, Pro, or Ultra) per task. Zen and other models stay available for explicit selection. Enterprise and dedicated deployment are available on request.

Ready to build with Hanzo Enso?

Enable Enso for your Hanzo Cloud organization, or talk to us about enterprise and dedicated deployment.