Frontier AI models, agent infrastructure, and the platform powering the next generation of intelligent applications.
Hanzo Cloud has no checked-in OpenAPI file. The live route table is the source, projected three ways, and a bijection test fails the build if any projection drifts. 1,362 operations across 963 paths, generated per process at request time.
We set out to see how fast our cloud data plane could go on an NVIDIA GB10. The answer taught us more about NIC receive queues than about our own code — and it more than doubled our throughput on the same wire.
Enso is Hanzo's learned router: no single frontier model leads every benchmark, so Enso routes each request to the model most likely to win it — and learns from your feedback as it goes. Microsecond routing on a CPU, a transparent per-request meter, and a 1% fee that pays for itself. Here is how it works, the measured results, and the cost-transparency angle a monolith cannot offer.
Hanzo Router places each request on the cheapest model that can still serve it — reuse a local model for free, load one that fits, or fall through to cloud. Here's the idea, what it feels like in production (qualitative), and the honest cost arithmetic (quantitative) — no fabricated benchmark.
A ClickHouse-derived columnar warehouse, rebuilt around storage/compute separation on Hanzo S3, profile-guided and link-time optimized to roughly half the memory, with gRPC stripped out and a path to leaderless post-quantum coordination.
How Hanzo Datastore replaces the NuRaft engine under ClickHouse Keeper with Lux Quasar — post-quantum, leaderless consensus linked directly into the C++ server — while keeping the ZooKeeper API that replication speaks.
Hanzo Cloud is built on Proof of AI — deterministic inference any validator can re-verify on-chain — and on a commitment that 25% of what the cloud earns flows back to the open-source developers whose code it runs.
LP-138 lands. accel is a capability-negotiated dispatch layer so accelerated crypto, ZK, and FHE are one codebase across runtimes.
A single SQLite file your agent reads from every runtime. Zero LLM round-trips for graph ops, FTS5 plus vector plus typed-edge extraction at write time.
A hybrid X25519 + ML-KEM-768 transport that drops in where CurveZMQ used to live. PQ by default, two backends, every ZMQ pattern supported.
Hanzo's flagship LLM — a 1.04 trillion parameter Mixture of Experts model fused from top language models, with distributed training and on-chain model attestation.
Machine-checked Lean 4 proofs for agent safety, MCP protocol, PoAI consensus, and platform operations
How BitDelta's 1-bit delta compression lets us serve 14 Zen model variants from shared GPU infrastructure — the math, the architecture, and the tradeoffs.
Twelve years of building AI infrastructure — from a real-time recommendation engine in 2014 to 94+ frontier models and a post-quantum blockchain in 2026. The technical story of how we got here.
Hanzo Network launches Business through Ultra enterprise cloud tiers — dedicated CPU machines from 8 to 96 vCPUs with 71–81% cost advantage over hyperscaler equivalents.
Hanzo AI introduces the first platform combining 100+ AI models, cloud compute, GPU access, and 260+ MCP tools under a single developer account.
Hanzo Bot opens on-demand H100 GPU access for AI agent developers at $3.48/hr — 72% less than AWS equivalent pricing.
Hanzo Bot launches with the cheapest AI agent hosting in the industry — starting at $5/mo with free credit, integrated access to 100+ models, and H100 GPUs at $3.48/hr.
Hanzo Network launches a developer cloud that undercuts DigitalOcean, AWS Lightsail, Vultr, and Linode by 10–54% — with zero egress fees, DDoS protection, and consistent global pricing included.
Hanzo Bot delivers speech-to-text at $0.0009/minute — 85% cheaper than OpenAI Whisper — plus image generation at the industry's lowest price of $0.00013/step.
The comprehensive breakdown of Hanzo pricing across AI models, cloud compute, GPU access, and tools — with head-to-head comparisons against every major competitor.
Hanzo AI launches the industry's first zero-markup multi-provider AI gateway — one API key for 100+ models from every major provider, plus 14 proprietary Zen models.
Hanzo AI launches Zen4 Coder, a 480B-parameter Mixture of Experts code model that activates only 35B parameters per token — delivering frontier code intelligence at a fraction of the compute cost.
Egress fees are a tax on success. Hanzo Network includes bandwidth at every tier — from 500 GB on the $5 plan to 120 TB on Ultra — with zero overage charges.
Qwen3.5-35B-A3B (Apache 2.0, multimodal MoE) and Inception's Mercury 2 are now available on the Hanzo AI Gateway.
Google's Gemini 3.1 Pro is now available on the Hanzo AI Gateway — featuring a 1M-token context window and multimodal reasoning across text, images, audio, video, and code.
Announcing the full Zen4 family: mini (4B) through ultra (1T MoE), all abliterated. Eight models covering every scale from edge to cloud, with no refusal behavior and full capability access.
Claude Sonnet 4.6 and Qwen3.5-397B-A17B are now available on the Hanzo AI Gateway — Anthropic's latest alongside the largest open-source multimodal MoE model.
ByteDance's Seed 2.0 Pro and Alibaba's Qwen3-Coder-Next are now available on the Hanzo AI Gateway — frontier multimodal reasoning and open-source coding agents.
GLM-5 and MiniMax M2.5 — two of the largest open-source models ever released — are now available through the Hanzo AI Gateway.
Google's Gemma 3n (multimodal on-device) and Gemma 3 270M (smallest Gemma ever) are now available on the Hanzo AI Gateway.
Zen4 is a complete lineup of open AI models spanning from 4B to over 1 trillion parameters, featuring consumer, coder, and ultra tiers.
Anthropic's Claude Opus 4.6 and OpenAI's GPT-5.3 Codex are now available through the Hanzo AI Gateway — same API, same key, zero markup.
Bringing AI capabilities to your browser - summarize pages, answer questions, and automate workflows.
Introducing Hanzo Cap Table — equity management, scenario modeling, and 409A readiness without Carta's price tag.
Introducing Hanzo Dataroom — secure document sharing with per-page analytics, watermarking, and granular access controls for fundraising and M&A.
Announcing the unified Hanzo ML Platform — training, serving, pipelines, feature store, model registry, and evaluation in one integrated system.
Zen Reranker delivers 7680-dimensional neural reranking with a 31.87x BitDelta compression ratio, designed for RAG pipelines and integrated with Hanzo and Zoo decentralized search networks.
Hanzo Dev launches — an AI coding agent that lives in your terminal, understands your codebase, and writes production-quality code.
Zen VL is a family of vision-language models at 4B, 8B, and 30B -- each with instruct and agent variants -- supporting OCR in 32 languages, GUI navigation, spatial grounding, and native function calling with visual context.
Introducing Zen Artist - our multimodal image generation model with unprecedented control and quality.
Zen Omni is a 30B MoE unified multimodal model with Thinker-Talker architecture, handling text, vision, and audio in a single model with real-time speech-to-speech at under 300ms latency.
How we built 260+ tools using Model Context Protocol, enabling AI models to interact with the world.
Zen Designer is a 235B MoE vision-language model with 22B active parameters, supporting image analysis, video understanding, OCR in 32 languages, and native design reasoning.
Zen Max is a 671B MoE reasoning model with 384 experts, 256K context, and abliterated base weights -- achieving AIME 2025 99.1%, SWE-Bench 71.3%, and BrowseComp 60.2%.
Hanzo AI has driven over $1 billion in cumulative revenue for our clients through AI-powered marketing and commerce infrastructure.
Zen Coder is a family of code-specialized models spanning 4B edge to 480B frontier, with 128K context, extended thinking up to 512K tokens, MCP integration, and support for 100+ programming languages.
Introducing Hanzo Operative — enabling AI agents to use computers like humans: screen capture, mouse, keyboard, and application interaction.
Hanzo Flow launches — a visual drag-and-drop builder for complex AI workflows. Connect models, tools, and APIs without writing integration code.
Zen Guard is an 8B multilingual safety classifier covering 119 languages and 9 harm categories, with a three-tier severity system and 5ms/token streaming latency.
Hanzo Console launches — a unified development environment for observing, debugging, and optimizing LLM applications across all providers.
Zen Embedding reaches #1 on the MTEB multilingual leaderboard with a 70.58 score, supporting 100+ languages and Matryoshka representation learning across four dimension sizes.
Introducing Hanzo Experiments — feature flags, A/B testing, and progressive rollouts with statistical rigor.
Zen Nano is a 0.6B on-device AI model built for mobile and embedded deployment, achieving 44K tokens/sec on M3 Max with a 40K context window.
Introducing Zen Musician - our AI music generation model that creates original compositions in any style.
Announcing Jin 2.0, a major upgrade to our unified multimodal AI architecture with expanded capabilities and improved performance.
How Hanzo integrated AI infrastructure into the Keek social platform — content recommendation, multimodal moderation, and creator tools processing 12M daily content items.
Introducing ACI, a decentralized network for AI compute that enables trustless, verifiable AI inference.
Introducing Zen Video - our generative video model for creating high-quality video content from text and images.
Hanzo AI partners with Personas Social Inc. to integrate AI-powered features into the Keek social platform.
Introducing Hanzo Notebooks — managed Jupyter workspaces with on-demand GPU access and pre-installed ML frameworks.
Introducing our investment in Candle, the Rust ML framework powering Hanzo's next-generation inference.
Introducing Hanzo Sign — legally binding electronic signatures with audit trails, templates, and API access at a fraction of DocuSign's price.
Introducing Hanzo CX — a unified customer experience platform with omnichannel inbox, AI-assisted responses, and full conversation history.
Building Zen Coder - our code generation model trained on real software engineering workflows.
How we built Zen Dub - an AI system for voice cloning and automatic dubbing in 100+ languages.
How we built infrastructure to serve billions of LLM requests for commerce applications.
Introducing Hanzo Auto — AI-powered workflow automation with 280+ integrations and a visual builder.
Introducing Hanzo CRM — a modern, AI-enhanced CRM for teams that want to close deals, not manage software.
Introducing Zen Agent - our framework for building AI systems that can take action in the real world.
Introducing Zen, our system for ensuring AI follows business rules, compliance requirements, and brand guidelines.
Introducing Hanzo Insights — product analytics with event tracking, funnels, cohorts, and session replay, without selling user data.
Introducing Hanzo Edge — a high-performance API gateway with sub-millisecond routing, rate limiting, and authentication at the edge.
Introducing Gym - our platform for distributed AI model training that turns the world's GPUs into a supercomputer.
Introducing Jin, our unified architecture for multimodal AI that powers the next generation of Hanzo intelligence.
How we brought large language models to edge devices using llama.cpp and what we learned along the way.
Launching the Hanzo Observability stack — unified metrics, logs, and distributed tracing with OpenTelemetry-native ingestion.
Bringing AI capabilities to Vim without sacrificing the terminal-native experience.
Introducing Identity NFTs: blockchain-based portable identity for commerce.
Announcing Hanzo's open-source strategy — why we're open-sourcing our core infrastructure and introducing revenue sharing for contributors.
Launching HKE — managed Kubernetes that eliminates the operational burden without limiting what you can do.
Why we built a custom ML framework and what we learned along the way.
How Hanzo helped Bellabeat scale their women's health wearable platform to $35M in projected revenue through CMO/CTO services, marketing automation, and data analytics.
Introducing Hanzo Search — full-text search with typo tolerance, faceting, and sub-millisecond query times.
Introducing the Hanzo Agent SDK: tools for building intelligent commerce assistants.
Launching Hanzo Platform — a PaaS that deploys your code from git push to production with zero DevOps configuration.
How we are combining vision, language, and structured data for next-generation commerce AI.
Launching Hanzo Tasks and Queues — durable workflow execution and message queues for reliable background processing.
How Hanzo delivered a 500x ROI marketing campaign for Damon Motorcycles, generating a 230% increase in qualified leads for their electric motorcycles.
How we built a context management protocol for multi-model AI pipelines — early research that shaped how we think about AI state.
Introducing Hanzo KMS — end-to-end encrypted secrets management with environment sync, rotation, and audit trails.
Launching Hanzo Storage — S3-compatible object storage with integrated CDN and no egress fees.
How Hanzo helped Casper Labs architect and launch an enterprise-grade blockchain — from the Rust pivot to mainnet launch and DEVxDAO founding.
Lux Network achieves full EVM compatibility with the C-Chain, enabling seamless migration of Ethereum dApps.
Launching Hanzo SQL — managed PostgreSQL with automatic scaling, high availability, and zero operational overhead.
Announcing the Hanzo Agent Framework: building blocks for intelligent automation in commerce.
Introducing Hanzo Base — a complete backend-as-a-service with database, auth, storage, and realtime capabilities.
Introducing Hanzo Cloud: fully managed commerce infrastructure with enterprise-grade reliability.
Launching Hanzo DNS and Networking — programmable DNS with service discovery and zero-trust virtual networks.
Announcing Hanzo Webhooks: reliable, secure, real-time event delivery for your integrations.
How Hanzo helped Unikrn raise 120,000 ETH in their token launch, combining blockchain infrastructure with regulatory compliance across 15+ jurisdictions.
Announcing official SDKs for JavaScript, Python, Ruby, Go, and PHP, plus the new SDK development kit.
Introducing Hanzo Functions — serverless compute with zero cold starts, built for event-driven commerce workloads.
Introducing the Hanzo API Gateway: unified access, better security, and improved developer experience.
Hanzo co-founded the first SEC-approved crowdfunding token offering, bridging traditional crowdfunding with tokenized securities.
How we are using machine learning to transform analytics from passive reporting to active decision support.
Launching Hanzo Payments — unified payment processing with intelligent routing, fraud detection, and multi-currency support.
Announcing Checkout 2.0, our redesigned checkout experience built for maximum conversion.
Introducing Hanzo Identity — enterprise-grade authentication and access management built for commerce platforms.
Lessons from a year of building and deploying AI systems for commerce.
Reflecting on Techstars Demo Day and $42M in client sales generated through the Hanzo platform.
Hanzo has been selected for Techstars 2017, accelerating our mission to power intelligent commerce.
Crowdstart evolves into Hanzo: a complete commerce platform for the modern era.
How we built collaborative filtering and content-based recommendations into the Crowdstart platform.
Launching Hanzo Analytics — real-time behavioral intelligence for commerce teams who need answers now, not tomorrow.
Introducing Earle, our genetic algorithm system for evolving high-performing marketing campaigns.
How we built real-time analytics into Crowdstart and what we learned about data-driven commerce.
Releasing Astle.js, our open-source library for building reactive commerce interfaces.
How we built Crowdstart as a cloud-native platform from day one, and why this matters for commerce.
Announcing Crowdstart, a new platform enabling creators to build sustainable businesses from successful crowdfunding campaigns.
How we built a publish-subscribe system that became the nervous system of our entire platform.
Building a key-value store that could handle the chaos of distributed systems - our second foundational project.
The first building block of what would become Hanzo - a flexible datastore designed for the next generation of applications.