Overview
HiDream-O1-Image-Pro is an image large model released by Zhixiang Future on May 19, 2026, at the first Technology Open Day 'Imaging the World'. Built on the next-generation **native full-modal model architecture Unified Transformer (UiT)**, with over **200 billion** parameters, it has **set new SOTA records** on multiple benchmarks.
This is a key step for Zhixiang Future from 'visual generation' to 'world model'—by unifying image pixels, text tokens, and task conditions into a continuous shared token space, the model can handle tasks such as general text-to-image, high-fidelity text rendering, and image editing in a unified manner. At the same time, Zhixiang Future announced the completion of a new round of hundreds of millions in financing (with participation from Shenzhen Capital, Jinpu Investment, Caixin Capital, Fuju Capital, etc.), two rounds in half a month, showing strong capital market recognition of the native full-modal direction.
According to the latest official website information, HiDream.ai has expanded into a globally leading AI platform supporting four modalities: text, image, video, and 3D model, and has launched the HiHarness enterprise-level multimodal AI platform and the HiBurst AIGC marketing tool.
Key Features
- Over 200 Billion Parameters: The parameter scale reaches the leading level of domestic image large models, with model capacity ensuring generation quality
- Unified Transformer (UiT) Architecture: Native full-modal architecture that unifies image pixels, text tokens, and task conditions into a shared token space
- Multi-Benchmark SOTA: Sets industry-leading records on multiple benchmarks including general text-to-image, high-fidelity text rendering, and image editing
- High-Fidelity Text Rendering: Industry-leading clarity and accuracy of Chinese and English text in images, suitable for posters, covers, and infographics
- Powerful Image Editing: Supports object replacement, style transfer, local modification, and other image editing capabilities
- Full-Modal Extension Capability: Prepares for unified modeling of images, video, text, and audio, representing an exploration toward the world model direction
- Open-Source Version Available: The 8B open-source version HiDream-O1-Image with the same architecture performs excellently on the Artificial Analysis text-to-image leaderboard
- Multimodal Generation Expansion: The platform now supports video and 3D model generation, covering four modalities
- Enterprise Platform HiHarness: Provides over 200 API services, supporting industry solutions and private MaaS deployment
- Marketing Tool HiBurst: AIGC marketing tool supporting viral script generation, storyboard control, and one-click publishing
Use Cases
- Design of posters, covers, and infographics requiring high-fidelity Chinese text rendering
- Batch generation of e-commerce product images and marketing materials
- Professional scenarios such as image editing, object replacement, and style transfer
- Enterprise users with needs for domestic AI images
- Academic research and prototype exploration (open-source version can be deployed locally)
- Content creators seeking domestic alternatives to Nano Banana / GPT-Image-2
- Video and 3D model generation (new platform features)
- Enterprise-level AI integration (via HiHarness platform)
- Rapid generation of marketing content (via HiBurst tool)
Pros
- New benchmark for domestic AI image large models (200 billion parameters)
- Native full-modal architecture with leading technical path (not just a dedicated 'text-to-image' model)
- Multi-benchmark SOTA with third-party validation of results
- Strong high-fidelity text rendering capability, especially suitable for Chinese scenarios
- 8B open-source version available for commercial use, lowering the barrier for domestic AI deployment
- Fast company financing pace (two rounds in half a month), with both technology and capital driving growth
- Evolution toward 'world model' with significant future potential
- Platform expanded to video and 3D model generation, covering four modalities
- Launch of enterprise platform HiHarness and marketing tool HiBurst, expanding application scenarios
Pricing
The Pro version is available through Zhixiang Future's official platform and API, with specific pricing subject to official announcements. The 8B open-source version HiDream-O1-Image with the same architecture has been open-sourced and can be downloaded for free for local deployment. Enterprise-level cooperation and private deployment can be negotiated separately through the HiHarness platform.
Summary
HiDream-O1-Image-Pro represents one of the highest levels of domestic AI image generation in 2026. In a field dominated by overseas models like Nano Banana / GPT-Image-2 / Imagen 4 / Midjourney, domestic models have found a differentiated breakthrough with 'native full-modal architecture + 200 billion parameters + multi-benchmark SOTA'. Zhixiang Future's technical path from 'visual generation' to 'world model' gives it a unique strategic position in the domestic AI image direction. According to the latest official website information, HiDream.ai has expanded into a globally leading AI platform supporting four modalities: text, image, video, and 3D model, and has launched the HiHarness enterprise platform and HiBurst marketing tool, further expanding application scenarios. For users seeking the latest SOTA results, high-fidelity Chinese text rendering, and domestic models, HiDream-O1-Image-Pro is one of the most noteworthy new products.