ZhiXiang Future Releases 200B+ Parameter Image Model HiDream-O1-Image-Pro, Accelerates Funding
ZhiXiang Future unveiled the HiDream-O1-Image-Pro, a 200B+ parameter native omnimodal image model based on the Unified Transformer architecture, achieving new SOTA across benchmarks. The company secured another round of hundred-million-level funding from investors including Shenzhen Capital Group and Jinpu Capital. It also introduced a 'Model + Agent' strategy with three agent products and signed partnerships with multiple enterprises, advancing toward a world model.
On May 19, 2026, ZhiXiang Future held its first open day in Beijing under the theme 'Imaging the World,' where it officially launched the HiDream-O1-Image-Pro, a closed-source image model with over 200 billion parameters. Built on the company's proprietary Unified Transformer (UiT) architecture, this native omnimodal model integrates raw image pixels, discrete text tokens, and task conditions into a shared continuous token space, enabling deep fusion of image, text, and multimodal tasks. The model achieves state-of-the-art performance across multiple benchmarks, including general text-to-image generation, high-fidelity text rendering, diverse scene generation, and image editing.
Alongside the model release, ZhiXiang Future announced a new round of financing amounting to hundreds of millions of yuan, led by Shenzhen Capital Group, Jinpu Capital, Caixin Capital, and Fuju Capital. This round follows another major funding announcement just two weeks earlier, bringing in a diverse group of investors including Anhui Provincial Investment, Hefei Investment, and Oriental Fortune Capital. The rapid fundraising reflects strong market confidence in the native omnimodal AI direction.
ZhiXiang Future's strategy revolves around a 'Model + Agent' dual-engine approach. The company's '1+1+3' business architecture consists of the HiDream model series as the base, the HiHarness enterprise service platform as the middleware, and three agent products targeting specific verticals: HiBurst for business marketing (supporting TikTok, Meta, Douyin, etc.), Zhenzan for professional AI film production (over 5,000 minutes of short animations produced), and vivago for social media creation (over 40 million users across 100+ countries).
During the open day, ZhiXiang Future also signed strategic cooperation agreements with Shanghai Film Group's New Vision Fund, BlueFocus, Beijing Jetsen Technology, and Beier Health, focusing on deep collaboration in model capabilities, agent development, and industry scenarios including film, marketing, cross-border e-commerce, IP operations, and healthcare.
According to CEO Mei Tao, the company chose the native omnimodal path because it integrates physical laws, spatial relationships, and causal logic from the start, enabling the model to understand, reason, and reconstruct the world rather than merely generate content. CTO Yao Ting emphasized that the UiT architecture allows 'Any to Any' generation, where any input modality can produce any output modality, a key capability for building a world model. ZhiXiang Future aims to evolve from visual generation to world modeling, unifying understanding, generation, and prediction of real-world states.