Google AI
The Times Australia

Times Media

HiDream Unveils HiDream-O1-Video-1.0, a Native Omnimodal Video Model Built for Physical Consistency

The model supports multimodal inputs including text, images and video, and generates high-fidelity 1080p videos of 5 to 20 seconds with natively synchronized audio

BEIJING, CHINA – Media OutReach Newswire – 17 September 2026 – HiDream.ai, an AI company specializing in foundation models and generative AI, announced the launch of HiDream-O1-Video-1.0, or HiDream V1, its native omnimodal video generation model.

HiDream Unveils HiDream-O1-Video-1.0, a Native Omnimodal Video Model Built for Physical Consistency

Designed around a deeper understanding of creative intent and real-world physics, HiDream V1 supports multimodal inputs including text, images and video. It can generate high-fidelity 1080p videos ranging from 5 to 20 seconds, with enhanced narrative planning, character consistency, physical plausibility and audiovisual synchronization.

In its debut on two independent international benchmarks, HiDream V1 ranked No. 4 globally on the Artificial Analysis Image to Video Leaderboard (With Audio) and No. 8 on the Arena.ai Image-to-Video leaderboard, placing it among the world's leading video generation models.

A Strong Debut on Two International AI Benchmarks
Artificial Analysis independently evaluates leading AI models through standardized and reproducible benchmark testing. Arena.ai uses anonymous head-to-head comparisons and user voting to assess model outputs. Together, the two platforms provide third-party perspectives on model performance based on both standardized testing and real-world user preferences.

HiDream V1's results demonstrate its competitiveness across key dimensions including visual quality, prompt adherence, motion quality, narrative coherence and audiovisual coordination.

The global AI video generation market has become increasingly competitive, with models developed by Chinese teams—including Seedance and MiniMax H3—maintaining strong positions on international leaderboards. HiDream V1's top-tier debut further expands China's presence among the world's leading video generation models.

"The next generation of video models will not be defined solely by higher resolution or longer duration," said Yao Ting, Chief Technology Officer of HiDream.ai. "What matters is whether a model can genuinely understand a creator's intent and how objects, actions and sounds interact in the real world. HiDream V1 was designed from the outset to represent text, video and audio within a unified framework. Our goal is to move video generation beyond simply looking sharp and moving smoothly toward understanding instructions, sustaining coherent performances and keeping sound aligned with visuals. Its performance on two independent international benchmarks provides encouraging validation of our native omnimodal approach."

Upgrades in Visual Fidelity, Narrative Coherence and Character Consistency
HiDream V1 introduces broad improvements in high-fidelity rendering, narrative continuity and character consistency. Rather than optimizing only for the quality of individual frames, the model is designed to maintain coherence across an entire video sequence.

Once characters, emotions, motivations and physical rules are placed within a continuous sequence, each element can affect the others. HiDream therefore uses a technical framework built around three stages: planning first, followed by joint generation, and then alignment through multimodal reward signals.

Under this approach, the model first plans the narrative and character states at a global level. It then jointly constrains visuals, movement and semantics during generation, before using multimodal reward signals to align visual quality, continuity and physical plausibility.

Native Omnimodal Architecture for Better Intent and Physics Understanding
The central challenge in AI video generation is shifting from the quality of individual frames to a broader understanding of complex instructions, continuous motion and the rules governing the physical world.

User prompts often contain multiple layers of information, including characters, actions, settings, camera directions, dialogue and sound. To handle these requirements more effectively, HiDream V1 incorporates multimodal intent understanding and planning. Before generation begins, the model structures the user's request across elements such as shot duration, setting, character state, movement, facial expression, composition, camera motion, dialogue and ambient sound.

Through this "understand, plan and generate" workflow, HiDream V1 plans narrative development and character states at a global level, while coordinating visuals, motion, semantics and audio during generation. This is designed to improve content completeness, character consistency and continuity between shots.

Physical reasoning is another core capability of HiDream V1. The model incorporates factors such as gravity, inertia, collisions, deformation, materials, lighting and spatial continuity into the generation process. As a result, object movement, character actions and environmental responses can more closely reflect the behavior of the real world.

This approach moves video generation beyond reproducing visual appearances and toward modeling how a dynamic world operates over time.

Content-Adaptive Video Duration
Most current video generation models require users to specify a fixed duration in advance, forcing the content to fit within a predetermined time window. This can lead to unnecessary pauses after an action has ended or the use of slow motion to fill the remaining time.

HiDream V1 instead incorporates duration into its narrative planning process. Within a range of 5 to 20 seconds, the model can determine an appropriate video length based on how an event unfolds, how long an action takes to complete and the pacing required by the content.

The capability is not simply a matter of adding or removing frames. It is designed to make duration serve the narrative, producing more natural pacing and more complete sequences. Users can focus on describing the event and intended outcome, while the model plans the corresponding temporal structure.

Joint Modeling of Text, Video and Audio
Conventional AI video generation systems often follow a staged workflow in which visuals are generated first and sound is added afterward. This can lead to mismatches between lip movements, physical actions, ambient sound and sound effects.

HiDream V1 uses a native omnimodal architecture to jointly model text, video and audio signals. The three modalities constrain and inform one another within a single generation process. Visual motion influences sound timing, dialogue and sound effects contribute to the emotional tone of a scene, and text conditions guide the generation process throughout.

This unified approach is designed to improve synchronization across character performance, environmental changes and physical actions.

During post-training, HiDream uses Diffusion Reinforcement Learning and a multimodal reward model aligned with human perception and aesthetic preferences. Generated content is evaluated across multiple dimensions, including visual quality, semantics, motion, physical plausibility and sound, helping improve overall consistency and realism.

Completing HiDream's Native Omnimodal World Model Portfolio
HiDream V1 is not a standalone model. It is a core component of HiDream's native omnimodal world model strategy.

Built on a unified UiT, or Unified Transformer, foundation, the company's portfolio now comprises four major model families:

  • HiDream-O1-Image, for image understanding and generation;
  • HiDream-O1-Video, for dynamic content generation and temporal storytelling;
  • HiDream-O1-World, for 3D environment simulation and real-time interaction; and
  • HiDream-O1-Embodied, for spatial reasoning, action planning and feedback in embodied AI.
Rather than operating as four separate technology stacks, the models are designed to evolve from the same unified architecture. Together, they cover spatial understanding, temporal modeling, 3D interaction and embodied action, creating an integrated system for perceiving, understanding, generating, interacting with and acting upon the physical world.

HiDream's models have previously achieved leading results on several independent benchmarks. HiDream-O1-Image 1.5, the commercial version of the company's image model, ranked No. 2 globally and No. 1 among Chinese models on the Artificial Analysis text-to-image leaderboard, while its open-source edition also reached No. 2. HiDream-O1-World ranked No. 1 overall on WBench, a benchmark for interactive world models. HiDream-O1-Embodied achieved first-place results in the disturbance adaptation and spatial reasoning categories of RoboColiseum.

"We are not building four isolated model product lines," said Dr. Mei Tao, founder and CEO of HiDream.ai. "We are building a native omnimodal system that shares a unified technical foundation and continues to evolve toward a deeper understanding of the real world. Competition among the next generation of foundation models will depend not only on improvements in individual capabilities, but also on whether different modalities can understand, generate and interact within a unified architecture."

From understanding and generating the world to simulating and acting within it, the launch of HiDream V1 completes a key part of HiDream's native omnimodal world model portfolio.

HiDream Completes Series C+ Financing
Alongside the launch of HiDream V1, HiDream announced the completion of a Series C+ financing round backed by New Micro Capital, Jiaozi Capital and ICBC Capital. The proceeds will support native omnimodal model research and development, product iteration, AI computing infrastructure and the expansion of industry applications.

New Micro Capital was jointly established by Shanghai New Micro Technology Group, Shanghai Science and Technology Venture Capital Group and its management team. The firm manages nearly RMB 10 billion in assets.

As the corporate venture capital arm of New Micro Technology Group, New Micro Capital is expanding its investments in artificial intelligence. Drawing on the group's industrial capabilities in specialty integrated-circuit manufacturing, it seeks to advance collaboration across AI infrastructure, model technologies and end-user applications.

Jiaozi Capital is a wholly owned state-backed equity investment platform under Chengdu Jiaozi Financial Holding Group. The firm focuses on artificial intelligence and other strategic technology sectors, using long-term capital and industrial resources to accelerate technology commercialization and ecosystem development.

With the launch of HiDream V1, HiDream's native omnimodal world model strategy is moving from a technical roadmap toward an integrated system supported by a unified foundation, a coordinated model portfolio and deployable product capabilities.

Availability
HiDream V1 is currently in internal testing and is expected to become available in the near future. Users can learn more at: https://hiharness.ai/

Hashtag: #HiDreamAI

The issuer is solely responsible for the content of this announcement.

About HiDream.ai

HiDream.ai is an artificial intelligence company focused on generative AI and native omnimodal world models. Built on its unified UiT architecture, the company's model portfolio spans image generation, video generation, interactive world simulation and embodied intelligence. HiDream aims to develop AI systems capable of understanding, generating, simulating and acting upon the real world.

Leaderboard disclaimer: Rankings cited in this release reflect publicly available results at the time of publication and may change as participating models, evaluation data or platform methodologies are updated.

Read more: HiDream Unveils HiDream-O1-Video-1.0, a Native Omnimodal Video Model Built for Physical Consistency

More Articles …

  1. Hungry for Choice: 85% of Singaporeans Fear Rising Costs as Food Delivery Shrinks to Two-Player Market - New Blackbox Study
  2. VX Logistics Transforms Fresh Fruit and Vegetables Industry through Application of AI and World-first Robot
  3. SWISS REJU Wins "Exquisite Body-Sculpting Reputable Brand Award" from HK01, Following Exceptional User Reviews and Recognition
  4. kapok Introduces AW26 Collection, Discovering Niche Design in Hong Kong’s Hidden Cultural Gem
  5. Semperis Secures IMDA Accreditation to Strengthen Identity Resilience in the AI Era
  6. Dah Sing Bank Releases the 2027 Investor Confidence Index
  7. ANN heritage e-book brings Asia closer through a digital celebration of ancient legacies, modern innovation, and shared identity
  8. Oriental Residence Bangkok: When Needs Are Understood Without a Word, Care Feels Effortlessly Natural
  9. DHL Express sees rise in heavyweight shipments in Asia Pacific as businesses circumvent unpredictability
  10. Tenchijin’s "KnoWaterleak" Surpasses 100 Cumulative Municipal Contracts in Japan
  11. Media OutReach Newswire PRISM Summit 2026 Champions Trust and Visibility for the Future of PR in the Age of AI
  12. LHN Energy to Roll Out Integrated Energy Infrastructure Solution Combining Solar PPA and DC Fast Charging
  13. Singapore and Korean Clinicians Call for More Individualised, Evidence-Based Postnatal Care as Women Enter Motherhood Later in Life
  14. Fastmarkets Chooses TMX Trayport as Technology Partner for Lithium Market Infrastructure
  15. SERV, Swiss Export Risk Insurance, Supports Guarantee of USD 212.5 Million for Capex of First Phosphate Mine Project in Quebec, Canada
  16. Arrow Electronics to showcase intelligent and sustainable technology solutions at electronica India 2026
  17. Tackling the 40% IV failure rate: CUHK students win the James Dyson Award 2026 with an affordable AI vein visualizer
  18. Response to the 2026 Policy Address and the Five-Year Plan (2026-2030) by Cushman & Wakefield
  19. Bupa Hong Kong Welcomes Hong Kong's First Five-Year Plan, Supporting Prevention, Primary Healthcare and Healthy Ageing to Build a Healthier Hong Kong
  20. Lever Style Announces the Sale of Champion System to ONDO
  21. Hang Lung Opens Xi Zhe Wuxi, Curio Collection by Hilton at Center 66, Further Tapping Wuxi’s Premium Clientele
  22. PHOENIX STAR 2026 CHINA CITY HONORS Unveiled in Guangzhou, Highlighting the Distinctive Value of Chinese City Brands
  23. Hong Kong's First Five-Year Plan sets out long-term vision for development
  24. Live from IAA Transportation 2026: JAC Motors Responds to Europe's Green Logistics Needs with Four Models
  25. Hong Kong’s 2026 Policy Address: A Strategic Vision for a Bright New Era
  26. The Carnival Fair Reaches 1,200 Events Milestone Across Singapore
  27. China's development philosophy resonates across Global South at BRICS Summit
  28. The Longines Hong Kong International Horse Show Returns to the AsiaWorld-Expo in January 2027, Cementing Its Status as Asia’s Ultimate Lifestyle and Sporting Destination
  29. ONYX Hospitality Group Strengthens Sustainable Hotel Operations Through Strategic Partnership with SP Group
  30. Flam Raises US$40 million (S$51 Million) in Series B Funding to Scale AI-Powered Interactive Content
  31. Skyro Unveils Refreshed Visual Identity as It Builds a Global Financial Ecosystem
  32. TAT Celebrates the Success of the Amazing Thailand Ambassador Campaign, Sparking a Viral "Feel All the feelings" Phenomenon with Over 448 Million Impressions
  33. VEC and dmg events Announce Strategic Partnership, Opening a New Chapter for the Vietnam International Industrial Fair (VIIF)
  34. The Iranian People Receive the 2026 Freedom Prize
  35. Xsolla Launches AI Toolkit, Enabling Engine-Agnostic Game Commerce Setup With AI Coding Tools
  36. TDCX Opens Lisbon Campus as Its Next European Growth Engine
  37. Over half of APAC businesses want greater control over strategic data in latest digital sovereignty push, new research shows
  38. WRISE Prestige Marks Two Years in Hong Kong with Strong Growth and Unveils the Evolution of the Intelligent Wealth Consultant
  39. The Makeover Guys and RHB Bank Introduce the Makeover Flexi Loan with Up to 120% Financing for Home Purchase and Renovation
  40. Thailand invites global visitors to Udon Thani International Horticultural Expo 2026
  41. FPG Fortune Prime Global Partners with AC Milan to Strengthen Its Global Presence
  42. "Rumah Kita, Ceria Bersama" — Easyhome Mall Joins Hands with the Community to Celebrate Malaysia Day
  43. Fuel Incorporation Facilitates Over US$1 Billion in Gasoil Transactions as Diesel Stockpiling Surges — and Expands Into Coal With Indonesian Mine Supply
  44. HKUST Welcomes the Prime Minister of Uzbekistan Strengthening Partnerships in Education, Research, and Talent Development
  45. Surge Announces 99.95% Battery-Grade Lithium Carbonate from Nevada North
  46. November 20-21 "Galaxy Stars of Gastronomy" Presents Michelin-starred Mastery: 12 Chefs | 24 Glittering Stars
  47. Hao Yi Tou's Tang Plaza Outlet Marks One Year of Business Class Reflexology
  48. Award-Winning WRITERS AT WORK Refines Signature Writing Programme for 2027 to Develop Future-Ready Thinkers
  49. Regent Hong Kong Celebrates Four Decades of Cantonese Culinary Excellence as Lai Ching Heen Preserves Its Legacy for Future Generations
  50. Azbil to Showcase Building Automation Solutions as a Gold Sponsor at Data Centre World Asia 2026

Times Magazine

Australia Needs Permission Budgets for AI Agents

The Times recently argued that Australia should keep building the data centres the AI economy requ...

Nikon Is Developing Nine New Cinema Lenses and a Major ZR Firmware Update

Nikon is developing a new series of nine NIKKOR Z CINEMA T1.9 VV cinema lenses, designed for cinem...

The Hormuz conflict enters a more dangerous phase — and Australia will pay for every voyage

The conflict around Iran and the Strait of Hormuz has entered a more dangerous phase, with the Uni...

Technology

Australia Needs Permission Budgets …

The Times recently argued that Australia should keep building the data centres the AI economy requ...

Local News

Fitstop Global Games to Bring 1,000…

The Australian-born fitness brand is bringing its global competition home, with athletes from across...

Culture

Two Years Later — a love story that quietly a…

Two Years Later looks, initially, like a romantic television series. It is considerably more inte...

Travel

School holiday pricing: fair market economics…

Every Australian family with school-aged children knows the pattern. Look at an airfare, hotel ro...

The Times Features

Two Years Later — a love story that quietly asks what r…

Two Years Later looks, initially, like a romantic television series. It is considerably more inte...

Free Family Fun Day Brings K-Pop, Face Painting and Sch…

K-Pop Demon Hunters kids disco and free face painting headline a day of activities for local familie...

Four Free Family Days Put the Fun Back Into School Holi…

Batmobile, Bluey & Bingo, a reptile show and a petting zoo across the September/October break ...