HiDream.ai Unveils World's First Native All-Modal Interactive World Model HiDream-O1-World, Tops WBench
HiDream.ai today unveiled the world's first native all-modal interactive world model, HiDream-O1-World, built on its proprietary UiT architecture. The model supports multimodal inputs including text, images, and interactions, and can generate dynamic worlds with long-term spatiotemporal and physical consistency. In the WBench benchmark jointly launched by Meituan LongCat and Fudan University, HiDream-O1-World topped the Navi sub-leaderboard on its first attempt, scoring 73.3 in physical dimensions and 88.0 in consistency, ranking first overall. This launch marks a shift in AI content generation from one-way output to explorable and interactive world models.