AI Resonance · 阅读最新日报 · AI 入门推荐

2026-09-24 · America/Los_Angeles · 论文 · #4

HappyWorld-Bench

Evaluating world models requires assessing both the quality of the worlds they generate and their consistency and responsiveness under exploration, interaction, and modification. We introduce HappyWorld-Bench, a comprehensive benchmark that evaluates whether generated worlds remain reliable as agents interact with them. Our design is built on a hierarchical capability framework of six world capabilities (W1-W6), from generative construction to unified world modeling, instantiated across three independent evaluation tracks: video world models, spatial world models, and embodied world models.…

HappyWorld-Bench

热度 46.8 / 100;排名与评分保留该期记录。