One Video, One World: Turning Monocular Video into Physical 4D Scenes
TLDR
OVOW reconstructs instance-level, simulation-ready 4D mesh scenes from a single monocular video without training, using a four-stage pipeline.
Reasoning
The paper presents a novel training-free pipeline for 4D reconstruction with instance separation and physical plausibility, which is a strength. However, evaluation is limited to synthetic benchmarks, and the core contribution is about reconstruction rather than world models, making relevance to the given keywords low.
Read-first score
Read-first score 23.4, weighted from topical fit, citation, graph, method, reproducibility, and recency signals. Original total remains 4.
Field roles
Frontier
Rank sensitivity
Stability: volatile; rank range: 43.