SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency
TLDR
Stable Video 4D introduces a unified latent video diffusion model to generate multi-view consistent novel view videos from monocular video, enabling efficient dynamic NeRF 4D generation.
Reasoning
The paper proposes a unified diffusion model for multi-frame and multi-view consistent dynamic 3D generation, with strong empirical results and user studies. However, the abstract relies on the Objaverse dataset and does not clarify whether real-world dynamic scenes are included, limiting evidence of generalizability.
Read-first score
Read-first score 28.8, weighted from topical fit, citation, graph, method, reproducibility, and recency signals. Original total remains 1.
Field roles
Candidate
Rank sensitivity
Stability: volatile; rank range: 80.