VideoPhy
Dataset Analysis
A benchmark to evaluate if text-to-video models follow physical commonsense for real-world activities, finding current models severely lacking.
Provenance
Collected from papers.
Derived from paper: VideoPhy: Evaluating Physical Commonsense for Video Generation