V-ReasonBench:面向视频生成模型的统一推理基准套件
TLDR
提出V-ReasonBench基准,用于评估生成模型在四个维度上的视频推理能力,使用合成和真实序列。
评分理由
The paper presents a well-structured benchmark with clear dimensions and evaluation of six models, but its focus is on reasoning evaluation rather than world model development. Strengths include reproducibility and real-world data; weakness is limited novelty in world model concepts.
Read-first 评分解释
综合优先阅读分 30.1,由主题、引用、图谱、方法、可复现性和近期性等信号加权得到。 原始总分保留为 9。
研究版图角色
前沿论文
排序敏感性
稳定性:volatile;排名波动范围:88。