复用与扩散:面向文本到视频生成的迭代去噪
TLDR
提出VidRD框架,利用潜在扩散模型迭代生成额外视频帧,提升文本到视频生成的时间一致性。
评分理由
The paper presents a clear method for extending video frames using latent diffusion and includes quantitative and qualitative evaluations. However, the abstract lacks explicit discussion of limitations or comparisons to broader world model frameworks, and the connection to world modeling is indirect.
Read-first 评分解释
综合优先阅读分 35.9,由主题、引用、图谱、方法、可复现性和近期性等信号加权得到。 原始总分保留为 5。
研究版图角色
候选论文
排序敏感性
稳定性:volatile;排名波动范围:90。