Cavia:可控相机的多视角视频扩散与视图集成注意力
In recent years, there have been remarkable breakthroughs in image-to-video generation. However, the 3D consistency and camera controllability of generated frames have remained unsolved. Recent...
近年来,图像到视频生成取得显著进展,但3D一致性和相机可控性问题仍未解决。为此,我们提出了Cavia框架,能够将输入图像转换为多个时空一致的视频,支持精确控制相机运动,同时保持物体运动。实验结果表明,Cavia在几何一致性和感知质量上优于现有方法。
