Lighting-grounded Video Generation with Renderer-based Agent Reasoning
基于渲染的视频生成与基于代理的推理
机构 * State Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(北京大学计算机学院多媒体信息处理国家重点实验室) ; National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(北京大学计算机学院国家视觉技术工程研究中心) ; Beijing Academy of Artificial Intelligence(北京智源人工智能研究院) ; OpenBayes Information Technology Co., Ltd.(北京开贝信息技术有限公司) ; School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
专题命中 视频生成 :video generation(title,abstract);video diffusion(abstract);分类 cs.CV
AI总结 本文提出LiVER框架,通过显式3D场景属性条件生成可控视频,结合轻量条件模块和渐进训练策略,实现高保真和精确控制,适用于图像到视频和视频到视频的合成。
Comments Accepted to CVPR 2026