SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations
SceneScribe-1M:一个具有全面几何和语义标注的大规模视频数据集
机构 * Shanghai Jiao Tong University(上海交通大学) ; Ant Group(蚂蚁集团) ; Visual Geometry Group, University of Oxford(牛津大学视觉几何组) ; Ningbo Institute of Digital Twin, Eastern Institute of Technology, Ningbo(宁波数字孪生研究院,东部技术研究院,宁波) ; Zhejiang Key Laboratory of Industrial Intelligence and Digital Twin(浙江工业智能与数字孪生重点实验室)
专题命中 视频数据与评测 :video generation(abstract);text-to-video(abstract);分类 cs.CV
AI总结 SceneScribe-1M通过提供一个包含一百万真实视频的数据集,支持3D几何感知与视频合成的统一资源,推动了单目深度估计、场景重建等下游任务及文本到视频合成等生成任务的发展。
Comments Accepted by CVPR 2026