arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

世界模型

面向环境建模、时序预测、仿真规划、具身智能和自动驾驶的世界模型方法与应用。

共收录 4308 信号源:cs.AI, cs.LG, cs.CV, cs.RO, cs.MA

1. 通用世界模型 4308 篇

1111.1106 2011-11-07 hep-th hep-ph 83%

Fermion mass hierarchy in a multiple warped braneworld model

R. S. Hundi, Soumitra SenGupta

专题命中 通用世界模型 :world model(title);world model(title)

Comments 15 Pages, Latex

详情

展开后加载摘要…

URL PDF HTML 收藏
0812.2545 2011-07-08 gr-qc 83%

Features of galactic halo in a brane world model and observational constraints

K. K. Nandi, A. I. Filippov, F. Rahaman, Saibal Ray, A. A. Usmani, M. Kalam, A. DeBenedictis

专题命中 通用世界模型 :world model(title);world model(title)

Comments 17 pages. Accepted for publication in Monthly Notices of the Royal Society; Figures available at http://www.sfu.ca/~adebened/research/brane_gal_rot_curves/

Journal ref Mon.Not.Roy.Astron.Soc.399:2079-2087,2009

详情

展开后加载摘要…

URL PDF HTML 收藏
1105.1064 2011-06-24 hep-ph hep-th 83%

Bulk Higgs and Gauge fields in a multiply warped braneworld model

Ashmita Das, R. S. Hundi, Soumitra SenGupta

专题命中 通用世界模型 :world model(title);world model(title)

Comments 15 Pages, Latex

Journal ref Phys.Rev.D83:116003,2011

详情

展开后加载摘要…

URL PDF HTML 收藏
1010.6079 2010-11-09 hep-th hep-ph 83%

Hierarchy problem and the cosmological constant in a five-dimensional Brans-Dicke brane world model

Mikhail N. Smolyakov

专题命中 通用世界模型 :world model(title);world model(title)

Comments 11 pages

Journal ref Gen.Rel.Grav.42:2799-2811,2010

详情

展开后加载摘要…

URL PDF HTML 收藏
0910.4766 2009-12-01 hep-th 83%

Doubling of background solution in 5D stabilized brane world model

Mikhail N. Smolyakov

专题命中 通用世界模型 :world model(title);world model(title)

Comments 7 pages, LaTeX, typos corrected

Journal ref JHEP 0911:077, 2009

详情

展开后加载摘要…

URL PDF HTML 收藏
gr-qc/0401049 2009-12-01 gr-qc hep-th 83%

Vacuum solutions of the gravitational field equations in the brane world model

T. Harko, M. K. Mak

专题命中 通用世界模型 :world model(title);world model(title)

Comments 13 pages, no figures, to appear in PRD

Journal ref Phys.Rev. D69 (2004) 064020

详情

展开后加载摘要…

URL PDF HTML 收藏
0804.4484 2009-12-01 gr-qc astro-ph hep-th 83%

Phantom-like behaviour in a brane-world model with curvature effects

Mariam Bouhmadi-Lopez, Paulo Vargas Moniz

专题命中 通用世界模型 :world model(title);world model(title)

Comments 14 pages, 7 figures. RevTeX 4. Discussion expanded, new appendix and references added. Version to appear in PRD

Journal ref Phys.Rev.D78:084019,2008

详情

展开后加载摘要…

URL PDF HTML 收藏
0708.1460 2009-12-01 gr-qc astro-ph hep-th 83%

Warm inflation in the DGP brane-world model

Sergio del Campo, Ramon Herrera

专题命中 通用世界模型 :world model(title);world model(title)

Comments 15 pages and 1 figure. Accepted for publication in Phys. Lett. B

Journal ref Phys.Lett.B653:122-128,2007

详情

展开后加载摘要…

URL PDF HTML 收藏
hep-th/0409083 2009-12-01 hep-th gr-qc hep-ph 83%

Quantum dynamics of particles in a discrete two-branes world model: Can matter particles exchange occur between branes?

Michael Sarrazin, Fabrice Petit

专题命中 通用世界模型 :world model(title);world model(title)

Comments 11 pages, no figures. Final version. Published in Acta Physica Polonica B

Journal ref Acta Phys.Polon. B36 (2005) 1933-1950

详情

展开后加载摘要…

URL PDF HTML 收藏
hep-th/0104049 2009-11-30 hep-th gr-qc hep-ph 83%

Second Order Perturbations in the Randall-Sundrum Infinite Brane-World Model

Hideaki Kudoh, Takahiro Tanaka

专题命中 通用世界模型 :world model(title);world model(title)

Comments 15 pages, 2 figures, comment and reference added, typos corrected

Journal ref Phys.Rev. D64 (2001) 084022

详情

展开后加载摘要…

URL PDF HTML 收藏
astro-ph/9907080 2009-11-30 astro-ph gr-qc 83%

A conserved variable in the perturbed hydrodynamic world model

J. Hwang

专题命中 通用世界模型 :world model(title);world model(title)

Comments 4 pages, no figure, To appear in Phys. Rev. D

Journal ref Phys.Rev. D60 (1999) 103512

详情

展开后加载摘要…

URL PDF HTML 收藏
gr-qc/0205002 2009-11-30 gr-qc hep-th 83%

Rigidity theorems in the braneworld model

Hideo Kodama

专题命中 通用世界模型 :world model(title);world model(title)

Comments 7 pages in LaTeX with the PTP style. No figure. To be published in the proceedings of the workshop "Braneworld -- Dynamics of spacetime with boundary --"

Journal ref Prog.Theor.Phys.Suppl.148:277-283,2003

详情

展开后加载摘要…

URL PDF HTML 收藏
hep-th/0104177 2009-11-30 hep-th gr-qc hep-ph 83%

A 6-D Brane World Model

Panagiota Kanti, Richard Madden, Keith A. Olive

专题命中 通用世界模型 :world model(title);world model(title)

Comments 22 pages, LaTeX file, no figures

Journal ref Phys.Rev.D64:044021,2001

详情

展开后加载摘要…

URL PDF HTML 收藏
hep-th/0110273 2009-11-30 hep-th 83%

Thick brane world model from perfect fluid

V. D. Ivashchuk, V. N. Melnikov

专题命中 通用世界模型 :world model(title);world model(title)

Comments 11 pages, Latex, to be published in Grav. Cosmol. 7, 241-245 (2001)

Journal ref Grav.Cosmol. 7 (2001) 241-245

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.01372 2026-08-19 cs.LG cs.AI cs.CV 版本更新 83%

BRo-JEPA: Learning Modular Transformations in Latent Space

BRo-JEPA:在潜空间中学习模算术

Divyansh Jha, Yuanfang Xie, Brennen Yu, Varan Mehra

机构 * Georgia Institute of Technology(佐治亚理工学院) NYU Langone Health(纽约大学Langone医疗中心)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出BRo-JEPA模型,通过在潜空间中施加模10算术的循环结构,实现零样本泛化,解决了标准模型无法外推未见操作的问题。

Comments 20 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.14530 2026-08-17 cs.CV cs.AI 新提交 83%

Marionette: Predicting World States, Rendering Geometry, Painting Appearance

Marionette:预测世界状态、渲染几何、绘制外观

Zian Meng, Zhen Li, Chuanhao Li, Qiang Li, Kaipeng Zhang

机构 * Alaya Lab(Alaya实验室) Shanghai Innovation Institute(上海创新研究院) Huazhong University of Science and Technology(华中科技大学)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 Marionette是面向带关节角色的交互式游戏的世界模型,通过显式建模3D世界状态、委托几何计算、合成外观,实现了可控的长时序行为,且外观生成无明显保真度损失。

Comments Project page: https://alayalab.github.io/Marionette/

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.01049 2026-08-04 cs.AI cs.CV cs.LG 新提交 83%

FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds

FactorJEPA:将整体未来分解为布局-智能体-交互通道以应对拥挤混乱的全球南方城市场景

Kapil Wanaskar, Gaytri Jena, Aman Chadha, Vinija Jain, Vasu Sharma, Amitava Das

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 针对全球南方拥挤混乱的城市场景,该研究构建了首个大规模数据集DENSEWORLD-115k,并提出FactorJEPA模型,通过分解结构提升JEPA性能,相关结果在不同主干网络上具有一致性。

Comments 42 pages. Independent preprint. Dataset: https://huggingface.co/datasets/anonymousML123/denseworld-115k ; checkpoints: https://huggingface.co/datasets/anonymousML123/factorjepa-outputs

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26217 2026-06-26 cs.LG cs.CV cs.RO 新提交 83%

Fast LeWorldModel

快速LeWorldModel

Yuntian Gao, Xiangyu Xu

机构 * Xi’an Jiaotong University(西安交通大学)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 提出Fast-LeWM,通过动作前缀并行预测未来潜在状态,替代LeWM的自回归滚动,降低规划时间并减少潜在误差累积。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03485 2026-06-17 cs.CV cs.AI cs.RO 版本更新 83%

Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion

Phys4D: 从视频扩散模型实现细粒度物理一致的4D建模

Haoran Lu, Shang Wu, Songling Liu, Jianshu Zhang, Maojiang Su, Guo Ye, Chenwei Xu, Lie Lu, Pranav Maneriker, Fan Du, Manling Li, Zhaoran Wang, Han Liu

机构 * Northwestern University(西北大学) Dolby Laboratories(杜比实验室)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 提出Phys4D流水线,通过三阶段训练(伪监督预训练、物理监督微调、强化学习校正)从视频扩散模型学习物理一致的4D世界表示,显著提升细粒度时空与物理一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18324 2026-06-16 cs.CV cs.AI cs.GR cs.LG stat.ML 版本更新 83%

Improved Baselines with Representation Autoencoders

改进的基于表示自动编码器的基线

Jaskirat Singh, Boyang Zheng, Zongze Wu, Richard Zhang, Eli Shechtman, Saining Xie

机构 * Adobe Research(Adobe研究院) ANU(澳大利亚国立大学) New York University(纽约大学)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文研究了基于表示自动编码器(RAE)的设计选择,发现三个见解,简化并改进了RAE。首先,研究了一种通用公式,将表示定义为最后k个编码器层的总和,而不是仅最终层。其次,研究了RAE与表示对齐(REPA)的假设,发现两者具有互补的工作机制。最后,改进了RAE在无分类器指导(CFG)中的表现,通过重新参数化DiT模型输出,实现了无需训练第二个模型的指导效果。RAEv2在ImageNet-256上达到了1.06的gFID,且训练效率显著提高。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09877 2026-06-02 cs.CV cs.AI cs.RO 83%

Genie 4D: Semantic-Prior-Guided 4D Dynamic Scene Reconstruction

Genie 4D:语义先验引导的4D动态场景重建

Yiru Yang, Zhuojie Wu, Nishant Kumar Singh, Max Schulthess

机构 * University of Zurich(苏黎世大学) ETH Zurich(苏黎世联邦理工学院)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 提出Genie 4D框架,结合实时视觉惯性高斯泼溅前端和前馈4D骨干网络,利用冻结的DINOv3特征作为结构先验抑制身份漂移,并通过条件扩散精炼器恢复高频细节,最终通过轻量级潜在动作头实现用户可控的4D世界模型重建。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21592 2026-05-26 cs.CV cs.AI cs.LG 83%

What Happens Next? Anticipating Future Motion by Generating Point Trajectories

接下来会发生什么?通过生成点轨迹预测未来运动

Gabrijel Boduljak, Laurynas Karazija, Iro Laina, Christian Rupprecht, Andrea Vedaldi

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 提出一种基于单张图像预测未来运动的方法,通过生成密集轨迹网格来捕捉场景动态和不确定性,相比现有方法更准确多样,并验证其在机器人等下游任务中的有效性。

Journal ref ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24631 2026-05-26 cs.LG cs.AI cs.CV 83%

Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion

超越生成先验:JEPA引导扩散的少数采样

Sol Park, Soobin Um

机构 * Department of Artificial Intelligence, Kookmin University, Seoul, South Korea(人工智能系,韩国全州大学,首尔)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 提出一种基于世界模型JEPA引导的扩散采样框架,通过近似策略实现高效计算,在无条件、类别条件和文本到图像生成中提升少数样本的保真度和语义有效性。

Comments ICML 2026, 21 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.04128 2026-05-21 cs.GR cs.AI cs.CL cs.CV cs.LG 83%

JoyAI-Image: Awaking Spatial Intelligence in Unified Multimodal Understanding and Generation

JoyAI-Image: 激活统一多模态理解和生成中的空间智能

Lin Song, Wenbo Li, Guoqing Ma, Wei Tang, Bo Wang, Yuan Zhang, Yijun Yang, Yicheng Xiao, Jianhui Liu, Yanbing Zhang, Guohui Zhang, Wenhu Zhang, Hang Xu, Nan Jiang, Xin Han, Haoze Sun, Maoquan Zhang, Haoyang Huang, Nan Duan

机构 * Joy Future Academy, JD(京东探索研究院)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出JoyAI-Image,一种统一的多模态基础模型,用于视觉理解、文本到图像生成和指令引导的图像编辑。该模型结合了空间增强的多模态大语言模型(MLLM)和多模态扩散Transformer(MMDiT),通过共享的多模态接口实现感知与生成的交互。构建可扩展的训练配方,结合统一指令微调、长文本渲染监督、空间 grounded 数据和通用及空间编辑信号,使模型具备广泛的多模态能力,同时增强几何感知推理和可控视觉合成。实验表明,JoyAI-Image在理解、生成、长文本渲染和编辑基准上达到最先进的性能。更重要的是,增强的理解、可控的空间编辑和新视角辅助推理之间的双向循环使模型超越一般视觉能力,向更强的空间智能发展。

Comments Code: https://github.com/jd-opensource/JoyAI-Image

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.16412 2026-05-19 cs.RO cs.CV 83%

SCAR: Self-Supervised Continuous Action Representation Learning

SCAR:自监督连续动作表示学习

Hongjia Liu, Fan Feng, Minghao Fu, Xinyue Wang, Haofei Lu, Biwei Huang

机构 * University of California, San Diego(加州大学圣地亚哥分校) KTH Royal Institute of Technology(皇家理工学院)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出SCAR框架,通过自监督学习统一动作表示,提升跨体素和任务的泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26694 2026-05-08 cs.RO cs.AI cs.CV 83%

Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising

从视频先验构建统一的4D世界动作模型

Jun Guo, Qiwei Li, Peiyan Li, Zilong Chen, Nan Sun, Yifei Su, Heyun Wang, Yuan Zhang, Xinghang Li, Huaping Liu

机构 * Tsinghua University(清华大学) Xiaomi Robotics(小米机器人) Peking University(北京大学) CASIA

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出X-WAM,统一了实时机器人动作执行与高保真的4D世界合成,在单一框架中解决传统统一世界模型仅建模2D像素空间且无法平衡动作效率与世界建模质量的问题。

Comments Project website: https://sharinka0715.github.io/X-WAM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00078 2026-05-04 cs.RO cs.CV cs.LG 83%

Being-H0.7: A Latent World-Action Model from Egocentric Videos

Being-H0.7: 一种从第一人称视频中学习的潜在世界-动作模型

Hao Luo, Wanpeng Zhang, Yicheng Feng, Sipeng Zheng, Haiweng Xu, Chaoyi Xu, Ziheng Xi, Yuhui Fu, Zongqing Lu

机构 * BeingBeyond Team(BeingBeyond团队)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出Being-H0.7,通过在感知与动作之间插入可学习的潜在查询,实现未来感知的VLA策略,无需生成未来帧,提升预测效率与部署性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22879 2026-04-28 cs.MA cs.AI cs.CR cs.LG 83%

Beyond Single-Agent Alignment: Preventing Context-Fragmented Violations in Multi-Agent Systems

超越单体对齐:防止多智能体系统中的上下文碎片化违规

Jie Wu, Ming Gong

机构 * Atlassian(Atlassian公司)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出Distributed Sentinel架构,通过Semantic Taint Token协议解决多智能体系统中因上下文碎片化导致的政策违规问题,实验证明其在跨域政策验证中的高效性与可靠性。

Comments 34 pages, 3 figures, 20 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24440 2026-03-26 cs.LG cs.AI cs.CV 83%

CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents

CUA-Suite:大规模人工标注的视频演示用于计算机使用代理

Xiangru Jian, Shravan Nayak, Kevin Qinghong Lin, Aarash Feizi, Kaixin Li, Patrice Bechard, Spandana Gella, Sai Rajeswar

机构 * ServiceNow University of Waterloo(多伦多大学) Mila Université de Montréal(蒙特利尔大学) McGill University(麦吉尔大学) University of Oxford(牛津大学) National University of Singapore(新加坡国立大学)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 CUA-Suite通过提供连续高质量视频演示和密集标注,解决计算机使用代理训练中视频数据不足的问题,包含约55小时的专家视频和600万帧数据,支持多模态研究。

Comments Project Page: https://cua-suite.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05848 2026-03-24 cs.CV cs.AI cs.RO 83%

Goal Force: Teaching Video Models To Accomplish Physics-Conditioned Goals

目标力:教视频模型实现物理条件化的目标

Nate Gillman, Yinghua Zhou, Zitian Tang, Evan Luo, Arjan Chakravarthy, Daksh Aggarwal, Michael Freeman, Charles Herrmann, Chen Sun

机构 * Brown University(布朗大学) Cornell University(康奈尔大学)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出Goal Force框架,通过显式力矢量和中间动力学定义目标,使视频模型能零样本泛化到复杂现实场景,实现基于物理的视频生成与规划。

Comments Camera ready version (CVPR 2026). Code and interactive demos at https://goal-force.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏