Self-Supervised Multi-Modal World Model with 4D Space-Time Embedding
具有4D空间-时间嵌入的自监督多模态世界模型
Lance Legel, Qin Huang, Brandon Voelker, Daniel Neamati, Patrick Alan Johnson, Favyen Bastani, Jeff Rose, James Ryan Hennessy, Robert Guralnick, Douglas Soltis, Pamela Soltis, Shaowen Wang
机构
*
Ecological Intelligence Lab(生态智能实验室)
;
School of Complex Adaptive Systems(复杂适应系统学院)
;
University of Houston(休斯顿大学)
;
Geosensing Systems Engineering & Sciences Lab(传感系统工程与科学实验室)
;
Stanford University(斯坦福大学)
;
Allen Institute for Artificial Intelligence(人工智能研究院)
;
Spatial Intelligence Lab(空间智能实验室)
;
Department of Computer Science(计算机科学系)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Florida Museum of Natural History(佛罗里达自然历史博物馆)
;
University of Florida(佛罗里达大学)
;
NSF Institute for Geospatial Understanding(国家科学基金会地理理解研究所)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
Stealth Fine-Tuning: Efficiently Breaking Alignment in RVLMs Using Self-Generated CoT
隐形微调:通过自动生成的CoT打破RVLMs的对齐
Le Yu, Zhengyue Zhao, Yawen Zheng, Yunhao Liu
机构
*
Machine Intelligence Laboratory, Sichuan University(四川大学人工智能实验室)
;
University of Wisconsin--Madison(威斯康星大学麦迪逊分校)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
Global Innovation Exchange, Tsinghua University(清华大学全球创新交流中心)