机构
*
The University of Hong Kong(香港大学)
;
Huawei Technologies Co., Ltd(华为技术有限公司)
;
University College London(伦敦大学学院)
;
The Chinese University of Hong Kong(香港中文大学)
Physics-based phenomenological characterization of cross-modal bias in multimodal models
基于物理现象的多模态模型跨模态偏差表征
Hyeongmo Kim, Sohyun Kang, Yerin Choi, Seungyeon Ji, Junhyuk Woo, Hyunsuk Chung, Soyeon Caren Han, Kyungreem Han
机构
*
B rain Science Institute(脑科学研究院)
;
Korea Institute of Science and Technology(韩国科学技术院)
;
Department of Physics and Astronomy(物理与天文学系)
;
Department of Computer Science and Engineering(计算机科学与工程系)
;
University of Science and Technology KIST School(科学技术KIST学院)
专题命中
代码与定理证明
:reasoning(abstract);分类 cs.AI
AI总结
本文提出基于物理现象的多模态模型跨模态偏差表征方法,揭示多模态输入可能强化模态主导性。
CommentsBest Paper Award at BiasinAI track in AAAI2026
LogicGraph : Benchmarking Multi-Path Logical Reasoning via Neuro-Symbolic Generation and Verification
LogicGraph : 通过神经符号生成与验证进行多路径逻辑推理的基准测试
Yanrui Wu, Lingling Zhang, Xinyu Zhang, Jiayu Chang, Pengyu Li, Xu Jiang, Jingtao Hu, Jun Liu
机构
*
School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院)
;
Department of Electrical Engineering, Stanford University(斯坦福大学电气工程系)
;
School of Computer Science and Technology, Tiangong University(天津大学计算机科学与技术学院)
;
Ministry of Education Key Laboratory of Intelligent Networks and Network Security, China(中国教育部长智网络与网络安全重点实验室)
;
Shaanxi Province Key Laboratory of Big Data Knowledge Engineering, China(陕西省大数据知识工程重点实验室)
CommentsFinal version of the article accepted for publication on Scientific Reports. 29 pages (13 pages are from appendix), 8 figures, 2 tables, code for experiments replication and supplementary material provided at https://github.com/jtyska/llm-robotics-article/
Efficient and Explainable End-to-End Autonomous Driving via Masked Vision-Language-Action Diffusion
高效的端到端自动驾驶:通过掩码视觉-语言-动作扩散
Jiaru Zhang, Manav Gagvani, Can Cui, Juntong Peng, Ruqi Zhang, Ziran Wang
机构
*
Institute for Physical Artificial Intelligence (IPAI), Purdue University(物理人工智能研究所(IPAI)、普渡大学)
;
College of Engineering, Purdue University(工程学院、普渡大学)
;
Department of Computer Science, Purdue University(计算机科学系、普渡大学)
机构
*
School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)
;
Shanghai Engineering Research Center of Intelligent Vision and Imaging(智能视觉与成像上海工程研究中心)
;
Tongyi Lab, Alibaba Group(阿里集团通义实验室)
Monte Carlo Tree Diffusion with Multiple Experts for Protein Design
结合多专家的蒙特卡洛树扩散用于蛋白质设计
Xuefeng Liu, Mingxuan Cao, Songhao Jiang, Xiao Luo, Xiaotian Duan, Mengdi Wang, Tobin R. Sosnick, Jinbo Xu, Rick Stevens
机构
*
University of Chicago(芝加哥大学)
;
Data Science Institute(数据科学研究所)
;
Department of Biochemistry and Molecular Biology(生物化学与分子生物学系)
;
Toyota Technological Institute at Chicago(芝加哥丰田技术研究所)
;
Argonne National Laboratory(阿贡国家实验室)
;
Princeton University(普林斯顿大学)
Multi-Round Human-AI Collaboration with User-Specified Requirements
多轮人机协作与用户指定要求
Sima Noorani, Shayan Kiyani, Hamed Hassani, George Pappas
机构
*
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
A Survey on the Optimization of Large Language Model-based Agents
基于大语言模型代理的优化综述
Shangheng Du, Jiabao Zhao, Jinxin Shi, Zhentao Xie, Xin Jiang, Yanhong Bai, Liang He
机构
*
Shanghai Institute of Artificial Intelligence for Education, East China Normal University(上海人工智能教育研究院,东华大学)
;
School of Computer Science and Technology, East China Normal University(计算机科学与技术学院,东华大学)
;
School of Computer Science and Technology, Donghua University(计算机科学与技术学院,东华大学)