RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
RadAgent:一种用于胸部CT逐步解读的工具型AI智能体
Mélanie Roschewitz, Kenneth Styppa, Yitian Tao, Jiwoong Sohn, Jean-Benoit Delbrouck, Benjamin Gundersen, Nicolas Deperrois, Christian Bluethgen, Julia E. Vogt, Bjoern Menze, Farhad Nooralahzadeh, Michael Krauthammer, Michael Moor
机构
*
Department of Biosystems Science and Engineering, ETH Zurich(生物系统科学与工程系,苏黎世联邦理工学院)
;
ETH AI Center, Zurich(ETH人工智能中心,苏黎世)
;
Department of Computer Science, ETH Zurich(计算机科学系,苏黎世联邦理工学院)
;
Faculty of Computer Science and Mathematics, Heidelberg University(计算机科学与数学学院,海德堡大学)
;
Stanford Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学人工智能在医学和影像中的中心)
;
Department of Radiology, Stanford University(放射科,斯坦福大学)
;
Department of Quantitative Biomedicine, University of Zurich(定量生物医学系,苏黎世大学)
;
Institute of Computer Science, Zurich University of Applied Sciences(应用科学大学计算机科学研究所)
A Deep Learning Model of Mental Rotation Informed by Interactive VR Experiments
基于交互式VR实验的心理旋转深度学习模型
Raymond Khazoum, Daniela Fernandes, Aleksandr Krylov, Qin Li, Stephane Deny
机构
*
Department of Computer Science, Aalto University, Espoo, Finland(奥卢大学计算机科学系,芬兰埃斯波)
;
Department of Neuroscience and Biomedical Engineering, Aalto University, Espoo, Finland(奥卢大学神经科学与生物医学工程系,芬兰埃斯波)
机构
*
Department of Computer Science and Software Engineering, Concordia University, Montreal, Canada(计算机科学与软件工程系,康科迪亚大学,蒙特利尔,加拿大)
;
Department of Electrical and Computer Engineering, Concordia University, Montreal, Canada(电气与计算机工程系,康科迪亚大学,蒙特利尔,加拿大)
CommentsThis preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this contribution will be published as part of the MICCAI 2026 proceedings in October
Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers
张量记忆:用于长程Transformer的固定大小循环状态
Kabir Swain, Sijie Han, Daniel Karl I. Weidele, Mauro Martino, Antonio Torralba
机构
*
Massachusetts Institute of Technology, Cambridge, MA, USA(麻省理工学院)
;
IBM Research, Cambridge, MA, USA(IBM研究院)
;
University of Toronto, Toronto, Canada(多伦多大学)
Ligand-Conditioned Discrete Diffusion for Protein Sequence-Structure Co-Design
配体条件离散扩散用于蛋白质序列-结构协同设计
Chen Wei, Fanding Xu, Minghao Sun, Zhiyuan Liu, Lin Wang, Tianrui Jia, Yihang Zhou, Yang Zhang
机构
*
Xi’an University of Posts & Telecommunications(西安邮电大学)
;
National University of Singapore(新加坡国立大学)
;
Xi’an Jiaotong University(西安交通大学)
;
Institute of Systems Medicine, Chinese Academy of Medical Sciences(中国医学科学院系统医学研究院)
MMSI-Bench: A Benchmark for Multi-Image Spatial Intelligence
MMSI-Bench:多图像空间智能基准
Sihan Yang, Runsen Xu, Yiman Xie, Sizhe Yang, Mo Li, Jingli Lin, Chenming Zhu, Xiaochen Chen, Haodong Duan, Xiangyu Yue, Dahua Lin, Tai Wang, Jiangmiao Pang
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Zhejiang University(浙江大学)
;
Tsinghua University(清华大学)
;
Shanghai Jiaotong University(上海交通大学)
;
University of Hong Kong(香港大学)
;
Beijing Normal University(北京师范大学)
AirVista-II: An Agentic System for Embodied UAVs Toward Dynamic Scene Semantic Understanding
AirVista-II:面向动态场景语义理解的具身无人机智能体系统
Fei Lin, Yonglin Tian, Tengchao Zhang, Jun Huang, Sangtian Guan, Fei-Yue Wang
机构
*
Department of Engineering Science, Faculty of Innovation Engineering, Macau University of Science and Technology(创新工程学院工程科学系,澳门科学技术大学)
;
State Key Laboratory for Management and Control of Complex Systems, Institute of Automation, Chinese Academy of Sciences(复杂系统管理与控制国家重点实验室,中国科学院自动化研究所)
;
State Key Laboratory for Management and Control of Complex Systems, Chinese Academy of Sciences(复杂系统管理与控制国家重点实验室,中国科学院)
Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models
Circle-RoPE: 用于大视觉-语言模型的锥形解耦旋转位置嵌入
Chengcheng Wang, Jianyuan Guo, Hongguang Li, Yuchuan Tian, Ying Nie, Chang Xu, Kai Han
机构
*
Huawei Noah's Ark Lab.(华为诺亚实验室)
;
City University of Hong Kong.(香港城市大学)
;
University of Sydney.(悉尼大学)
;
State Key Lab of General AI, School of Intelligence Science and Technology, Peking University(北京大学人工智能国家重点实验室,智能科学与技术学院)
CommentsExtended version of the paper published at LREC 2026 (Palma de Mallorca, Spain), with expanded VLM baselines and inter-annotator agreement analysis
Journal refProceedings of the 15th Language Resources and Evaluation Conference (LREC 2026), Palma de Mallorca, Spain