OmniSpace: Efficient Geometry Awareness for Autonomous Vehicles MLLMs
OmniSpace: 自动驾驶多模态大语言模型的高效几何感知
Hao Vo, Phu Loc Nguyen, Khoa Vo, Sieu Tran, Duc Minh Nguyen, Ngo Xuan Cuong, Nghi D. Q. Bui, Anh Nguyen, Duy Minh Ho Nguyen, Ngan Le
机构
*
University of Arkansas(阿肯色大学)
;
Google Research, Google(谷歌研究院)
;
University of Liverpool(利物浦大学)
;
Max Planck Research School for Intelligent Systems(马克斯·普朗克智能系统研究所)
HDRAgent: An Agentic Framework for Multi-Exposure HDR Imaging
HDRAgent: 一种用于多曝光HDR成像的智能体框架
Weiyu Zhou, Tao Hu, Yijian Wang, Xiaogang Xu, Ruixing Wang, Qingsen Yan
机构
*
School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院)
;
Shenzhen Research Institute, Northwestern Polytechnical University(西北工业大学深圳研究院)
;
Zhejiang University(浙江大学)
;
Camera Group, DJI(大疆相机部门)
PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow
PathoSage:通过经验感知的代理工作流实现病理学多源证据裁决
Chengyang Zhang, Wenchuan Zhang, Bo Li, Mengran Li, Bob Zhang, Yuhao Yi, Hong Bu, Jiancheng Lv
机构
*
College of Computer Science, Sichuan University(四川大学计算机科学学院)
;
Department of Pathology and Institute of Clinical Pathology, West China Hospital, Sichuan University(四川大学华西医院病理科/临床病理研究所)
;
Department of Computer and Information Science, University of Macau(澳门大学计算机与信息科学系)
;
School of Intelligent Systems Engineering, Sun Yat-sen University(中山大学智能工程学院)
Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning
通过自主经验探索与事后经验利用赋能GUI智能体任务规划
Tianyi Men, Zhuoran Jin, Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao
机构
*
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所复杂系统认知与决策智能重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
机构
*
Shanghai University(上海大学)
;
Shanghai Engineering Research Center of Motion Picture Special Effects(上海电影特效工程技术研究中心)
;
Institute for Math and AI, Wuhan School of Artificial Intelligence, Wuhan University(武汉大学数学与人工智能研究院武汉人工智能学院)
Multi-Modal Conditioned High-Resolution Transformer for Urban Electromagnetic Field Map Prediction Download PDF
面向城市电磁场地图预测的多模态条件高分辨率Transformer
Do-Eon Kim, Dongryul Park, Seungyoung Ahn, Namwoo Kang, Seong-heum Kim, Seongsin Kim
机构
*
Soongsil University(崇实大学)
;
Cho Chun Shik Graduate School of Mobility, Korea Advanced Institute of Science and Technology(韩国科学技术院赵春植移动研究生院)
;
Department of Intelligent Semiconductors, Soongsil University(崇实大学智能半导体系)
SciOrch: Learning to Orchestrate Expert LLMs for Solving Frontier Multimodal Scientific Reasoning Tasks
SciOrch: 学习编排专家大语言模型以解决前沿多模态科学推理任务
Jingru Guo, Xiangyuan Xue, Lian Zhang, Wanghan Xu, Siki Chen, Philip Torr, Wanli Ouyang, Lei Bai, Zhenfei Yin
机构
*
Imperial College London(伦敦帝国学院)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
University of Oxford(牛津大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
Comments7 pages, 7 figures. Accepted at the 2026 IEEE ICDL Conference. Cite as: L. Philipp, F. M. López, and J. Triesch, "Embodiment Shapes Rolling Behavior in a Multimodal Infant Model", in 2026 IEEE International Conference on Development and Learning (ICDL). IEEE, 2026, pp. 1-7