机构
*
College of Computer Science and Artificial Intelligence, Shanghai Key Laboratory of Intelligent Information Processing, Fudan University(复旦大学计算机科学与人工智能学院,上海智能信息处理重点实验室)
;
University of Oxford(牛津大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
Exploring the Underwater World Segmentation without Extra Training
探索无需额外训练的水下世界分割
Bingyu Li, Tao Huo, Da Zhang, Zhiyuan Zhao, Junyu Gao, Xuelong Li
机构
*
Institute of Artificial Intelligence (TeleAI), China Telecom, China(人工智能研究院(TeleAI),中国电信,中国)
;
University of Science and Technology of China, China(中国科学技术大学,中国)
;
Northwestern Polytechnical University, China(西北工业大学,中国)
专题命中
视觉定位与Grounding
:multimodal large language model(abstract);分类 cs.CV、cs.AI
Proactive Rejection and Grounded Execution: A Dual-Stage Intent Analysis Paradigm for Safe and Efficient AIoT Smart Homes
主动拒绝与 grounded 执行:一种双阶段意图分析范式用于安全高效的 AIoT 智能家居
Xinxin Jin, Zhengwei Ni, Zhengguo Sheng, Victor C. M. Leung
机构
*
School of Information and Electronic Engineering (Sussex Artificial Intelligence Institute), Zhejiang Gongshang University(信息与电子工程学院(Sussex人工智能研究院),浙江工商大学)
;
Sussex Artificial Intelligent Institute, Zhejiang Gongshang University(Sussex人工智能研究院,浙江工商大学)
;
Department of Engineering and Design, University of Sussex(工程与设计系, Sussex大学)
;
Artificial Intelligence Research Institute, Shenzhen MSU-BIT University(人工智能研究院,深圳MSU-BIT大学)
;
College of Computer Science and Software Engineering, Shenzhen University(计算机科学与软件工程学院,深圳大学)
;
Department of Electrical and Computer Engineering, The University of British Columbia(电气与计算机工程系,不列颠哥伦比亚大学)
Swadesh Jana, Cansu Sancaktar, Tomáš Daniš, Georg Martius, Antonio Orvieto, Pavel Kolev
机构
*
University of Tübingen(图宾根大学)
;
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
ELLIS Institute Tübingen(图宾根ELLIS研究所)
;
Tübingen AI Center(图宾根人工智能中心)
CommentsAccepted at ICLR 2026 Workshop on AI with Recursive Self-Improvement (RSI 2026) as Spotlight, and ICLR 2026 Workshop on Lifelong Agents (LLA 2026)
Wavelet-based Frame Selection by Detecting Semantic Boundary for Long Video Understanding
基于语义边界的小波帧选择用于长视频理解
Wang Chen, Yuhui Zeng, Yongdong Luo, Tianyu Xie, Luojun Lin, Jiayi Ji, Yan Zhang, Xiawu Zheng
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学)
;
College of Computer and Data Science, Fuzhou University(福州大学计算机与数据科学学院)
Open-Vocabulary Octree-Graph for 3D Scene Understanding
开放词汇八叉树图用于3D场景理解
Zhigang Wang, Yifei Su, Chenhui Li, Dong Wang, Yan Huang, Bin Zhao, Xuelong Li
机构
*
Northwestern Polytechnical University(西北工业大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
CASIA
;
TeleAI
MedArena: Comparing LLMs for Medicine-in-the-Wild Clinician Preferences
MedArena: 比较医疗领域LLM的临床医生偏好
Eric Wu, Kevin Wu, Jason Hom, Paul H. Yi, Angela Zhang, Alejandro Lozano, Jeff Nirschl, Jeff Tangney, Kevin Byram, Braydon Dymm, Narender Annapureddy, Eric Topol, David Ouyang, James Zou
机构
*
Department of Electrical Engineering, Stanford University(斯坦福大学电气工程系)
;
Department of Biomedical Data Science, Stanford University(斯坦福大学生物医学数据科学系)
;
Division of Hospital Medicine, Department of Medicine, Stanford School of Medicine(斯坦福医学院医学部住院医学科)
;
Department of Radiology, St. Jude Children's Research Hospital(圣 Jude 儿童研究医院放射科)
;
University of California, San Francisco(旧金山大学)
;
Department of Pathology and Laboratory Medicine, University of Wisconsin School of Medicine and Public Health(威斯康星大学医学与公共卫生学院病理学与实验室医学系)
;
Doximity, San Francisco, CA, USA(Doximity公司)
;
Department of Medicine, Division of Rheumatology and Immunology, Vanderbilt University Medical Center(范德比尔特大学医学中心医学系风湿病与免疫学科)
;
Department of Neurology, Charleston Area Medical Center(查尔斯顿医疗中心神经科)
;
Department of Translational Medicine, Scripps Research Translational Institute(斯克里普斯研究转化研究所转化医学系)
;
Kaiser Permanente Division of Research(凯撒医疗集团研究部)
On Theoretically-Driven LLM Agents for Multi-Dimensional Discourse Analysis
关于理论驱动的LLM代理在多维话语分析中的应用
Maciej Uberna, Michał Wawer, Jarosław A. Chudziak, Marcin Koszowy
机构
*
Laboratory of The New Ethos, Warsaw University of Technology, Poland(新伦理实验室,华沙理工大学,波兰)
;
Faculty of Electronics and Information Technology, Warsaw University of Technology, Poland(电子与信息技术学院,华沙理工大学,波兰)
Comments8 pages, 4 figures, 3 tables. This is the accepted version of the paper presented at the 18th International Conference on Agents and Artificial Intelligence (ICAART 2026), Marbella, Spain
Journal refProceedings of the 18th International Conference on Agents and Artificial Intelligence (ICAART 2026)