RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis
RoboGaze: 通过结构化视觉-语言分析评估机器人世界模型
Minh-Loi Nguyen, Nghiem Tuong Diep, Hung Khang Nguyen, Minh Le, Doanh Le Thien, Hoang H. Tran, Dung D. Le, Vu N. Duong, Daniel Sonntag, An Thai Le, Duy Minh Ho Nguyen, Vien Anh Ngo, Tran Van Nhiem
机构
*
Ho Chi Minh City University of Science, Vietnam National University(越南国家科学大学胡志明市大学)
;
VinRobotics
;
VinUniversity
;
German Research Center for AI (DFKI)(人工智能研究中心(DFKI))
;
University of Stuttgart(斯图加特大学)
;
Max Planck Research School for Intelligent Systems(智能系统马克斯·普朗克研究学校)
;
Technische Universität Darmstadt(达姆施塔特技术大学)
Comments07 pages, 01 figure, accepted for presentation at the IEEE International Conference on Communication, Computing, Networking, and Control in Cyber-Physical Systems (CCNCPS 2026)
CommentsThis preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this contribution is published in Advances in Information Retrieval, ECIR 2026, Lecture Notes in Computer Science, vol. 16485, pp. 458-473, and is available online at https://doi.org/10.1007/978-3-032-21324-2_35
Journal refAdvances in Information Retrieval, ECIR 2026. Lecture Notes in Computer Science, vol. 16485, pp. 458-473. Springer, Cham (2026)
ORCA: Open-ended Response Correctness Assessment for Audio Question Answering
ORCA:面向音频问答的开放性回答正确性评估
Šimon Sedláček, Sara Barahona, Bolaji Yusuf, Laura Herrera-Alarcón, Santosh Kesiraju, Cecilia Bolaños, Alicia Lozano-Diez, Sathvik Udupa, Fernando López, Allison Ferner, Ramani Duraiswami, Jan Černocký
机构
*
Brno University of Technology(布尔诺科技大学)
;
Universidad Autónoma de Madrid(马德里自治大学)
;
University of Buenos Aires(布宜诺斯艾利斯大学)
;
Tufts University(塔夫茨大学)
;
University of Maryland(马里兰大学)
机构
*
Qwen Large Model Application Team, Alibaba(阿里云大模型应用团队)
;
Alibaba Zhejiang University(阿里巴巴浙江大学)
;
University of Waterloo(多伦多大学)
;
Vector Institute(向量研究所)
;
Nanjing University(南京大学)
;
Binjiang Institute of Zhejiang University(浙江大学滨江学院)
TopoAgent: An Agentic Framework for Automated Topology Learning in Medical Imaging
TopoAgent: 医学影像中自动拓扑学习的智能体框架
Guangyu Meng, Pengfei Gu, Xueyang Li, Yiyu Shi, Erin Wolf Chambers, Danny Z. Chen
机构
*
Dept. of Computer Science and Engineering, University of Notre Dame(内布拉斯加大学达灵顿分校计算机科学与工程系)
;
Dept. of Computer Science, The University of Texas Rio Grande Valley(德克萨斯大学里奥格兰德谷分校计算机科学系)
SFBench: The SciFy Scientific Feasibility Benchmark
SFBench:SciFy科学可行性基准
Cash Costello, James Mayfield, Elsbeth Turcan, Christine Piatko, Christina K. Pikas, Justin Rokisky, Sam Scheck, Chris Ribaudo, Ritwik Bose, Alex Memory
AURORA: Asymmetry and Update-Induced Rotation for Robust Hallucination Detection in Large Language Models
AURORA:用于大型语言模型中鲁棒幻觉检测的不对称性与更新诱导旋转
Zishuai Zhang, Hainan Zhang, Zhiming Zheng
机构
*
School of Artificial Intelligence, Beihang University, China(北京航空航天大学人工智能学院)
;
Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing, Beihang University, China(北京航空航天大学未来区块链与隐私计算先进创新中心)
BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards
BV-Blend: 不确定性加权历史基线用于稳定无评论家可验证奖励强化学习
Yupeng Chang, Yuan Wu, Yi Chang
机构
*
School of Artificial Intelligence, Jilin University(吉林大学人工智能学院)
;
Engineering Research Center of Knowledge-Driven Human-Machine Intelligence, MOE, China(知识驱动人机智能工程研究中心,教育部,中国)
;
International Center of Future Science, Jilin University(未来科学国际中心,吉林大学)