机构
*
University of Science and Technology of China(中国科学技术大学)
;
SenseTime Research(商汤科技研究院)
;
National University of Singapore(新加坡国立大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
Overview of the NLPCC 2026 Shared Task 1: Difficulty-Aware Multilingual and Multimodal Medical Instructional Video Understanding Evaluation
NLPCC 2026共享任务1概述:难度感知多语言多模态医学教学视频理解评估
Shenxi Liu, Kan Li, Mingyang Zhao, Yuhang Tian, Bin Li
机构
*
School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与工程学院)
;
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算学系)
;
Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
MMR-V: What's Left Unsaid? A Benchmark for Multimodal Deep Reasoning in Videos
MMR-V:未言明的是什么?视频中多模态深度推理的基准测试
Kejian Zhu, Zhuoran Jin, Hongbang Yuan, Jiachun Li, Shangqing Tu, Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao
机构
*
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences, Beijing, China(认知与决策智能复杂系统重点实验室,自动化研究所,中国科学院,北京,中国)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)
;
Tsinghua University(清华大学)
Unsupervised Multimodal Clustering for Semantics Discovery in Multimodal Utterances
用于多模态话语语义发现的无监督多模态聚类
Hanlei Zhang, Hua Xu, Fei Long, Xin Wang, Kai Gao
机构
*
State Key Laboratory of Intelligent Technology and Systems, Department of Computer Science and Technology, Tsinghua University(智能技术与系统国家重点实验室,计算机科学与技术系,清华大学)
;
School of Information Science and Engineering, Hebei University of Science and Technology(信息科学与工程学院,河北科技大学)
;
Samton (Jiangxi) Technology Development Co.,Ltd(江西松通科技发展有限公司)
Learning Sparse Latent Predictive Foundation Model for Multimodal Neuroimaging
学习用于多模态神经影像的稀疏潜在预测基础模型
Haoxu Huang, Long Chen, Jingyun Chen, Jinu Hyun, James Ryan Loftus, Kara Melmed, Daniel Orringer, Jennifer Frontera, Seena Dehkharghani, Arjun Masurkar, Narges Razavian
机构
*
New York University, Center for Data Science(纽约大学数据科学中心)
;
NYU Grossman School of Medicine, Department of Radiology(纽约大学格罗斯曼医学院放射学系)
;
State University of New York at Binghamton, School of Computing(纽约州立大学宾汉姆顿分校计算机学院)
;
NYU Grossman School of Medicine, Department of Neurology(纽约大学格罗斯曼医学院神经病学系)
;
NYU Grossman School of Medicine, Department of Neurosurgery(纽约大学格罗斯曼医学院神经外科学系)
;
NYU Grossman School of Medicine, Department of Pathology(纽约大学格罗斯曼医学院病理学系)
;
School of Medicine, Department of Radiology, Stanford(斯坦福大学医学院放射学系)
;
NYU Grossman School of Medicine, Department of Neuroscience(纽约大学格罗斯曼医学院神经科学系)
;
NYU Grossman School of Medicine, Neuroscience Institute(纽约大学格罗斯曼医学院神经科学研究所)
Multi-Dimensional Quality Assessment for AI-Generated Human-Centric Videos: Dataset and Model
人工智能生成的以人为中心的视频的多维质量评估:数据集与模型
Sijing Wu, Yunhao Li, Huiyu Duan, Yucheng Zhu, Xiongkuo Min, Patrick Le Callet, Guangtao Zhai
机构
*
Institute of Image Communication and Network Engineering, Shanghai Jiao Tong University(上海交通大学图像通信与网络工程研究所)
;
USC-SJTU Institute of Cultural and Creative Industry, Shanghai Jiao Tong University(上海交通大学南加州大学文化创意产业学院)
;
Polytech Nantes, Université de Nantes(法国南特大学高等理工学院)
Local Brushstroke Quality Assessment via Vision-Language Feedback
通过视觉-语言反馈进行局部笔触质量评估
Mio Mitamura, Hirokatsu Kataoka
机构
*
Tokyo Institute of Science High School(东京理科大学附属高中)
;
National Institute of Advanced Industrial Science and Technology (AIST)(国立先进工业科学技术研究所)
;
Visual Geometry Group, University of Oxford(牛津大学视觉几何组)
CommentsThis work was carried out in the EssilorLuxottica "Smart Eyewear Lab", a Joint Research Center between EssilorLuxottica and Politecnico di Milano
A2RL V\textsubscript{max}: The A2RL autonomous racing dataset for long-range, high-speed perception and multi-vehicle interaction
A2RL V下标max:用于远程、高速感知和多车辆交互的A2RL自动驾驶数据集
Marvin Klemp, Dominic Ebner, Cornelius Schröder, Davide Malvezzi, László Turányi, Riccardo Donati, Ilia Schminik, Xia Ning, Yanxin Zhou, Matthew Flagg, Christoph Stiller, Markus Lienkamp, Marko Bertogna, Gergely Bári, Andreas Birk, Ren Jin, Chen Lv, Johannes Betz
机构
*
Institute of Measurement and Control Systems, Karlsruhe Institute of Technology(测量与控制系统研究所,卡尔斯鲁厄理工学院)
;
Institute of Automotive Technology, Technical University of Munich(汽车技术研究所,慕尼黑工业大学)
;
Professorship of Autonomous Vehicle Systems, Technical University of Munich(自动驾驶车辆系统教授职位,慕尼黑工业大学)
;
University of Modena and Reggio Emilia(摩德纳大学和雷焦艾米利亚大学)
;
Humda Lab, Széchenyi István University(胡姆达实验室,塞切尼·伊什特万大学)
;
Politecnico di Milano(米兰理工大学)
;
Constructor University(康斯坦丁大学)
;
Beijing Institute of Technology(北京理工大学)
;
Nanyang Technological University(南洋理工大学)
;
Code 19 Racing(代码19赛车)
;
Munich Institute of Robotics and Machine Intelligence (MIRMI), Technical University of Munich(慕尼黑机器人与机器智能研究所(MIRMI),慕尼黑工业大学)