机构
*
University of Science and Technology of China(中国科学技术大学)
;
City University of Hong Kong(香港城市大学)
;
Kuaishou Technology(快手科技)
;
Zhejiang University(浙江大学)
SridBench: Benchmark of Scientific Research Illustration Drawing of Image Generation Model
Yifan Chang, Yukang Feng, Jianwen Sun, Jiaxin Ai, Chuanhao Li, S. Kevin Zhou, Kaipeng Zhang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Nankai University(南开大学)
;
Wuhan University(武汉大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
机构
*
AIDS and SIAR, University of Science and Technology of China(艾滋病与SIAR,中国科学技术大学)
;
ESAT-PSI, KU Leuven(ESAT-PSI,比利时鲁汶大学)
;
Institute for Advanced Algorithms Research(先进算法研究所)
;
The Hong Kong Polytechnic University(香港理工大学)
AgentRecBench: Benchmarking LLM Agent-based Personalized Recommender Systems
Yu Shang, Peijie Liu, Yuwei Yan, Zijing Wu, Leheng Sheng, Yuanqing Yu, Chumeng Jiang, An Zhang, Fengli Xu, Yu Wang, Min Zhang, Yong Li
机构
*
Tsinghua University(清华大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
University of Science and Technology of China(中国科学技术大学)
;
National University of Singapore(新加坡国立大学)
MV-CoLight: Efficient Object Compositing with Consistent Lighting and Shadow Generation
Kerui Ren, Jiayang Bai, Linning Xu, Lihan Jiang, Jiangmiao Pang, Mulin Yu, Bo Dai
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Nanjing University(南京大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Science and Technology of China(中国科学技术大学)
;
The University of Hong Kong(香港大学)
机构
*
Great Bay University(大西洋大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Dongguan Key Laboratory for Intelligence and Information Technology(东莞智能与信息技术重点实验室)
;
GRGBanking Equipment Co., Ltd.(GRGBanking设备有限公司)
;
South China University of Technology(华南理工大学)
;
Lappeenranta University of Technology(拉普兰塔理工大学)
;
Macao Polytechnic University(澳门理工学院)
;
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs
Jiakang Yuan, Tianshuo Peng, Yilei Jiang, Yiting Lu, Renrui Zhang, Kaituo Feng, Chaoyou Fu, Tao Chen, Lei Bai, Bo Zhang, Xiangyu Yue
机构
*
Fudan University(复旦大学)
;
MMLab, The Chinese University of Hong Kong(中大香港人工智能实验室)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Nanjing University(南京大学)
机构
*
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知大学科学与技术大学实验室)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究所)
机构
*
State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究院,合肥综合性国家科学中心)
The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition
Ming Gao, Shilong Wu, Hang Chen, Jun Du, Chin-Hui Lee, Shinji Watanabe, Jingdong Chen, Siniscalchi Sabato Marco, Odette Scharenborg
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Northwestern Polytechnical University(西北工业大学)
;
University of Palermo(巴勒莫大学)
;
Delft University of Technology(代尔夫特理工大学)
CommentsAccepted by Interspeech 2025. Camera-ready version
Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval
Rong-Cheng Tu, Wenhao Sun, Hanzhe You, Yingjie Wang, Jiaxing Huang, Li Shen, Dacheng Tao
机构
*
College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算机与数据科学学院)
;
School of Information Science and Technology, University of Science and Technology of China, Hefei, China(中国科学技术大学信息科学与技术学院)
;
Sun Yat-sen University Shenzhen Campus, School of Cyber Science and Technology, Shenzhen, China(中山大学深圳校区计算机科学与技术学院)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
National Engineering Research Center of Speech and Language Information Processing(语音与语言信息处理国家级工程研究中心)
;
Interdisciplinary Research Center for Linguistic Sciences(语言科学交叉研究中心)
机构
*
Joint SDU-NTU Centre for Artificial Intelligence Research&School of Software(山东大学与南京工业大学人工智能研究中心及软件学院)
;
University of Science and Technology of China(中国科学技术大学)
;
Shanghai Jiaotong University(上海交通大学)
;
Nanjing University(南京大学)
;
Nanjing University of Posts and Telecommunications(南京邮电大学)
;
Jiangnan University(江南大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
Disentangled Human Body Representation Based on Unsupervised Semantic-Aware Learning
Lu Wang, Xishuai Peng, S. Kevin Zhou
机构
*
X-ray Product, Siemens Shanghai Medical Equipment Ltd.(西门子上海医疗设备有限公司X射线产品部)
;
University of Science and Technology of China, School of Biomedical Engineering & Suzhou Institute for Advanced Research(中国科学技术大学生物医学工程学院及苏州先进研究院)
A General Knowledge Injection Framework for ICD Coding
Xu Zhang, Kun Zhang, Wenxin Ma, Rongsheng Wang, Chenxu Wu, Yingtai Li, S. Kevin Zhou
机构
*
School of Biomedical Engineering, Division of Life Sciences and Medicine, USTC(生物医学工程学院,生命科学与医学系,中国科学技术大学)
;
MIRACLE Center, Suzhou Institute for Advance Research, USTC(MIRACLE中心,苏州先进研究院,中国科学技术大学)
;
Jiangsu Provincial Key Laboratory of Multimodal Digital Twin Technology(江苏省多模态数字孪生技术重点实验室)
;
State Key Laboratory of Precision and Intelligent Chemistry, USTC(精密与智能化学国家重点实验室)
机构
*
Zhejiang University(浙江大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Fudan University(复旦大学)
;
State Key Laboratory of Industrial Control Technology(工业控制技术国家重点实验室)
PMQ-VE: Progressive Multi-Frame Quantization for Video Enhancement
ZhanFeng Feng, Long Peng, Xin Di, Yong Guo, Wenbo Li, Yulun Zhang, Renjing Pei, Yang Wang, Yang Cao, Zheng-Jun Zha
机构
*
USTC(中国科学技术大学)
;
Max Planck Institute(马克斯·普朗克研究所)
;
CUHK(香港中文大学)
;
SJTU(上海交通大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Chang’an University(长安大学)
Distilling Textual Priors from LLM to Efficient Image Fusion
Ran Zhang, Xuanhua He, Ke Cao, Liu Liu, Li Zhang, Man Zhou, Jie Zhang
机构
*
Hefei University of Technology(合肥工业大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Hefei Institutes of Physical Science, Chinese Academy of Sciences(中国科学院合肥物质科学研究院)