机构
*
School of Biomedical Engineering, Division of Life Sciences and Medicine, University of Science and Technology of China(生物医学工程学院,生命科学与医学系,中国科学技术大学)
;
Medical Imaging, Robotics, Analytic Computing & Learning (MIRACLE) Lab, YRD-RIGHT, USTC Suzhou Institute for Advanced Research(医学影像、机器人、分析计算与学习(MIRACLE)实验室,YRD-RIGHT,中国科学技术大学苏州研究院)
;
Jiangsu Provincial Key Laboratory of Multimodal Digital Twin Technology(江苏省多模态数字孪生技术重点实验室)
;
Biomedical Basic Research Center (BBRC) of Jiangsu Province(江苏省生物医学基础研究中心)
;
Department of Radiology, The First Affiliated Hospital of USTC, Division of Life Sciences and Medicine, USTC(放射科,中国科学技术大学第一附属医院,生命科学与医学系,中国科学技术大学)
;
Anhui IFLYTEK CO., Ltd(安徽科大讯飞股份有限公司)
;
School of Medicine, Stanford University(医学院,斯坦福大学)
;
State Key Laboratory of Precision and Intelligent Chemistry, Hefei, Anhui, China(安徽省精密与智能化学重点实验室,合肥,安徽,中国)
UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment
UF-AMA: 通过自适应多模态对齐的跨域情感识别统一框架
Zheng Wang, Shuo Wang, Junhong Wang
机构
*
Institute of Advanced Technology, University of Science and Technology of China(中国科学技术大学先进技术研究院)
;
Department of Electronic Engineering and Information Science, University of Science and Technology of China(中国科学技术大学电子工程与信息科学系)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)
BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization
BitsMoE: 面向MoE大语言模型量化的频谱能量引导比特分配
Jiayu Zhao, Zihan Teng, Minhao Fan, Tianrui Ma, Wentao Ren, Song Chen, Weichen Liu
机构
*
School of Microelectronics, University of Science and Technology of China(中国科学技术大学微电子学院)
;
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)
SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting
SCOPE: 信号校准的在线策略蒸馏增强与双路径自适应加权
Binbin Zheng, Xing Ma, Yiheng Liang, Jingqing Ruan, Xiaoliang Fu, Kepeng Lin, Benchang Zhu, Ke Zeng, Xunliang Cai
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Meituan LongCat Interaction Team(美团 LongCat 交互团队)
;
Nanjing University(南京大学)
;
Fudan University(复旦大学)
;
Huazhong University of Science and Technology(华中科技大学)
LookWise: Knowing When and Where to Look for Fine-Grained Visual Reasoning in Multimodal Large Language Models
LookWise: 知道何时何地关注多模态大语言模型中的细粒度视觉推理
Yuxiang Shen, Hailong Huang, Zhenkun Gao, Xueheng Li, Man Zhou, Chengjun Xie, Haoxuan Che, Xuanhua He, Jie Zhang
机构
*
Institute of Intelligent Machines, Hefei Institutes of Physical Science, Chinese Academy of Sciences(智能机器研究所,合肥物理科学研究院,中国科学院)
;
University of Science and Technology of China(中国科学技术大学)
;
Zhejiang University(浙江大学)
;
East China Normal University(华东师范大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
看、规划、回退:面向鲁棒机器人操作的进度感知视觉-语言-动作模型
Tingjun Dai, Mingfei Han, Tingwen Du, Zhiheng Liu, Zihao Zhang, Zhihui Li, Salman Khan, Jun Yu, Xiaojun Chang
机构
*
School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学)
;
University of Technology Sydney(新南威尔士大学)
;
Department of Computer Vision, Mohamed Bin Zayed University of Artificial Intelligence(人工智能与计算机视觉系,Mohamed Bin Zayed人工智能大学)
;
The University of Hong Kong(香港大学)
;
Institute of AI for Industry, Chinese Academy of Sciences(产业人工智能研究所,中国科学院)
;
School of Intelligent Science and Engineering, Harbin Institute of Technology (Shenzhen)(智能科学与工程学院,哈尔滨工业大学(深圳))