机构
*
Ant Group(蚂蚁集团)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Xi'an Polytechnic University(西安理工大学)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Nagoya University(名古屋大学)
;
Institute of Science Tokyo(东京科学大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Wuhan University of Technology(武汉理工大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Huawei Noah’s Ark Lab(华为诺亚方舟实验室)
机构
*
National and Local Joint Engineering Laboratory of Computer Aided Design, School of Software Engineering, Dalian University(大连大学软件工程学院计算机辅助设计国家地方联合工程实验室)
;
Department of Radiology, Xinhua Hospital Affiliated to Dalian University(大连大学附属新华医院放射科)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
University of California, San Francisco(加州大学旧金山分校)
;
Yale University(耶鲁大学)
;
The Hong Kong Polytechnic University(香港理工大学)
VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment
VLAFlow:通过协同训练和未来潜在对齐的视觉-语言-动作模型统一训练框架
Guoyang Xia, Fengfa Li, Hongjin Ji, Lei Ren, Fangxiang Feng, Kun Zhan, Yan Xie
机构
*
Li Auto Inc.(理想汽车)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
机构
*
Hong Kong University of Science and Technology(香港科学与技术大学)
;
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Peking University(北京大学)
;
Independent Researcher(独立研究者)
Action-Conditioned World Model for Goal Plane Probe Guidance in Robotic Ultrasound
用于机器人超声中目标平面探头引导的动作条件世界模型
Siqi Fan, Mingcong Chen, Ran Liu, Zixuan Yang, Xiaoyu Fu, Xiaoqing Gao, Yunhui Liu, Hongbin Liu
机构
*
Department of Mechanical and Automation Engineering, Chinese University of Hong Kong(香港中文大学机械与自动化工程学系)
;
Department of Biomedical Engineering, City University of Hong Kong(香港城市大学生物医学工程学系)
;
Centre for Artificial Intelligence and Robotics Hong Kong Institute of Science & Innovation, Chinese Academy of Sciences(中国科学院香港创新研究院人工智能与机器人中心)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Department of Surgical and Interventional Engineering, King’s College London(伦敦国王大学外科与介入工程学系)
;
Department of Vascular Ultrasound, Xuanwu Hospital, the Capital Medical University(首都医科大学宣武医院血管超声科)
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook
大型音频语言模型综述:通用性、可信度与展望
Kaiwen Luo, Zhenhong Zhou, Leyan Wang, Liang Lin, Tianyu Shao, Yuanhe Zhang, Yang Xiao, Yuxuan Li, Miao Yu, Kailin Lyu, Jiaming Zhang, Li Sun, Songze Li, Yueming Wu, Ting Dang, Xiaojun Jia, Dongrui Liu, Kai Li, Rohan Kumar Das, Siyuan Liang, Xinfeng Li, Qiankun Li, Jing Chen, Xingjun Ma, Kun Wang, Junhao Dong, Deqing Zou, Yu Cheng, Xia Hu, Zhigang Zeng, Sen Su, Yang Liu, Yu-Gang Jiang, Philip S. Yu, Yew-Soon Ong
机构
*
Nanyang Technological University(南洋理工大学)
;
Independent Researcher(独立研究者)
;
The University of Melbourne(墨尔本大学)
;
North China Electric Power University(华北电力大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Huazhong University of Science and Technology(华中科技大学)
;
Tsinghua University(清华大学)
;
Fortemedia Singapore(富媒体新加坡)
;
Tencent(腾讯)
;
Fudan University(复旦大学)
;
Wuhan University(武汉大学)
;
Chinese University of Hong Kong(香港中文大学)
;
Chongqing University of Posts and Telecommunications(重庆邮电大学)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
MMLab
;
The Chinese University of Hong Kong(香港中文大学)
;
ICMAT Spanish National Research Council(西班牙国家科研 council)
;
The High School Affiliated to Renmin University of China(中国人民大学附属高中)
;
Ren Hui Academy of Beijing(北京润辉学院)
Restore Text First, Enhance Image Later: Two-Stage Scene Text Image Super-Resolution with Glyph Structure Guidance
先恢复文本,后增强图像:基于字形结构引导的两阶段场景文本图像超分辨率
Minxing Luo, Linlong Fan, Wang Qiushi, Ge Wu, Yiyan Luo, Yuhang Yu, Jinwei Chen, Yaxing Wang, Qingnan Fan, Jian Yang
机构
*
VCIP, CS, Nankai University(南开大学计算机科学与技术学院)
;
vivo Mobile Communication Co. Ltd(vivo移动通信有限公司)
;
SDS, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)信息科学与技术学院)
机构
*
National Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家实验室,南京大学)
;
School of Artificial Intelligence, Nanjing University(人工智能学院,南京大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
The Chinese University of Hong Kong, Shenzhen(香港大学深圳校区)
;
Southern University of Science and Technology(南方科技大学)
So-Fake: Benchmarking and Explaining Social Media Image Forgery Detection
So-Fake:社交媒体图像伪造检测的基准构建与可解释性研究
Zhenglin Huang, Xiangtai Li, Xi Yang, Bei Peng, Xiaowei Huang, Baoyuan Wu, Dacheng Tao, Ming-Hsuan Yang, Guangliang Cheng
机构
*
University of Liverpool, UK(利物浦大学)
;
Nanyang Technological University(南洋理工大学)
;
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
The University of Sheffield(谢菲尔德大学)
LLM-Grounded Dynamic Task Planning with Hierarchical Temporal Logic for Human-Aware Multi-Robot Handover
基于分层时序逻辑的LLM引导动态任务规划用于人感知多机器人协作
Shuyuan Hu, Tao Lin, Kai Ye, Tianwei Zhang
机构
*
The Shenzhen Institute of Artificial Intelligence and Robotics for Society(深圳人工智能与机器人社会研究院)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
The Chinese University of Hong Kong-Shenzhen(香港中文大学(深圳))
;
Tsinghua Shenzhen International Graduate School(清华大学深圳国际 Graduate School)
机构
*
Department of Information Engineering, The Chinese University of Hong Kong(香港中文大学信息工程系)
;
John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院)
;
California Institute of Technology(加州理工学院)
polyDAG: Polynomial Acyclicity Constraints for Efficient Continuous Causal Discovery in Visual Semantic Graphs
polyDAG:用于视觉语义图中高效连续因果发现的多项式无环性约束
Wenhao Zhang, Ramin Ramezani, Tao Han, Kai Hwang, Minyi Guo
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
Fudan University(复旦大学)
;
The Chinese University of Hong Kong(香港中文大学)
PrinciplismQA: A Philosophy-Grounded Approach to Assessing LLM-Human Clinical Medical Ethics Alignment
PrinciplismQA: 一种基于哲学的评估LLM与人类临床医学伦理对齐的方法
Chang Hong, Minghao Wu, Qingying Xiao, Yuchi Wang, Xiang Wan, Guangjun Yu, Benyou Wang, Yan Hu
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
National Health Data Institute, Shenzhen(深圳国家健康数据研究院)
;
Shenzhen Research Institute of Big Data(深圳大数据研究院)
机构
*
Zhejiang University, China(浙江大学)
;
University of California, Berkeley, USA(加州大学伯克利分校)
;
MMLab, Chinese University of Hong Kong, China(香港中文大学MMLab)