机构
*
Lanzhou University(兰州大学)
;
National University of Singapore(新加坡国立大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Tsinghua University(清华大学)
;
University of New South Wales(新南威尔士大学)
Large Language Model Aided Birt-Hogg-Dube Syndrome Diagnosis with Multimodal Retrieval-Augmented Generation
基于多模态检索增强生成的大型语言模型辅助Birt-Hogg-Dube综合征诊断
Haoqing Li, Jun Shi, Xianmeng Chen, Qiwei Jia, Rui Wang, Wei Wei, Hong An, Xiaowen Hu
机构
*
School of Computer Science and Technology(计算机科学与技术学院)
;
Department of Pulmonary and Critical Care Medicine(呼吸与危重症医学科)
;
Center for Diagnosis and Management of Rare Diseases(罕见病诊断与管理中心)
;
the First Affiliated Hospital of USTC(USTC第一附属医院)
;
Division of Life Sciences and Medicine(生命科学与医学系)
;
USTC
;
WanNan Medical College(皖南医学院)
机构
*
KOKONI, Moxin (Huzhou) Tech. Co., LTD(摩西(湖州)科技有限公司)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
;
School of Information Engineering, Huzhou University(湖州大学信息工程学院)
;
School of Electrical and Electronic Engineering, Nanyang Technological University(新加坡南洋理工大学电子与电气工程学院)
;
College of Computing and Data Science, Nanyang Technological University(新加坡南洋理工大学计算与数据科学学院)
;
School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)
COLI: A Hierarchical Efficient Compressor for Large Images
COLI:一种用于大图像的分层高效压缩器
Haoran Wang, Hanyu Pei, Yang Lyu, Kai Zhang, Li Li, Feng-Lei Fan
机构
*
Frontier of Artificial Networks (FAN) Lab, Department of Data Science, City University of Hong Kong(前沿人工智能网络实验室,数据科学系,香港城市大学)
;
Molecular Imaging Business Unit, Shanghai United Imaging Healthcare Co., Ltd(分子影像业务部,上海联合影像医疗科技股份有限公司)
;
MoE Key Laboratory of Brain-Inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知关键实验室,中国科学技术大学)
A Unified Multi-Dynamics Framework for Perception-Oriented Modeling in Tendon-Driven Continuum Robots
为腱驱动连续机器人感知导向建模设计的统一多动态框架
Ibrahim Alsarraj, Yuhao Wang, Abdalla Swikir, Cesare Stefanini, Dezhen Song, Zhanchi Wang, Ke Wu
机构
*
Robotics Department, Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(Mohamed bin Zayed大学人工智能研究院机器人部门)
;
Hefei National Research Center for Physical Sciences at the Microscale, University of Science and Technology of China (USTC)(中国科学技术大学微尺度物理科学国家级研究中心)
机构
*
Department of Computer Science and Engineering, State University of New York at Buffalo(纽约州立大学布法罗分校计算机科学与工程系)
;
Division of CEMSE, King Abdullah University of Science and Technology(卡普兰大学科学与技术大学CEMSE分校)
;
Institute of Artificial Intelligence and Blockchain, Guangzhou University(广州大学人工智能与区块链研究院)
;
School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)
Fine-Grained GRPO for Precise Preference Alignment in Flow Models
细粒度GRPO用于流模型中的精确偏好对齐
Yujie Zhou, Pengyang Ling, Jiazi Bu, Yibin Wang, Yuhang Zang, Jiaqi Wang, Li Niu, Guangtao Zhai
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Fudan University(复旦大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Shanghai Innovation Institute(上海创新研究院)
Understanding Fine-tuning in Approximate Unlearning: A Theoretical Perspective
理解近似反向学习中的微调:一种理论视角
Meng Ding, Rohan Sharma, Changyou Chen, Jinhui Xu, Kaiyi Ji
机构
*
Department of Computer Science and Engineering(计算机科学与工程系)
;
State University of New York at Buffalo(纽约州立大学布法罗分校)
;
School of Information Science and Technology(信息科学与技术学院)
;
University of Science and Technology of China(中国科学技术大学)
Llama2Vec: Unsupervised Adaptation of Large Language Models for Dense Retrieval
Llama2Vec: 无监督适应大型语言模型用于密集检索
Zheng Liu, Chaofan Li, Shitao Xiao, Yingxia Shao, Defu Lian
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong Polytechnic University(香港理工大学)
REArtGS: Reconstructing and Generating Articulated Objects via 3D Gaussian Splatting with Geometric and Motion Constraints
REArtGS: 通过3D高斯散点与几何和运动约束进行姿态物体的重建与生成
Di Wu, Liu Liu, Zhou Linli, Anran Huang, Liangtu Song, Qiaojun Yu, Qi Wu, Cewu Lu
机构
*
Hefei Institutes of Physical Science Chinese Academy of Sciences(合肥物理研究所中国科学院)
;
University of Science and Technology of China(中国科学技术大学)
;
Hefei University of Technology(合肥工业大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
ByteDance(字节跳动)
机构
*
School of Biomedical Engineering, Division of Life Sciences and Medicine, University of Science and Technology of China(中国科学技术大学生物医学工程学院)
;
Division of Life Science, Department of Chemical and Biological Engineering, State Key Laboratory of Nervous System Disorders, The Hong Kong University of Science and Technology(香港科技大学生命科学系)
;
SIAT-HKUST Joint Laboratory of Cell Evolution and Digital Health, HKUST Shenzhen-Hong Kong Collaborative Innovation Research Institute(深圳-香港联合创新研究院细胞进化与数字健康联合实验室)
;
Department of Hepatobiliary Surgery, The First Affiliated Hospital of USTC, Division of Life Sciences and Medicine, University of Science and Technology of China(中国科学技术大学附属第一医院肝胆外科)
;
Center for Medical Imaging, Robotics, Analytic Computing & Learning (MIRACLE), Suzhou Institute for Advanced Research, USTC, Suzhou, Jiangsu, China(中国科学技术大学苏州先进研究所)
;
Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Jiangsu Provincial Key Laboratory of Multimodal Digital Twin Technology, Suzhou, Jiangsu, China(江苏省多模态数字孪生技术重点实验室)
;
Key Laboratory of Precision and Intelligent Chemistry, USTC, Hefei, Anhui, China(中国科学技术大学精准与智能化学重点实验室)
;
Department of Pathology, Chinese PLA General Hospital, Beijing, China(中国人民解放军总医院病理科)
VividFace: High-Quality and Efficient One-Step Diffusion For Video Face Enhancement
VividFace: 高质量和高效的一步扩散用于视频面部增强
Shulian Zhang, Yong Guo, Long Peng, Ziyang Wang, Ye Chen, Wenbo Li, Xiao Zhang, Yulun Zhang, Jian Chen
机构
*
South China University of Technology(华南理工大学)
;
Max Planck Institute for Informatics(马克斯·普朗克研究所(信息学))
;
University of Science and Technology of China(中国科学技术大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Nanjing University of Science and Technology(南京理工大学)
;
Shanghai Jiao Tong University(上海交通大学)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
National Key Laboratory of Deep Space Exploration, Deep Space Exploration Laboratory(国家空间科学探测重点实验室,深空探测实验室)
;
Sangfor Technologies
Text2Loc++: Generalizing 3D Point Cloud Localization from Natural Language
Yan Xia, Letian Shi, Yilin Di, Joao F. Henriques, Daniel Cremers
机构
*
School of Artificial Intelligence and Data Science, University of Science and Technology of China(人工智能与数据科学学院,中国科学技术大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Visual Geometry Group, University of Oxford(牛津大学视觉几何组)
CommentsThis paper builds upon and extends our earlier conference paper Text2Loc presented at CVPR 2024