LRANet++: Low-Rank Approximation Network for Accurate and Efficient Text Spotting
LRANet++:用于准确高效文本定位的低秩近似网络
Yuchen Su, Zhineng Chen, Yongkun Du, Zuxuan Wu, Hongtao Xie, Yu-Gang Jiang
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
Institute of Trustworthy Embodied AI, College of Intelligent Robotics and Advanced Manufacturing, Fudan University(复旦大学可信具身人工智能研究院)
;
School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)
Reasoning Path Divergence: A New Metric and Curation Strategy to Unlock LLM Diverse Thinking
推理路径分歧:一种新的度量和筛选策略,以解锁LLM多样化思考
Feng Ju, Zeyu Qin, Rui Min, Zhitao He, Lingpeng Kong, Yi R. Fung
机构
*
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
University of Science and Technology of China(中国科学技术大学)
;
The University of Hong Kong(香港大学)
High-Precision Transformer-Based Visual Servoing for Humanoid Robots in Aligning Tiny Objects
高精度基于Transformer的视觉伺服控制用于人形机器人对齐微小物体
Jialong Xue, Wei Gao, Yu Wang, Chao Ji, Dongdong Zhao, Shi Yan, Shiwu Zhang
机构
*
Institute of Humanoid Robots, Department of Precision Machinery and Precision Instrumentation, University of Science and Technology of China(人形机器人研究院,精密机械与精密仪器系,中国科学技术大学)
机构
*
Unmanned System Research Institute at Northwestern Polytechnical University(西北工业大学无人系统研究院)
;
School of Electronic, Electrical, and Communication Engineering, University of Chinese Academic of Sciences(中国科学院大学电子电气与通信工程学院)
;
Max Planck Institute for Informatics (MPI-INF)(马克斯·普朗克研究所(信息研究所))
;
University of Science and Technology of China(中国科学技术大学)
OmniVaT: Single Domain Generalization for Multimodal Visual-Tactile Learning
OmniVaT:单域泛化用于多模态视觉-触觉学习
Liuxiang Qiu, Hui Da, Yuzhen Niu, Tiesong Zhao, Yang Cao, Zheng-Jun Zha
机构
*
Fujian Key Laboratory for Intelligent Processing and Wireless Transmission of Media Information(福建智能媒体信息处理与无线传输重点实验室)
;
College of Physics and Information Engineering(物理与信息工程学院)
;
Fuzhou University(福州市大学)
;
College of Computer and Data Science(计算机与数据科学学院)
;
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition(MoE脑启发智能感知与认知重点实验室)
;
University of Science and Technology of China(中国科学技术大学)
机构
*
University of Science and Technology of China(科学技术大学)
;
Hefei University of Technology(合肥工业大学)
;
Hefei Xiaosheng Intelligent Technology Co., Ltd.(合肥小生智能科技有限公司)
;
Cylingo Group(Cylingo集团)
;
The University of Electro-Communications(电通大学)
机构
*
University of New South Wales(新南威尔士大学)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
Feng-Qi Cui, Zhen Lin, Xinlong Rao, Anyang Tong, Shiyao Li, Fei Wang, Changlin Chen, Bin Liu
机构
*
University of Science and Technology of China(科学技术大学)
;
Hefei University of Technology(合肥工业大学)
;
IAI, Hefei Comprehensive National Science Center(IAI合肥国家科学中心)
机构
*
State Key Laboratory of Fire Science(火灾科学国家重点实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Fire Research Division(火灾研究部)
;
National Institute of Standards and Technology(美国国家标准与技术研究院)
Exploiting the Prior of Generative Time Series Imputation
利用生成时间序列填补的先验
YuYang Miao, Chang Li, Zehua Chen
机构
*
Imperial College London, Department of Electronic and Electrical Engineering(帝国理工学院伦敦校区电子与电气工程系)
;
University of Science and Technology of China(中国科学技术大学)
;
Tsinghua University, Department of CST(清华大学计算机科学与技术系)
;
Shengshu AI(盛数人工智能)
机构
*
Qing Yuan Research Institute, Shanghai Jiao Tong University(上海交通大学庆元研究院)
;
Shanghai Innovation Institute(上海创新研究院)
;
Department of Pathology, The First Affiliated Hospital of USTC, Division of Life Sciences and Medicine, University of Science and Technology of China(中国科学技术大学生命科学与医学学院病理科)
;
Intelligent Pathology Institute, Division of Life Sciences and Medicine(生命科学与医学学院智能病理研究所)
;
Department of Pathology, Fudan University Shanghai Cancer Center(复旦大学上海癌症中心病理科)
;
Department of Oncology, Shanghai Medical College, Fudan University(复旦大学上海医学院肿瘤科)
;
Institute of Pathology, Fudan University(复旦大学病理研究所)
;
Department of Pathology, The First Affiliated Hospital with Nanjing Medical University(南京医科大学第一附属医院病理科)
Replay Failures as Successes: Sample-Efficient Reinforcement Learning for Instruction Following
重播失败作为成功:用于指令跟随的样本高效强化学习
Kongcheng Zhang, Qi Yao, Shunyu Liu, Wenjian Zhang, Min Cen, Yang Zhou, Wenkai Fang, Yiru Zhao, Baisheng Lai, Mingli Song
机构
*
Zhejiang University(浙江大学)
;
Nanyang Technological University(南洋理工大学)
;
Dalian University of Technology(大连理工大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Alibaba Cloud Computing(阿里云计算)
;
Chinese Academy of Sciences(中国科学院)
AI总结
HiR 提出了一种样本高效强化学习框架,通过将失败尝试视为成功来提升指令跟随任务的性能,利用双偏好学习实现高效优化。
SymMaP: Improving Computational Efficiency in Linear Solvers through Symbolic Preconditioning
SymMaP:通过符号预条件化提高线性求解器的计算效率
Hong Wang, Jie Wang, Minghao Ma, Haoran Shao, Haoyang Liu
机构
*
University of Science and Technology of China(中国科学技术大学)
;
CAS Key Laboratory of Technology in GIPAS, University of Science and Technology of China(中国科学技术大学国家实验室(GIPAS技术))
机构
*
School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院)
;
Meituan Inc.(美团公司)
;
Institute of Advanced Technology, University of Science and Technology of China(中国科学技术大学先进技术研究院)
;
School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院)
;
Research Center for Space Computing System, Zhejiang Lab(浙江实验室空间计算系统研究中心)
;
College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院)
;
School of Robotics and Automation, Nanjing University(南京大学机器人与自动化学院)
;
College of Intelligent Robotics and Advanced Manufacturing, Fudan University(复旦大学智能机器人与先进制造学院)
Unleashing Foundation Vision Models: Adaptive Transfer for Diverse Data-Limited Scientific Domains
释放基础视觉模型:面向多样化数据受限科学领域的自适应迁移
Qiankun Li, Feng He, Huabao Chen, Xin Ning, Kun Wang, Zengfu Wang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
AnnLab, Institute of Semiconductors, Chinese Academy of Sciences(中国科学院半导体研究所)
;
Nanyang Technological University(南洋理工大学)
The CCF AATC 2025 Speech Restoration Challenge: A Retrospective
2025年CCF AATC语音恢复挑战:回顾
Junan Zhang, Mengyao Zhu, Xin Xu, Hui Bu, Zhenhua Ling, Zhizheng Wu
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Audio Department, Huawei CBG(华为CBG音频部门)
;
Beijing AISHELL Technology Co., Ltd.(北京艾斯HELL科技有限公司)
;
University of Science and Technology of China(中国科学技术大学)
GroupDebate: Enhancing the Efficiency of Multi-Agent Debate Using Group Discussion
GroupDebate: 通过群体讨论提升多智能体辩论效率
Tongxuan Liu, Xingyu Wang, Weizhe Huang, Wenjiang Xu, Yuting Zeng, Lei Jiang, Hailong Yang, Jing Li
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Beihang University(北航)
机构
*
DP Technology Beijing China(北京DP技术有限公司)
;
AI for Science Institute Beijing China(北京AI for Science研究院)
;
Shanghai Jiao Tong University Shanghai China(上海交通大学)
;
Beihang University Beijing China(北京航空航天大学)
;
Peking University Beijing China(北京大学)
;
Institute of Theoretical Physics Chinese Academy of Sciences Beijing China(中国科学院理论物理研究所)
;
Shanghai Innovation Institute Shanghai China(上海创新研究院)
;
East China Normal University Shanghai China(华东师范大学)
;
Zhongguancun Academy Beijing China(中关村学院)
;
University of Science and Technology of China Hefei China(中国科学技术大学)
;
Tongji University Shanghai China(同济大学)
;
The Hong Kong University of Science and Technology Hong Kong China(香港科技大学)
ActionFlow: A Pipelined Action Acceleration for Vision Language Models on Edge
ActionFlow:面向边缘设备的视觉语言模型流水线动作加速
Yuntao Dai, Hang Gu, Teng Wang, Qianyu Cheng, Yifei Zheng, Zhiyong Qiu, Lei Gong, Wenqi Lou, Xuehai Zhou
机构
*
School of Computer Science and Technology, University of Science and Technology of China(计算机科学与技术学院,中国科学技术大学)
;
Suzhou Institute for Advanced Research, University of Science and Technology of China(苏州市先进研究院,中国科学技术大学)
;
IEIT SYSTEMS Co., Ltd.(IEIT SYSTEMS公司)