Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information
通过关键信息的分步注意力混合层蒸馏提升小模型的推理能力
Yao Chen, Jiawei Sheng, Wenyuan Zhang, Tingwen Liu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
TopFeaRe: Locating Critical State of Adversarial Resilience for Graphs Regarding Topology-Feature Entanglement
TopFeaRe: 在拓扑-特征纠缠视角下定位图的对抗鲁棒性临界状态
Xinxin Fan, Wenxiong Chen, Quanliang Jing, Chi Lin, Shaoye Luo, Wenbo Song, Yunfeng Lu
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Dalian University of Technology(大连理工大学)
;
Beihang University(北京航空航天大学)
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos
SurgMotion: 一种用于外科视频通用理解的视频原生基础模型
Jinlin Wu, Felix Holm, Chuxi Chen, An Wang, Yaxin Hu, Xiaofan Ye, Zelin Zang, Miao Xu, Lihua Zhou, Huai Liao, Danny T. M. Chan, Ming Feng, Wai S. Poon, Hongliang Ren, Dong Yi, Nassir Navab, Gaofeng Meng, Jiebo Luo, Hongbin Liu, Zhen Lei
机构
*
Center for Artificial Intelligence and Robotics, Hong Kong Institute of Science and Innovation, Chinese Academy of Sciences, Hong Kong, China(人工智能与机器人中心,香港科学与创新研究院,中国科学院,香港,中国)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences, Beijing, China(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院,北京,中国)
;
Computer Aided Medical Procedures, Technical University of Munich, Munich, Germany(医学辅助程序,慕尼黑技术大学,德国慕尼黑)
;
Electronic Engineering Department, The Chinese University of Hong Kong, Hong Kong, China(电子工程系,香港中文大学,香港,中国)
;
Neuromedical Centre, Hong Kong University Shenzhen Hospital, Shenzhen, China(神经医学中心,香港大学深圳医院,深圳,中国)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
;
Department of Respiratory Medicine, The First Affiliated Hospital of Sun Yat-sen University, Guangzhou, China(呼吸科,中山大学附属第一医院,广州,中国)
Revisiting Entropy Regularization: Adaptive Coefficient Unlocks Its Potential for LLM Reinforcement Learning
重新审视熵正则化:自适应系数解锁其在大语言模型强化学习中的潜力
Xiaoyun Zhang, Xiaojian Yuan, Di Huang, Wang You, Chen Hu, Jingqing Ruan, Ai Jian, Kejiang Chen, Xing Hu
机构
*
State Key Lab of Processors, Institute of Computing Technology, CAS(处理器国家重点实验室,计算技术研究所,中国科学院)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
StepFun Inc(StepFun公司)
机构
*
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
DSLAB, School of Information Science & Engineering, Lanzhou University(兰州大学信息科学与工程学院DSLAB)
;
School of Mathematical Science, Jiangsu University(江苏大学数学学院)
;
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(中国科学技术大学脑科学与智能感知教育部重点实验室)
;
Agricultural Information Institute, Chinese Academy of Agricultural Sciences(中国农业科学院农业信息研究所)
;
Key Laboratory of Agricultural Big Data, Ministry of Agriculture and Rural Affairs(农业农村部农业大数据重点实验室)
An Empirical Study of Validating Synthetic Data for Text-Based Person Retrieval
基于文本的人检索中合成数据验证的实证研究
Min Cao, Yuxin Lu, Ziyin Zeng, Dong Yi, Jinqiao Wang, Mang Ye
机构
*
School of Computer Science and Technology, Soochow University(苏州大学计算机科学与技术学院)
;
Frontis
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心)
;
School of Artifcial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Wuhan AI Research, Wuhan, China(武汉人工智能研究所,武汉,中国)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
From Limited Labels to Open Domains:An Efficient Learning Method for Drone-view Geo-Localization
从有限标签到开放领域:一种用于无人机视角地理定位的高效学习方法
Zhongwei Chen, Zhao-Xu Yang, Hai-Jun Rong, Jiawei Lang, Guoqi Li
机构
*
State Key Laboratory for Strength and Vibration of Mechanical Structures(强度与振动机械结构国家重点实验室)
;
Shaanxi Key Laboratory of Environment and Control for Flight Vehicle(飞行器环境与控制陕西省重点实验室)
;
School of Aerospace Engineering(航空工程学院)
;
Xi’an Jiaotong University(西安交通大学)
;
Institute of Automation(自动化研究所)
;
Chinese Academy of Sciences(中国科学院)
;
School of Artificial Intelligence(人工智能学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Peng Cheng Laboratory(鹏城实验室)
机构
*
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS部)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Meituan(美团)
;
The Hong Kong Polytechnic University(香港理工大学)
;
CAIR, HKISI, CAS(中国科学院CAS HKISI CAIR)
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
Technology and Engineering Center for Space Utilization, Chinese Academy of Sciences(中国科学院空间利用技术与工程中心)
Geoparsing: Diagram Parsing for Plane and Solid Geometry with a Unified Formal Language
几何解析:一种统一形式语言用于平面和立体几何的图表解析
Peijie Wang, Ming-Liang Zhang, Jun Cao, Chao Deng, Dekang Ran, Hongda Sun, Pi Bu, Xuan Zhang, Yingyao Wang, Jun Song, Bo Zheng, Fei Yin, Cheng-Lin Liu
机构
*
MAIS, Institute of Automation of Chinese Academy of Sciences(中国科学院自动化研究所MAIS)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Future Living Lab of Alibaba(阿里巴巴未来生活实验室)
Yao Chen, Yilong Chen, Yinqi Yang, Junyuan Shang, Zhenyu Zhang, Zefeng Zhang, Shuaiyi Nie, Shuohuan Wang, Yu Sun, Hua Wu, HaiFeng Wang, Tingwen Liu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Baidu Inc.(百度公司)
Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality
按需语言,知识为核心:通过编码器-解码器翻译模型组成LLM以实现可扩展的多语言性
Mengyu Bu, Yang Feng
机构
*
Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences(智能信息处理重点实验室,计算技术研究所,中国科学院)
;
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
All Changes May Have Invariant Principles: Improving Ever-Shifting Harmful Meme Detection via Design Concept Reproduction
所有变化可能都有不变原则:通过设计概念再现改进永动有害迷因检测
Ziyou Jiang, Mingyang Li, Junjie Wang, Yuekai Huang, Jie Huang, Zhiyuan Chang, Zhaoyang Li, Qing Wang
机构
*
State Key Laboratory of Complex System Modeling and Simulation Technology(复杂系统建模与仿真技术国家重点实验室)
;
Science and Technology on Integrated Information System Laboratory Institute of Software Chinese Academy of Sciences(软件研究所信息集成系统技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval
跌入陷阱,获得智慧:通过误判风险模式检索的认知引导有害迷因检测
Wenshuo Wang, Ziyou Jiang, Junjie Wang, Mingyang Li, Jie Huang, Yuekai Huang, Zhiyuan Chang, Feiyan Duan, Qing Wang
机构
*
State Key Laboratory of Complex System Modeling and Simulation Technology(复杂系统建模与仿真技术国家重点实验室)
;
Science and Technology on Integrated Information System Laboratory(集成信息系统技术研究所)
;
Institute of Software Chinese Academy of Sciences(中国科学院软件研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space
从P(y|x)到P(y):在预训练空间中研究强化学习
Yuqiao Tan, Minzheng Wang, Bo Liu, Zichen Liu, Tian Liang, Shizhu He, Jun Zhao, Kang Liu
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
National University of Singapore(新加坡国立大学)
;
Tencent AI Lab(腾讯AI实验室)
ExpSeek: Self-Triggered Experience Seeking for Web Agents
ExpSeek:面向Web代理的自触发经验寻求
Wenyuan Zhang, Xinghua Zhang, Haiyang Yu, Shuaiyi Nie, Bingli Wu, Juwei Yue, Tingwen Liu, Yongbin Li
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Tongyi Lab , Alibaba Group(阿里云实验室,阿里巴巴集团)
机构
*
School of Informatics, Xiamen University(厦门大学信息学院)
;
The Hong Kong Polytechnic University(香港理工大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Alibaba Group(阿里巴巴集团)
;
Key Laboratory of Digital Protection and Intelligent Processing of Intangible Cultural Heritage of Fujian and Taiwan (Xiamen University), Ministry of Culture and Tourism, China(福建省和台湾非物质文化遗产数字化保护与智能处理重点实验室(厦门大学),中华人民共和国文化和旅游部,中国)
机构
*
National Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institution of Automation, Chinese Academy of Sciences(认知与复杂系统决策智能国家实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Beijing National Research Center for Information Science and Technology, Tsinghua University(北京信息科学与技术国家研究中心,清华大学)
FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing
FABLE:面向无结构模型编辑的细粒度事实锚定
Peng Wang, Biyu Zhou, Xuehai Tang, Jizhong Han, Songlin Hu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络与信息安全学院)
WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering
WikiSeeker: 重新思考视觉语言模型在基于知识的视觉问答中的作用
Yingjian Zhu, Xinming Wang, Kun Ding, Ying Wang, Bin Fan, Shiming Xiang
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统国家重点实验室 (MAIS))
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning
CamReasoner:通过结构化空间推理强化相机运动理解
Hang Wu, Yujun Cai, Zehao Li, Haonan Ge, Bowen Sun, Junsong Yuan, Yiwei Wang
机构
*
University of California, Merced(加州大学梅尔德分校)
;
University of Queensland(昆士兰大学)
;
Ant Group(蚂蚁集团)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
University at Buffalo, State University of New York(纽约州立大学布法罗分校)
Uncovering and Aligning Anomalous Attention Heads to Defend Against NLP Backdoor Attacks
揭示并对齐异常注意力头以防御NLP后门攻击
Haotian Jin, Yang Li, Haihui Fan, Lin Shen, Xiangfang Li, Bo Li
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
State Key Laboratory of Cyberspace Security Defense(网络空间安全防御国家重点实验室)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
From Attenuation to Attention: Variational Information Flow Manipulation for Fine-Grained Visual Perception
从衰减到注意:用于细粒度视觉感知的变分信息流操控
Jilong Zhu, Yang Feng
机构
*
Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences (ICT/CAS)(智能信息处理重点实验室,计算技术研究所,中国科学院(ICT/CAS))
;
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences (ICT/CAS)(人工智能安全国家重点实验室,计算技术研究所,中国科学院(ICT/CAS))
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
SEATrack: Simple, Efficient, and Adaptive Multimodal Tracker
SEATrack: 简单、高效且自适应的多模态跟踪器
Junbin Su, Ziteng Xue, Shihui Zhang, Kun Chen, Weiming Hu, Zhipeng Zhang
机构
*
School of Artificial Intelligence (School of Software), Yanshan University(燕山大学人工智能学院(软件学院))
;
School of Software, Beihang University(北航软件学院)
;
Hebei Key Laboratory of Computer Virtual Technology and System Integration(河北省计算机虚拟技术与系统集成重点实验室)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),CASIA)
;
School of Artificial Intelligence, UCAS(中国科学院大学人工智能学院)
;
AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院自动化实验室)