arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Chinese Academy of Sciences(中国科学院大学)

共收录 1952
2406.11290 2026-04-14 cs.IR cs.AI cs.CL cs.LG

An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs

一种受哲学相关性启发的迭代效用判断框架

Hengran Zhang, Keping Bi, Jiafeng Guo, Xueqi Cheng

机构 * State Key Laboratory of AI Safety(人工智能安全国家重点实验室) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 本文提出ITEM框架,通过迭代提升信息检索中相关性排序、效用判断和答案生成的性能,实验表明在多个数据集上均优于基线模型。

Comments Accepted to ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09532 2026-04-13 cs.CV cs.AI

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

见仁见智:在标签噪声下鲁棒的视觉引导跨模态提示学习

Zibin Geng, Xuefeng Jiang, Jia Li, Zheng Li, Tian Wen, Lvhua Wu, Sheng Sun, Yuwei Wang, Min Liu

机构 * Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) University of Chinese Academy of Sciences(中国科学院大学) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) PCALab, VCIP, College of Computer Science, Nankai University(南开大学计算机学院PCALab, VCIP)

AI总结 本文提出VisPrompt框架,通过跨模态注意力机制将视觉语义注入提示表示,提升在标签噪声下的鲁棒性,实验表明其在多个数据集上表现优于现有基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09474 2026-04-13 cs.RO cs.AI

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

SafeMind: 一种考虑风险的可微控制框架用于自适应和安全的四足运动

Zukun Zhang, Kai Shu, Mingqiao Mo

机构 * The University of Hong Kong(香港大学) Alibaba(阿里巴巴) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 SafeMind通过整合概率控制障碍函数与语义上下文理解及元适应性风险校准,实现了在不确定环境下自适应且安全的四足运动控制,实验显示其在安全性与能耗上均优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09442 2026-04-13 cs.CL

UIPress: Bringing Optical Token Compression to UI-to-Code Generation

UIPress:将光学令牌压缩引入UI到代码生成

Dasen Dai, Shuoqi Li, Ronghao Chen, Huacan Wang, Biao Wu, Qizhen Lan

机构 * The Chinese University of Hong Kong(香港中文大学) Peking University(北京大学) University of Chinese Academy of Sciences(中国科学院大学) University of Technology Sydney(悉尼科技大学)

AI总结 本文提出UIPress,一种轻量级学习压缩模块,通过结合深度可分离卷积、元素引导的空间重加权和Transformer细化,将UI截图的视觉令牌从约6700压缩到256,结合LoRA实现表示差距桥接,提升效率并优于现有方法。

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09324 2026-04-13 cs.CV

Structure-Aware Fine-Grained Gaussian Splatting for Expressive Avatar Reconstruction

具有结构意识的细粒度高斯点散布用于表现力Avatar重建

Yuze Su, Hongsong Wang, Jie Gui, Liang Wang

机构 * School of Cyber Science and Engineering, Southeast University(东南大学网络空间安全学院) School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education(教育部新一代人工智能技术及其跨学科应用重点实验室(东南大学)) Purple Mountain Laboratories(紫金山实验室) Engineering Research Center of Blockchain Application, Supervision And Management (Southeast University), Ministry of Education(教育部区块链应用监管工程研究中心(东南大学)) State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS)(多模态人工智能系统全国重点实验室) Institute of Automation, Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

AI总结 本文提出SFGS方法,通过空间三平面和时间六平面捕捉动态特征,结合结构感知高斯模块和残差细化模块,实现高保真全身体素Avatar重建,优于现有方法。

Comments The code is on Github: https://github.com/Su245811YZ/SFGS

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09308 2026-04-13 cs.AI

Constraint-Aware Corrective Memory for Language-Based Drug Discovery Agents

基于约束的纠正记忆:用于基于语言的药物发现代理

Maochen Sun, Youzhi Zhang, Gaofeng Meng

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉科学学院)

AI总结 本文提出CACM框架,通过精确的集合级诊断和简洁的记忆写回机制,提升语言基药物发现代理的可靠性,实验表明其在目标级成功率提升36.4%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07720 2026-04-13 cs.AI

Towards Knowledgeable Deep Research: Framework and Benchmark

走向知识导向的深度研究:框架与基准

Wenxuan Liu, Zixuan Li, Long Bai, Chunmao Zhang, Fenghui Zhang, Zhuo Chen, Wei Li, Yuxin Zuo, Fei Wang, Bingbing Xu, Xuhui Jiang, Jin Zhang, Xiaolong Jin, Jiafeng Guo, Tat-Seng Chua, Xueqi Cheng

机构 * State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所人工智能安全国家重点实验室) National University of Singapore(新加坡国立大学) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 本文提出知识导向深度研究(KDR)任务,设计混合知识分析框架(HKA)并构建KDR-Bench基准,通过结构化与非结构化知识生成多模态报告,实验表明HKA在通用和知识导向指标上优于现有方法,且在视觉增强指标上超越Gemini DR代理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08920 2026-04-13 cs.IR cs.AI cs.CL cs.LG

Beyond Relevance: Utility-Centric Retrieval in the LLM Era

超越相关性:在大语言模型时代以效用为中心的检索

Hengran Zhang, Minghao Tang, Keping Bi, Jiafeng Guo

机构 * State Key Laboratory of AI Safety, ICT, CAS(中国科学院计算技术研究所人工智能安全国家重点实验室) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 本文探讨了在大语言模型时代,检索系统从以相关性为中心转向以效用为中心的转变,分析了效用与LLM需求及代理RAG的联系,提供设计检索系统的理论基础和实践指导。

Comments Accepted by SIGIR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08597 2026-04-13 cs.DB cs.AI

STIndex: A Context-Aware Multi-Dimensional Spatiotemporal Information Extraction System

STIndex:一种基于上下文的多维时空信息提取系统

Wenxiao Zhang, Yu Liu, Qiang sun, Yihao Ding, Sirui Li, Yanbing Liu, Jin B. Hong, Wei Liu

机构 * The University of Western Australia(西澳大利亚大学) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) University of Chinese Academy of Sciences(中国科学院大学) Murdoch University(莫道克大学)

AI总结 STIndex通过多维时空数据仓库提升无结构数据的提取效率,结合大语言模型实现上下文感知提取,提升时空实体识别精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08256 2026-04-13 cs.CL cs.AI

HyperMem: Hypergraph Memory for Long-Term Conversations

HyperMem:用于长期对话的超图记忆

Juwei Yue, Chuanrui Hu, Jiawei Sheng, Zuyi Zhou, Wenyuan Zhang, Tingwen Liu, Li Guo, Yafeng Deng

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院) EverMind AI

AI总结 HyperMem通过超图结构建模高阶关联,提升长期对话中的记忆检索效率与准确性,实验表明其在LoCoMo基准上达到92.73%的LLM-as-a-judge准确率。

Comments ACL 2026 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04736 2026-04-13 cs.AI cs.AR cs.PL

ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning

ChipSeek: 通过集成EDA的强化学习优化Verilog生成

Zhirong Chen, Kaiyan Chang, Zhuolin Li, Cangyuan Li, Xinyang He, Chujie Chen, Mengdi Wang, Haobo Xu, Yinhe Han, Huawei Li, Ying Wang

机构 * SKLP, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所智能计算机研究中心) University of Chinese Academy of Sciences(中国科学院大学) University of Electronic Science and Technology of China(电子科技大学) Beijing Institute of Technology(北京理工大学)

AI总结 ChipSeek通过集成EDA的强化学习框架,优化RTL代码生成,提升功能正确性和PPA指标。

Comments Accepted by ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08363 2026-04-10 cs.SD

CapTalk: Unified Voice Design for Single-Utterance and Dialogue Speech Generation

CapTalk:单语句和对话语音生成的统一语音设计

Xiaosu Su, Zihan Sun, Peilei Jia, Jun Gao

机构 * University of Chinese Academy of Sciences(中国科学院大学) Hello Group Inc.

AI总结 CapTalk提出一种统一的文本-音频自回归框架,实现单语句和对话语音设计,通过分层变分条件模块平衡语音稳定性和上下文适应性,提升表达可控性和上下文恰当性。

Comments 14 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08147 2026-04-10 cs.SD cs.CV

Semantic Noise Reduction via Teacher-Guided Dual-Path Audio-Visual Representation Learning

通过教师引导的双路径实现语义噪声削减

Linge Wang, Yingying Chen, Bingke Zhu, Lu Zhou, Jinqiao Wang

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Objecteye Inc.(北京眼神科技有限公司)

AI总结 本文提出TG-DP框架,通过分离重建与对齐路径,减少语义噪声,提升音频视频表示学习效果,实现零样本检索性能提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08088 2026-04-10 cs.CV

Coordinate-Based Dual-Constrained Autoregressive Motion Generation

基于坐标双约束的自回归运动生成

Kang Ding, Hongsong Wang, Jie Gui, Liang Wang

机构 * School of Cyber Science and Engineering, Southeast University(东南大学网络空间安全学院) School of Computer Science and Engineering, Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications, Ministry of Education, Southeast University(东南大学计算机科学与工程学院、教育部新一代人工智能技术及其跨学科应用重点实验室) Purple Mountain Laboratories(紫金山实验室) Engineering Research Center of Blockchain Application, Supervision And Management (Southeast University), Ministry of Education(教育部区块链应用监管工程研究中心(东南大学)) State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统全国重点实验室) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

AI总结 本文提出CDAMD框架,通过坐标输入和双约束因果掩码提升文本到运动的生成质量与语义一致性。

Comments Code is available at: https://github.com/fly-dk/CDAMD

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07728 2026-04-10 cs.CV cs.GR cs.RO

GEAR: GEometry-motion Alternating Refinement for Articulated Object Modeling with Gaussian Splatting

GEAR: 基于高斯点划法的几何-运动交替优化用于关节物体建模

Jialin Li, Bin Fu, Ruiping Wang, Xilin Chen

机构 * Key Laboratory of AI Safety of Chinese Academy of Sciences (CAS), Institute of Computing Technology, CAS(中国科学院人工智能安全重点实验室,中国科学院计算技术研究所) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 GEAR通过交替优化几何与运动,提升关节物体重建的精度和泛化能力,尤其在复杂多关节物体上表现优异。

Comments Accepted to CVPRF2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07664 2026-04-10 cs.CV eess.IV

Monocular Depth Estimation From the Perspective of Feature Restoration: A Diffusion Enhanced Depth Restoration Approach

单目深度估计:从特征修复的角度出发:一种扩散增强的深度修复方法

Huibin Bai, Shuai Li, Hanxiao Zhai, Yanbo Gao, Chong Lv, Yibo Wang, Haipeng Ping, Wei Hua, Xingyu Gao

机构 * School of Software, Shandong University(山东大学软件学院) Shandong University-WeiHai Research Institute of Industrial Technology(山东大学威海工业技术研究院) School of Control Science and Engineering, Shandong University(山东大学控制科学与工程学院) Shandong Institute of Information Technology Industry Development(山东省信息技术产业发展研究院) Research Institute of Interdisciplinary Innovation, Zhejiang Lab(之江实验室交叉创新研究院) Institute of Microelectronics, Chinese Academy of Sciences(中国科学院微电子研究所) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 本文从特征修复角度提出扩散增强深度修复方法,通过改进编码器特征提升深度估计性能,在KITTI基准上取得显著提升。

Comments Accepted by IEEE TMM

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06950 2026-04-10 cs.CV

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation

让 MLLMs 失明:MLLM 内容审核中的对抗走私攻击

Zhiheng Li, Zongyang Ma, Yuntong Pan, Ziqi Zhang, Xiaolei Lv, Bo Li, Jun Gao, Jianing Zhang, Chunfeng Yuan, Bing Li, Weiming Hu

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(中国科学院自动化研究所多模态人工智能系统国家重点实验室) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京市多模态信息超智能安全重点实验室) Hellogroup University of Washington(华盛顿大学) Jilin University(吉林大学) ShanghaiTech University(上海科技大学)

AI总结 本文揭示了MLLM在内容审核中面临的新威胁——对抗走私攻击,通过SmuggleBench基准测试发现多种模型易受攻击,提出通过感知和推理角度分析其根本原因并探索缓解策略。

Comments Accepted to ACL 2026. 19 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06747 2026-04-10 cs.AI

TurboAgent: An LLM-Driven Autonomous Multi-Agent Framework for Turbomachinery Aerodynamic Design

TurboAgent:一种基于大语言模型的自主多智能体框架用于涡轮机械气动设计

Juan Du, Yueteng Wu, Pan Zhao, Yuze Liu, Min Zhang, Xiaobin Xu, Xinglong Zhang

机构 * Digital Twin Research Center, Institute of Engineering Thermophysics, Chinese Academy of Sciences(中国科学院工程热物理研究所数字孪生研究中心) National Key Laboratory of Science and Technology on Advanced Light-duty Gas-Turbine, Chinese Academy of Sciences(中国科学院先进轻型燃气轮机科学与技术国家重点实验室) University of Chinese Academy of Sciences(中国科学院大学) The University of Hong Kong(香港大学)

AI总结 本文提出TurboAgent框架,通过大语言模型实现涡轮机械气动设计的端到端自主设计,结合生成设计、性能预测、多目标优化和物理验证,实现数据驱动的协作流程,提升设计效率和精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04666 2026-04-10 cs.AI cs.CR

Know Thy Enemy: Securing LLMs Against Prompt Injection via Diverse Data Synthesis and Instruction-Level Chain-of-Thought Learning

知彼者智:通过多样化数据合成和指令级推理学习增强大语言模型抗提示注入能力

Zhiyuan Chang, Mingyang Li, Yuekai Huang, Ziyou Jiang, Xiaojun Jia, Qian Xiong, Junjie Wang, Zhaoyang Li, Qing Wang

机构 * State Key Laboratory of Complex System Modeling and Simulation Technology(复杂系统建模与仿真技术国家重点实验室) Science and Technology on Integrated Information System Laboratory, Institute of Software Chinese Academy of Sciences(中国科学院软件研究所综合信息系统技术国家级重点实验室) University of Chinese Academy of Sciences(中国科学院大学) Nanyang Technological University(南洋理工大学) Beijing Forestry University(北京林业大学)

AI总结 本文提出InstruCoT方法,通过多样化数据合成和指令级推理学习提升大语言模型对提示注入攻击的防御能力,实验表明其在行为偏差、隐私泄露和有害输出方面显著优于基线方法。

Comments 19 pages, 6 figures; accepted by ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11548 2026-04-10 cs.CR cs.AI

One Shot Dominance: Knowledge Poisoning Attack on Retrieval-Augmented Generation Systems

单文档支配:面向检索增强生成系统的知识污染攻击

Zhiyuan Chang, Mingyang Li, Xiaojun Jia, Junjie Wang, Yuekai Huang, Ziyou Jiang, Yang Liu, Qing Wang

机构 * State Key Laboratory of Intelligent Game(智能游戏国家重点实验室) Science and Technology on Integrated Information System Laboratory, Institute of Software Chinese Academy of Sciences(中国科学院软件研究所综合信息系统技术国家级重点实验室) University of Chinese Academy of Sciences(中国科学院大学) Nanyang Technological University(南洋理工大学)

AI总结 本文提出AuthChain攻击方法,通过单文档污染实现对检索增强生成系统更有效的知识污染攻击,提升攻击成功率并保持隐蔽性。

Comments 15pages, 4 figures; accepted by EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20718 2026-04-10 cs.CV cs.AI

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

MM-MoralBench: 一种多模态道德评估基准用于大型视觉-语言模型

Bei Yan, Jie Zhang, Zhiyuan Chen, Shiguang Shan, Xilin Chen

机构 * University of Chinese Academy of Sciences(中国科学院大学)

AI总结 为评估大型视觉-语言模型的道德对齐性,提出MM-MoralBench基准,通过多模态场景测试模型在六个道德基础上的判断与分类能力,揭示模型存在显著的道德偏见问题。

Comments Accepted by Pattern Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06954 2026-04-09 cs.CV

Compression as an Adversarial Amplifier Through Decision Space Reduction

压缩作为决策空间缩减中的对抗放大器

Lewis Evans, Harkrishan Jandu, Zihan Ye, Yang Lu, Shreyank N Gowda

机构 * University of Nottingham(诺丁汉大学) University of Chinese Academy of Sciences(中国科学院大学) Xiamen University(厦门大学) School of Computer Science, University of Nottingham(诺丁汉大学计算机科学学院)

AI总结 研究发现压缩可增强深度图像分类器的对抗鲁棒性,通过决策空间缩减机制,压缩导致分类边距收缩并增加扰动敏感性,实验验证了压缩循环部署中的关键漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06746 2026-04-09 cs.CL

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference

StructKV: 保持结构骨架以实现可扩展的长上下文推理

Zhirui Chen, Peiyang Liu, Ling Shao

机构 * UCAS-Terminus AI Lab, University of Chinese Academy of Sciences, China(中国科学院大学UCAS-Terminus AI实验室) National Engineering Research Center for Software Engineering, Peking University(北京大学国家软件工程研究中心)

AI总结 StructKV通过全局入度中心性、动态枢轴检测和结构传播与解耦方法,有效保留长距离依赖性和检索鲁棒性,提升长上下文推理效率。

Comments Accepted to ACL 2026 Findings, 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06696 2026-04-09 cs.AI

AgentGate: A Lightweight Structured Routing Engine for the Internet of Agents

AgentGate: 一种轻量级的结构化路由引擎用于物联网代理

Yujun Cheng, Enfang Cui, Hao Qin, Zhiyuan Liang, Qi Xu

机构 * School of Artificial Intelligence, University of Science and Technology Beijing(北京科技大学人工智能学院) China Telecom Research Institute(中国电信研究院) Hangzhou Institute for Advanced Study, University of Chinese Academy of Sciences(中国科学院大学杭州高等研究院)

AI总结 本文提出AgentGate,一种轻量级结构化路由引擎,用于在资源受限条件下高效且隐私友好的代理系统,通过分阶段决策和结构化接地提升路由效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06685 2026-04-09 cs.CL cs.AI

ChemVLR: Prioritizing Reasoning in Perception for Chemical Vision-Language Understanding

ChemVLR:在化学视觉语言理解中优先考虑推理

Xuanle Zhao, Xinyuan Cai, Xiang Cheng, Xiuyi Chen, Bo Xu

机构 * The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所复杂系统认知与决策智能重点实验室) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

AI总结 ChemVLR通过细化化学描述符识别,提升化学视觉语言模型的推理能力,实现更清晰的推理路径,实验显示其在分子和反应任务中达到SOTA性能。

Comments Accepted by ACL 2026 Findings, Preprint Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.06662 2026-04-09 cs.CV cs.LG

Towards Robust Content Watermarking Against Removal and Forgery Attacks

面向对抗性移除与伪造攻击的鲁棒内容水印技术

Yifan Zhu, Yihan Wang, Xiao-Shan Gao

机构 * Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院) University of Chinese Academy of Sciences(中国科学院大学) University of Waterloo(滑铁卢大学)

AI总结 本文提出Instance-Specific watermarking with Two-Sided detection方法,通过动态控制水印注入时间和模式提升鲁棒性,有效对抗移除和伪造攻击。

Comments 14 pages, 5 figures, CVPR 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03128 2026-04-09 cs.LG cs.CL

Self-Distilled RLVR

自蒸馏强化学习验证奖励

Chenxu Yang, Chuanyu Qin, Qingyi Si, Minghui Chen, Naibin Gu, Dingyu Yao, Zheng Lin, Weiping Wang, Jiaqi Wang, Nan Duan

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)

AI总结 本文提出RLSD方法,结合自蒸馏和RLVR,通过自蒸馏获取token级策略差异以确定细粒度更新幅度,同时利用RLVR获取环境反馈的可靠更新方向,提升收敛性和稳定性。

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.29441 2026-04-09 cs.CV

EarthEmbeddingExplorer: A Web Application for Cross-Modal Retrieval of Global Satellite Images

地球嵌入探索器:一种用于全球卫星图像跨模态检索的Web应用

Yijie Zheng, Weijie Wu, Bingyue Wu, Long Zhao, Guoqing Li, Mikolaj Czerkawski, Konstantin Klemmer

机构 * Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院空天信息创新研究院) University of Chinese Academy of Sciences(中国科学院大学) Institute of Geographic Sciences and Natural Resources Research, Chinese Academy of Sciences(中国科学院地理科学与资源研究所) Asterisk Labs(Asterisk实验室) LGND AI, Inc.(LGND AI公司) University College London(伦敦大学学院)

AI总结 本文介绍EarthEmbeddingExplorer,一种Web应用,通过跨模态查询实现全球卫星图像的动态检索,帮助研究人员将研究成果转化为实际应用。

Comments ICLR 2026 Workshop ML4RS Tutorial Track (oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20049 2026-04-09 cs.CV

Data-Free Class-Incremental Gesture Recognition with Prototype-Guided Pseudo Feature Replay

无数据的类增量手势识别与原型引导的伪特征重放

Hongsong Wang, Ao Sun, Jie Gui, Liang Wang

机构 * School of Computer Science and Engineering, Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications, Ministry of Education, Southeast University(东南大学计算机科学与工程学院、教育部新一代人工智能技术及其跨学科应用重点实验室) School of Cyber Science and Engineering, Engineering Research Center of Blockchain Application, Supervision And Management (Southeast University), Ministry of Education, Southeast University(东南大学网络空间安全学院、教育部区块链应用监管工程研究中心(东南大学)) New Laboratory of Pattern Recognition (NLPR), State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所模式识别国家重点实验室、多模态人工智能系统全国重点实验室) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

AI总结 本文提出原型引导的伪特征重放框架,解决无数据的类增量手势识别问题,通过伪特征生成、旧类原型重放、新类截断交叉熵和持续分类器重训练,提升模型鲁棒性和泛化能力。

Comments Code is on https://github.com/sunao-101/PGPFR-3/

Journal ref IEEE Transactions on Image Processing, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.11789 2026-04-09 cs.LG cs.DC cs.SY eess.SY math.PR

Decentralized Online Learning for Random Inverse Problems Over Graphs

基于图的随机逆问题的去中心化在线学习

Xiwei Zhang, Tao Li, Yan Chen, Qianyuan Long

机构 * No.2 High School of East China Normal University(华东师范大学第二附属中学) School of Mathematical Sciences, East China Normal University(华东师范大学数学科学学院) Key Laboratory of Management, Decision and Information Systems, Institute of Systems Science, Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院系统科学研究所管理、决策与信息系统重点实验室) School of Mathematical Sciences, University of Chinese Academy of Sciences(中国科学院大学数学科学学院)

AI总结 本文提出了一种去中心化在线学习算法,用于网络图上的分布式随机逆问题,统一了Hilbert空间中的分布式参数估计与RKHS中的最小均方问题,并证明了算法在满足持久激励条件下的渐近稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏