arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Peking University(北京大学)

2026-04-27 至 2026-04-27 共收录 11
2604.21724 2026-04-27 cs.CL

Beyond N-gram: Data-Aware X-GRAM Extraction for Efficient Embedding Parameter Scaling

超越N-gram:面向数据的X-GRAM提取用于高效的嵌入参数扩展

Yilong Chen, Yanxi Xie, Zitian Gao, He Xin, Yihao Xiao, Jason Klein Liu, Haoming Luo, Yifan Luo, Zhengmao Ye, Tingwen Liu, Xin Zhao, Ran Tao, Bryan Dai

机构 * Peking University(北京大学) IQuest Research(IQuest研究)

AI总结 本文提出X-GRAM框架,通过频率感知的动态token注入方法,提升嵌入参数扩展效率,实验表明在0.73B和1.15B规模下,X-GRAM在准确率上优于基线模型。

Comments 29 pages, 9 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00813 2026-04-27 cs.CV cs.AI cs.RO

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale

DVGT-2:面向大规模自动驾驶的视觉-几何-动作模型

Sicheng Zuo, Zixun Xie, Wenzhao Zheng, Shaoqing Xu, Fang Li, Hanbing Li, Long Chen, Zhi-Xin Yang, Jiwen Lu

机构 * Tsinghua University(清华大学) Xiaomi EV(小米电动车) University of Macau(澳门大学) Peking University(北京大学)

AI总结 本文提出VGA范式,通过密集3D几何信息提升自动驾驶决策,引入DVGT-2实现在线处理,结合时间因果注意力和滑动窗口策略,提升效率与重建性能。

Comments Code is available at https://github.com/wzzheng/DVGT

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22591 2026-04-27 cs.RO

RedVLA: Physical Red Teaming for Vision-Language-Action Models

RedVLA:面向视觉-语言-动作模型的物理红队测试

Yuhao Zhang, Borong Zhang, Jiaming Fan, Jiachen Shen, Yishuai Cai, Yaodong Yang, Jiaming Ji

机构 * Institute for AI, Peking University(人工智能研究院,北京大学) State Key Laboratory of General Artificial Intelligence, Peking University(通用人工智能国家重点实验室,北京大学)

AI总结 RedVLA通过两阶段流程系统揭示VLA模型中的不安全行为,通过风险场景合成和风险放大技术,在10次优化迭代内达到95.5%的检测准确率,提出轻量级安全防护方案SimpleVLA-Guard。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22548 2026-04-27 stat.AP cs.LG

Multi-output Extreme Spatial Model for Complex Aircraft Production Systems

多输出极值空间模型用于复杂飞机生产系统

Cheolhei Lee, Xing Wang, Xiaowei Yue, Jianguo Wu

机构 * Grado Department of Industrial and Systems Engineering, Virginia Polytechnic Institute and State University(弗吉尼亚理工学院和州立大学工业与系统工程系) Department of Mathematics, Illinois State University(伊利诺伊州立大学数学系) Department of Industrial Engineering, Tsinghua University(清华大学工业工程系) Department of Industrial Engineering and Management, Peking University(北京大学工业工程与管理系)

AI总结 本文提出多输出极值空间模型,用于复杂飞机生产系统中的极值预测与风险分析,通过双域线性函数捕捉动态,提升极端事件预测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22542 2026-04-27 cs.CL cs.AI

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

可控的口语对话生成:一种基于LLM的评分系统用于K-12非母语英语学习者

Haidong Yuan, Haokun Zhao, Wanshi Xu, Songjun Cao, Qingyu Zhou, Long Ma, Hongjie Fan

机构 * Peking University(北京大学) Tencent(腾讯) China University of Political Science and Law(中国政法大学) Independent Researcher(独立研究者) Fudan University(复旦大学)

AI总结 本文提出一种基于LLM的评分系统,通过四级分级系统控制词汇复杂度,提升非母语K-12英语学习者的口语对话质量与教学价值。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22363 2026-04-27 cs.RO cs.AI

LeHome: A Simulation Environment for Deformable Object Manipulation in Household Scenarios

LeHome:一种用于家庭场景中可变形物体操控的仿真环境

Zeyi Li, Yushi Yang, Shawn Xie, Kyle Xu, Tianxing Chen, Yuran Wang, Zhenhao Shen, Yan Shen, Yue Chen, Wenjun Li, Yukun Zheng, Chaorui Zhang, Siyi Lin, Fei Teng, Hongjun Yang, Ming Chen, Steve Xie, Ruihai Wu

机构 * Peking University(北京大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Lightwheel The University of Hong Kong(香港大学)

AI总结 LeHome旨在解决家庭场景中可变形物体操控的挑战,提供高保真动态和真实交互,支持多种机器人形态,聚焦低成本机器人,实现家庭任务的端到端评估。

Comments ICRA2026 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22273 2026-04-27 cs.AI

When Does LLM Self-Correction Help? A Control-Theoretic Markov Diagnostic and Verify-First Intervention

当大语言模型自我纠正有助于什么?一种控制论马尔可夫诊断和先验证干预

Aofan Liu, Jingxiang Meng

机构 * Peking University(北京大学) University of Chicago(芝加哥大学)

AI总结 本文通过控制论马尔可夫模型分析自我纠正的条件,发现近零误差阈值决定其有效性,并通过实验验证先验证提示能有效减少误差。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22229 2026-04-27 cs.LG cs.AI

Preserve Support, Not Correspondence: Dynamic Routing for Offline Reinforcement Learning

保留支持,而非对应:动态路由用于离线强化学习

Zhancun Mu, Guangyu Zhao, Yiwu Zhong, Chi Zhang

机构 * School of Intelligence Science and Technology, Peking University(北京理工大学智能科学与技术学院)

AI总结 本文提出DROL,一种基于潜在条件的一步演员,通过动态路由在保持数据支持区域的同时提升性能,实验证明其在OGBench和D4RL上的有效性。

Comments 17 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17801 2026-04-27 cs.CV

View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity

通过双路径结构对应和语义连续性实现视图一致的3D场景编辑

Pufan Li, Bi'an Du, Shenghe Zheng, Junyi Yao, Wei Hu

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学王萱计算机技术研究所) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系)

AI总结 本文提出了一种视图一致的3D场景编辑框架,通过引入双路径一致性机制和配对多视角编辑数据集,提升复杂场景的编辑性能。

Comments Preprint. 10 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10079 2026-04-27 cs.CL

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

为何监督微调失效:对大语言模型中不完全学习现象的系统研究

Chao Xue, Yao Wang, Mengqiao Liu, Di Liang, Xingsheng Han, Peiyang Liu, Xianjie Wu, Chenyao Lu, Lei Jiang, Yu Lu, Haibo Shi, Shuang Liang, Minlong Peng, Flora D. Salim

机构 * University of New South Wales(新南威尔士大学) Tencent Hunyuan(腾讯文言) Tencent Yuanbao(腾讯元宝) UESTC(电子科技大学) Peking University(北京大学)

AI总结 本文系统研究了大语言模型微调中不完全学习现象,揭示了五个导致学习不完整的原因,并提出诊断优先框架和缓解策略,证明监督微调的局限性。

Comments Accepted by ACL 2026 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.01907 2026-04-27 cs.CV cs.AI

Lifting Unlabeled Internet-level Data for 3D Scene Understanding

提升互联网级未标记数据用于3D场景理解

Yixin Chen, Yaowei Zhang, Huangyue Yu, Junchao He, Yan Wang, Jiangyong Huang, Hongyu Shen, Junfeng Ni, Shaofei Wang, Baoxiong Jia, Song-Chun Zhu, Siyuan Huang

机构 * State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI) Beijing University of Posts and Telecommunications(北京邮电大学) Peking University(北京大学) Beijing Institute of Technology(北京理工大学) Tsinghua University(清华大学)

AI总结 本文通过设计数据引擎利用互联网未标记视频生成训练数据,提升3D场景理解模型性能,验证了在不同感知粒度任务中的有效性。

Comments CVPR 2026. Project page: https://sv-pp.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏