arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Huazhong University of Science and Technology(华中科技大学)

2026-05-14 至 2026-05-14 共收录 6
2605.13757 2026-05-14 cs.RO

FrameSkip: Learning from Fewer but More Informative Frames in VLA Training

FrameSkip: 从更信息丰富的较少帧中学习以在VLA训练中提升性能

Bin Yu, Shijie Lian, Xiaopeng Lin, Zhaolong Shen, Yuliang Wei, Changti Wu, Hang Yuan, Haishan Liu, Bailing Wang, Cong Huang, Kai Chen

机构 * Harbin Institute of Technology(哈尔滨理工大学) Zhongguancun Academy(中关村学院) Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院) Huazhong University of Science and Technology(华中科技大学) East China Normal University(华东师范大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Beihang University(北航) DeepCybo

AI总结 本文提出FrameSkip框架,通过选择高重要性帧来优化VLA训练,提升成功率与保留率的平衡,实现在三个基准测试中达到76.15%的成功率。

Comments GitHub: https://github.com/ZGC-EmbodyAI/FrameSkip

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.13473 2026-05-14 cs.LG cs.CL

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention

OSDN:通过可证明的在线预条件化改进Delta规则

Chenyu Zhou, Hongpei Li, Yuerou Liu, Jianghao Lin, Dongdong Ge, Yinyu Ye

机构 * Shanghai Jiao Tong University(上海交通大学) Northwestern University(西北大学) Huazhong University of Science and Technology(华中科技大学) Stanford University(斯坦福大学)

AI总结 OSDN通过在线预条件化改进Delta规则,通过在线更新对角预条件器提升特征曲率处理,理论证明其收敛性和泛化能力,在大规模参数下提升上下文回忆性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.13153 2026-05-14 cs.AI

Strikingness-Aware Evaluation for Temporal Knowledge Graph Reasoning

面向时间知识图谱推理的显著性感知评估

Rikui Huang, Shengzhe Zhang, Wei Wei

机构 * School of Computer Science & Technology, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院) Institute of Artificial Intelligence, Huazhong University of Science and Technology(华中科技大学人工智能研究院) School of Artificial Intelligence & Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)

AI总结 本文提出显著性感知评估框架,通过引入基于规则的显著性测量框架量化事件显著性,改进时间知识图谱推理的评估方法,实验表明不同模型在不同显著性事件上的表现差异。

Comments Accepted to IJCAI-ECAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.11679 2026-05-14 cs.AI

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion

通过偏好维度扩展解释并打破安全与有益性天花板

ShiYing Huang, Liang Lin, Yuer Li, Kaiwen Luo, Zhenhong Zhou, An Zhang, Junhao Dong, Kun Wang, Zhigang Zeng

机构 * Huazhong University of Science and Technology(华中科技大学) Nanyang Technological University(南洋理工大学) Tsinghua University(清华大学) Chongqing University(重庆大学)

AI总结 本文提出MORA方法,通过多维奖励整合解决多目标对齐中的安全与有益性矛盾,实验显示在序列对齐中提升5%-12.4%,同时对齐中整体奖励提升4.6%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06387 2026-05-14 cs.LG cs.AI

Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level

非对称在线策略蒸馏:在令牌层面弥合利用与模仿

Nan Jia, Haojin Yang, Xing Ma, Jiesong Lian, Shuailiang Zhang, Weipeng Zhang, Ke Zeng, Xunliang Cai, Zequn Sun

机构 * Huazhong University of Science and Technology(华中科技大学) Peking University(北京大学) Meituan(美团)

AI总结 本文提出非对称在线策略蒸馏,通过局部散度最小化改进传统策略蒸馏,提升数学推理任务表现,实现更稳定的策略熵和持续工具使用能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00029 2026-05-14 cs.LG cs.AI

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

LoRA-Mixer:通过序列注意力路由协调模块化LoRA专家

Wenbing Li, Zikai Song, Hang Zhou, Yunyao Zhang, Junqing Yu, Wei Yang

机构 * Huazhong University of Science and Technology(华中科技大学)

AI总结 LoRA-Mixer通过将任务特定的LoRA专家路由到注意力模块的核心投影矩阵中,实现细粒度的token级专业化,同时保持与Transformer和状态空间模型的兼容性,且在15个基准测试中表现优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏