arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Learning Representations · 会议 · Machine Learning

2026-04-13 至 2026-04-13 共收录 10
2604.09389 2026-04-13 cs.LG cs.CL

Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder

更多数据值得付出成本吗?在极小注意力解码器中的数据集扩展定律

Götz-Henrik Wiegand, Lorena Raichle, Rico Städeli, Tomas Hrycej, Bernhard Bermeitinger, Siegfried Handschuh

机构 * Institute of Computer Science, University of St. Gallen(圣加仑大学计算机科学研究所) Institute of Computer Science in Vorarlberg, University of St. Gallen(圣加仑大学福拉尔贝格计算机科学研究所)

AI总结 本文通过极小注意力解码器研究数据集规模扩展定律,发现数据量增加带来性能提升但存在边际效益递减,为数据与计算成本平衡提供指导。

Comments Presented as a paper at 3rd DATA-FM workshop @ ICLR 2026, Brazil. Published at 13th IEEE Swiss Conference on Data Science and AI (SDS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09369 2026-04-13 q-bio.BM cs.LG q-bio.QM

Biologically-Grounded Multi-Encoder Architectures as Developability Oracles for Antibody Design

基于生物原理的多编码架构作为抗体设计的可开发性预言机

Simon J. Crouzet

AI总结 本文提出CrossAbSense框架,通过结合冻结的蛋白质语言模型编码器和可配置的注意力解码器,提升抗体可开发性评估的准确性,显著改进了三种评估 assay 的性能。

Comments ICLR 2026 Workshop on Generative and Experimental Perspectives for Biomolecular Design

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.23045 2026-04-13 cs.AI

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

AI的混乱:模型智能与任务复杂性如何影响对齐问题

Alexander Hägele, Aryo Pradipta Gema, Henry Sleight, Ethan Perez, Jascha Sohl-Dickstein

机构 * Anthropic Fellows Program(Anthropic 研究员计划) EPFL(瑞士联邦理工学院洛桑) University of Edinburgh(爱丁堡大学) Constellation Anthropic

AI总结 研究探讨了高智能AI模型在复杂任务中失败的机制,发现模型规模越大,失败行为越不一致,强调对齐研究在防止奖励黑客和目标误指定中的重要性。

Comments ICLR 2026. 10 pages main text, 40 total, 27 figures. v2: typos, improved writing, references

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13792 2026-04-13 cs.CV

VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel Optimization

VisionLaw: 通过双层优化从视觉观测中推断可解释的内在动力学

Jiajing Lin, Shu Jiang, Qingyuan Zeng, Zhenzhong Wang, Min Jiang

机构 * School of Informatics, Xiamen University(厦门大学信息学院) Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究所) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 本文提出VisionLaw框架,通过双层优化从视觉数据中推断可解释的内在动力学,解决了传统方法在可解释性和泛化能力上的不足。

Comments Accepted by ICLR 2026; Project Page: https://github.com/JiajingLin/VisionLaw

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08772 2026-04-13 physics.ao-ph cs.LG

CERBERUS: A Three-Headed Decoder for Vertical Cloud Profiles

CERBERUS:垂直云廓线的三头解码器

Emily K. deJong, Nipun Gunawardena, Kevin Smalley, Hassan Beydoun, Peter Caldwell

机构 * Lawrence Livermore National Laboratory(劳伦斯利弗莫尔国家实验室)

AI总结 CERBERUS通过三头编码器-解码器架构,利用卫星温度、地面气象数据和时间信息生成垂直雷达反射率分布,提升云过程建模的准确性与不确定性估计。

Comments Accepted for oral presentation at 2026 ICLR workshop on Machine Learning for Remote Sensing

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08761 2026-04-13 cs.CV

State Space Models are Effective Sign Language Learners: Exploiting Phonological Compositionality for Vocabulary-Scale Recognition

状态空间模型是有效的手语学习者:利用语音学的组合性进行词汇级识别

Bryan Cheng, Austin Jin, Jasper Zhang

机构 * William A. Shine Great Neck South High School(威廉·A·夏因大颈南高中)

AI总结 本文提出PHONSSM模型,通过语音学分解提升手语识别性能,实现词汇级识别,优于现有方法,尤其在少样本场景下表现更佳。

Comments 8 pages, 3 figures. Accepted to workshop on Algorithmic Fairness Across Alignment Procedures and Agentic Systems at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08574 2026-04-13 cs.LG cs.AI

Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching

通过嵌入匹配高效表示mRNA的基因组模型压缩

Rasched Haidari, Sam Martin, Maxime Allard

机构 * Helical London, UK(Helical 伦敦,英国)

AI总结 本文提出通过嵌入匹配方法将先进基因组基础模型的mRNA表示压缩至更小模型,减少200倍,验证了嵌入级蒸馏优于logit方法,展示了压缩模型在mRNA-bench上的最优性能。

Comments Accepted at the Tiny Papers Track for the Machine Learning for Genomics Explorations Workshop at ICLR 2026 an the Gen2 Workshop at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08553 2026-04-13 cs.LG cs.AI cs.CL

GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN Feedback

GNN-as-Judge: 通过GNN反馈释放LLM在图学习中的潜力

Ruiyao Xu, Kaize Ding

机构 * Northwestern University(西北大学)

AI总结 本文提出GNN-as-Judge框架,通过结合GNN的结构归纳偏差,解决LLM在低资源条件下生成可靠伪标签和缓解标签噪声的问题,提升图学习性能。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04072 2026-04-13 cs.CL cs.AI

SkillFactory: Self-Distillation For Learning Cognitive Behaviors

SkillFactory:用于学习认知行为的自我蒸馏

Zayne Sprague, Jack Lu, Manya Wadhwa, Sedrick Keh, Mengye Ren, Greg Durrett

机构 * New York University(纽约大学) Toyota Research Institute(丰田研究所)

AI总结 SkillFactory通过监督微调阶段学习认知技能,为强化学习奠定基础,提升模型在复杂任务中的泛化能力与鲁棒性。

Comments Published at ICLR 2026; code at https://github.com/Zayne-sprague/SkillFactory

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24718 2026-04-13 cs.CV cs.LG

Generative View Stitching

生成视图拼接

Chonghyuk Song, Michal Stary, Boyuan Chen, George Kopanas, Vincent Sitzmann

机构 * MIT CSAIL(麻省理工学院计算机科学与人工智能实验室) Runway ML

AI总结 本文提出生成视图拼接(GVS),通过并行采样确保生成场景与预定义相机轨迹一致,解决视频生成中因未来条件缺失导致的碰撞问题,实现稳定且一致的相机引导视频生成。

Comments Published at ICLR 2026. Camera-ready Submission. Project website: https://andrewsonga.github.io/gvs

详情

展开后加载摘要…

URL PDF HTML 收藏