arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Washington(华盛顿大学)

2026-06-10 至 2026-06-10 共收录 7
2606.11169 2026-06-10 cs.DC cs.AI 新提交

Piper: A Programmable Distributed Training System

Piper: 可编程的分布式训练系统

Megan Frisella, Shubham Tiwari, Andy Ruan, Yi Pan, Parker Gustafson, Mat Jacob, Gilbert Bernstein, Stephanie Wang

机构 * University of Washington(华盛顿大学) University of Washington and Shanghai Jiao Tong University(华盛顿大学和上海交通大学)

AI总结 提出Piper系统,通过解耦策略与运行时实现,允许用户用少量注解和调度指令声明分布式训练策略,并编译为设备执行计划,支持常见策略并实现组合策略的联合调度优化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10887 2026-06-10 cs.CV 新提交

Listen, Look, and Learn: Learning Without Forgetting through SAM-Audio

听、看、学:通过SAM-Audio实现无遗忘学习

Avi Gupta, Nilotpal Sinha, Vishnu Raj, Sambuddha Saha, Pratik Joshi, Koteswar Rao Jerripothula, Tammam Tillo

机构 * University of Washington(华盛顿大学)

AI总结 提出一种利用SAM-Audio多模态先验的类增量学习方法,通过引导注意力机制和双层蒸馏策略,在音频-视觉场景中缓解灾难性遗忘,性能优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10302 2026-06-10 cs.CL 新提交

Where You Inject Diversity Matters: A Unified Framework for Diverse Generation

注入多样性的位置至关重要:统一框架下的多样化生成

Cheng Zhang, Rui Xin, Chudi Zhong

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校) University of Washington(华盛顿大学)

AI总结 提出统一框架,通过多样性源和传输分数衡量测试时多样化生成方法,并基于此提出全自动规范级方法,在五个开放任务中提升输出多样性且保持质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07936 2026-06-10 cs.CL cs.AI 版本更新

Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation

黄金标准的幻觉:长文本生成中人类评估协议的大规模分析

Katelyn Xiaoying Mei, Yi-Li Hsu, Minjoon Choi, Zongwan Cao, Chenjun Xu, Bingbing Wen, Su Lin Blodgett, Lucy Lu Wang

机构 * University of Washington(华盛顿大学) National Tsing Hua University(国立清华大学) Seoul National University(首尔大学) Mila - Québec AI Institute(米拉-魁北克人工智能研究所) Allen Institute for AI(艾伦人工智能研究所)

AI总结 通过分析2023-2025年*CL会议论文中的人类评估协议,发现报告不透明和可重复性差的问题,并提出改进建议。

Comments Accepted to ACL 2026 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.06744 2026-06-10 cs.LG cs.GT cs.MA econ.TH 版本更新

Learn to Match: Two-Sided Matching with Temporally Extended Feedback

学会匹配:具有时间扩展反馈的双边匹配

Haijing Zong, Yancheng Liang, Boyang Zhou, Natasha Jaques

机构 * Department of Economics, University of Washington(华盛顿大学经济系) Paul G. Allen School of Computer Science & Engineering, University of Washington(华盛顿大学保罗·G·艾伦计算机科学与工程学院)

AI总结 提出一个具有时间扩展反馈的双边匹配框架,将其建模为部分可观测马尔可夫博弈,并基于多智能体强化学习构建Learn2Match基准,实验表明独立PPO优于bandit基线,但存在信息摩擦损失。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17907 2026-06-10 cs.CL cs.AI 版本更新

Improving Topic Modeling by Distilling Soft Labels from Language Models

DSL-Topic:通过从语言模型中蒸馏软标签改进主题建模

Raymond Li, Amirhossein Abaskohi, Chuyuan Li, Gabriel Murray, Giuseppe Carenini

机构 * University of Washington(华盛顿大学)

AI总结 提出DSL框架,通过从语言模型蒸馏软标签来增强主题模型训练,利用上下文感知的软标签重构信号,显著提升主题连贯性和分配准确性。

Comments 22 pages, 5 figures. Camera-ready version for ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11517 2026-06-10 cs.CL cs.DC cs.LG

Learning to Keep a Promise: Scaling Language Model Decoding Parallelism with Learned Asynchronous Decoding

学习承诺:通过学习异步解码扩展语言模型解码并行性

Tian Jin, Ellie Y. Cheng, Zack Ankner, Nikunj Saunshi, Blake M. Elias, Amir Yazdanbakhsh, Jonathan Ragan-Kelley, Suvinay Subramanian, Michael Carbin

机构 * DeepMind, London, UK(深度思维公司,伦敦,英国) Google Research, New York, NY, USA(谷歌研究院,纽约,纽约州,美国) Stanford University, Stanford, CA, USA(斯坦福大学,斯坦福,加利福尼亚州,美国) University of Toronto, Toronto, Ontario, Canada(多伦多大学,多伦多,安大略省,加拿大) University of Washington, Seattle, WA, USA(华盛顿大学,西雅图,华盛顿州,美国)

AI总结 本文提出PASTA系统,通过学习使语言模型识别语义独立性,提升解码并行性,实验证明在解码速度和响应质量上优于现有方法。

Comments 15 pages

Journal ref Proceedings of the 42nd International Conference on Machine Learning (ICML), PMLR 267:27941-27956, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏