arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Massachusetts Institute of Technology(麻省理工学院)

2026-08-07 至 2026-08-07 共收录 4
2607.28617 2026-08-07 cs.AI cs.CL cs.CY cs.HC 版本更新

AISPA: User-Centric System Prompt Auditing for Large Language Model Applications

AISPA:面向大语言模型应用的以用户为中心的系统提示审计框架

Xiangning Lin, Shenzhe Zhu, Shu Yang, Zhenyu Zhang, Haoqian Zhang, Yipeng Zhao, Chengxuan Qian, Tianwei Wang, Ziheng Zhang, Zhenlong Yuan, Dingcheng Wang, Juncheng Wu, Yuan Si, Jiaxin Liu, Baolong Bi, Robert Mahari, Tobin South, Dazza Greenwood, Zexue He, Rishi Bommasani, Sophia Kazinnik, Andreas Haupt, Samuele Marro, Erik Brynjolfsson, Alex Pentland, Jiaxin Pei

机构 * Stanford University(斯坦福大学) CMU(卡内基梅隆大学) UT Austin(德克萨斯大学奥斯汀分校) University of Toronto(多伦多大学) UCSB(加利福尼亚大学圣巴巴拉分校) WashU(华盛顿大学) OSU(俄亥俄州立大学) UCSC(加利福尼亚大学圣克鲁兹分校) Northwestern University(西北大学) UIUC(伊利诺伊大学厄巴纳-香槟分校) KAUST(阿卜杜拉国王科技大学) MIT(麻省理工学院) University of Oxford(牛津大学) Institute for Decentralized AI(去中心化人工智能研究所)

AI总结 本文提出以用户为中心的AISPA框架,审计88款商业AI产品的3249条系统提示指令,发现其设计差异大、保护指令范围浅、长度增长但仍存问题指令,凸显系统提示需更高透明度与监督。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.01170 2026-08-07 cs.LG cs.AI cs.CL stat.AP stat.ML 版本更新

Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning

在线推理校准:测试时训练使可泛化的符合性大语言模型推理

Cai Zhou, Zekai Wang, Menghua Wu, Qianyu Julie Zhu, Flora C. Shi, Chenyu Wang, Ashia Wilson, Tommi Jaakkola, Stephen Bates

机构 * Department of Electrical Engineering and Computer Science (MIT EECS)(麻省理工学院电气工程与计算机科学系) Computer Science and Artificial Intelligence Laboratory (MIT CSAIL)(麻省理工学院计算机科学与人工智能实验室) Laboratory for Information and Decision Systems (MIT LIDS)(麻省理工学院信息与决策系统实验室) Computational Science and Engineering (MIT CSE)(麻省理工学院计算科学与工程) Massachusetts Institute of Technology(麻省理工学院)

AI总结 本文提出ORCA框架,通过测试时训练和符合性预测校准推理过程,提升大语言模型在分布变化下的效率和泛化能力,实验证明其在不同任务中具有更高的性能。

Comments Published as a conference paper at COLM 2026; 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21124 2026-08-07 cs.SD cs.AI eess.AS 版本更新

PhaseCoder: Microphone Geometry-Agnostic Spatial Audio Understanding for Multimodal LLMs

PhaseCoder: 无关麦克风几何的多模态大语言模型空间音频理解

Artem Dementyev, Wazeer Zulfikar, Sinan Hersek, Pascal Getreuer, Anurag Kumar, Vivek Kumar

机构 * Google DeepMind(谷歌DeepMind) Media Lab, MIT(媒体实验室,麻省理工学院) Google AR(谷歌AR)

AI总结 PhaseCoder是一种无需麦克风几何结构的Transformer空间音频编码器,能够生成空间嵌入并使LLM实现复杂空间推理和转录任务。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03631 2026-08-07 cs.LG math-ph math.MP nlin.CD q-bio.NC 版本更新

Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations

混沌系统的科学机器学习发现神经群体的支配方程

Anthony G. Chesebro, David Hofmann, Vaibhav Dixit, Earl K. Miller, Richard H. Granger, Alan Edelman, Christopher V. Rackauckas, Lilianne R. Mujica-Parodi, Helmut H. Strey

机构 * Department of Biomedical Engineering and Laufer Center for Physical and Quantitative Biology, State University of New York at Stony Brook, NY, USA(生物医学工程系和物理与定量生物学拉夫中心,石溪大学纽约州立大学 Stony Brook 分校) Athinoula A. Martinos Center for Biomedical Imaging, Massachusetts General Hospital and Harvard Medical School, MA, USA(Athinoula A. Martinos 生物医学成像中心,麻省总医院和哈佛医学院) Computer Science and Artificial Intelligence Lab, Massachusetts Institute of Technology, MA, USA(计算机科学与人工智能实验室,麻省理工学院) Picower Institute for Learning and Memory, Massachusetts Institute of Technology, MA, USA(记忆学习研究所,麻省理工学院) Psychological and Brain Sciences, Dartmouth College, NH, USA(心理学与脑科学系,达特茅斯学院) Santa Fe Institute, NM, USA(圣菲研究所)

AI总结 该研究提出PEM-UDE方法,通过科学机器学习从混沌系统中发现可解释的支配方程,成功恢复噪声污染数据中的动态,并在神经群体中推导出符合生物约束的新方程。

Comments 54 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏