arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Stanford University(斯坦福大学)

共收录 2251
2510.24591 2025-11-25 cs.CL astro-ph.IM

ReplicationBench: Can AI Agents Replicate Astrophysics Research Papers?

ReplicationBench: 人工智能代理能否复制天体物理学研究论文?

Christine Ye, Sihan Yuan, Suchetha Cooray, Steven Dillmann, Ian L. V. Roque, Dalya Baron, Philipp Frank, Sergio Martin-Alvarez, Nolan Koblischke, Frank J Qu, Diyi Yang, Risa Wechsler, Ioana Ciuca

机构 * Stanford University(斯坦福大学) University of Toronto(多伦多大学)

AI总结 ReplicationBench旨在评估人工智能代理复制天体物理学研究论文的能力,通过测试代理在实验设置、推导、数据分析和代码库等任务上的表现,揭示代理在科学研究中的可靠性及挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17796 2025-11-25 cs.LG stat.ML

SING: SDE Inference via Natural Gradients

SING:通过自然梯度进行SDE推断

Amber Hu, Henry Smith, Scott Linderman

机构 * Stanford University(斯坦福大学)

AI总结 SING通过自然梯度变分推断提升隐式SDE模型中的状态推断和非线性漂移函数估计精度。

Comments To appear in Advances in Neural Processing Information Systems (NeurIPS), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18302 2025-11-25 cs.AI

The Catastrophic Paradox of Human Cognitive Frameworks in Large Language Model Evaluation: A Comprehensive Empirical Analysis of the CHC-LLM Incompatibility

人类认知框架在大语言模型评估中的灾难性悖论:对CHC-LLM不兼容性的全面实证分析

Mohan Reddy

机构 * Stanford University(斯坦福大学)

AI总结 本研究揭示了人类认知框架与大语言模型评估之间的不兼容性悖论,指出模型在结晶知识任务上的低准确率与高IQ评分的矛盾,提出发展原生机器认知评估框架的必要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17822 2025-11-25 cs.LG cs.DS stat.ML

High-Accuracy List-Decodable Mean Estimation

高精度可列表解码均值估计

Ziyun Chen, Spencer Compton, Daniel Kane, Jerry Li

机构 * University of Washington(华盛顿大学) Stanford University(斯坦福大学) University of California, San Diego(加州大学圣迭戈分校)

AI总结 本文提出高精度可列表解码学习方法,通过非平凡的理论和算法保证,实现身份协方差高斯分布均值的高精度估计。

Comments Abstract shortened to meet arXiv requirement

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17818 2025-11-25 cs.LG cs.AI

APRIL: Annotations for Policy evaluation with Reliable Inference from LLMs

APRIL:利用LLM可靠推断进行策略评估的标注

Aishwarya Mandyam, Kalyani Limaye, Barbara E. Engelhardt, Emily Alsentzer

机构 * Stanford University(斯坦福大学)

AI总结 APRIL利用大语言模型生成医疗领域的反事实注释,以提高离线策略评估的准确性与可扩展性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10059 2025-11-25 cs.RO

Ionospheric and Plasmaspheric Delay Characterization for Lunar Terrestrial GNSS Receivers with Global Core Plasma Model

月球地表GNSS接收机的电离层和等离子层延迟特性分析:基于全球核心等离子体模型

Keidai Iiyama, Grace Gao

机构 * Department of Aeronautics and Astronautics, Stanford University, California, United States(航空与宇航系,斯坦福大学,加州,美国)

AI总结 本文基于全球核心等离子体模型和定制射线追踪算法,研究了月球地表GNSS接收机的电离层和等离子层延迟特性,揭示了不同太阳和地磁条件下延迟的影响因素。

Comments Submitted NAVIGATION: Journal of the Institute of Navigation

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17475 2025-11-24 physics.flu-dyn cs.LG

Addressing A Posteriori Performance Degradation in Neural Network Subgrid Stress Models

应对神经网络子网格应力模型的后验性能退化

Andy Wu, Sanjiva K. Lele

机构 * Department of Aeronautics and Astronautics, Stanford University(航空与宇航系,斯坦福大学) Department of Mechanical Engineering, Stanford University(机械工程系,斯坦福大学)

AI总结 本文提出通过训练数据增强和减少输入复杂度来改善神经网络子网格应力模型的后验性能,使其更符合先验评估。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15946 2025-11-24 cs.CV

Automated Interpretable 2D Video Extraction from 3D Echocardiography

自动从3D超声心动图中提取可解释的2D视频

Milos Vukadinovic, Hirotaka Ieki, Yuki Sahashi, David Ouyang, Bryan He

机构 * University of California, Los Angeles(加州大学洛杉矶分校) Kaiser Permanente Division of Research(凯撒医疗集团研究部) Cedars-Sinai Medical Center(西达赛西医学中心) Stanford University(斯坦福大学)

AI总结 本文提出了一种自动从3D超声心动图中提取标准2D视频的方法,通过深度学习和医学专家启发式方法提高解读效率和准确性。

Comments 12 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17323 2025-11-24 cs.SD cs.AI cs.CL cs.MM

MusicAIR: A Multimodal AI Music Generation Framework Powered by an Algorithm-Driven Core

MusicAIR: 一种由算法驱动核心的多模态AI音乐生成框架

Callie C. Liao, Duoduo Liao, Ellie L. Zhang

机构 * Stanford University(斯坦福大学) George Mason University(乔治·马歇尔大学)

AI总结 MusicAIR通过算法驱动的核心生成音乐,能从歌词、文本和图像生成符合音乐理论的旋律谱,提升音乐创作效率并降低入门门槛。

Comments Accepted by IEEE Big Data 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16842 2025-11-24 cs.AI cs.CL cs.LG

Fantastic Bugs and Where to Find Them in AI Benchmarks

人工智能基准中的非凡虫子及其寻找方法

Sang Truong, Yuheng Tu, Michael Hardy, Anka Reuel, Zeyu Tang, Jirayu Burapacheep, Jonathan Perera, Chibuike Uwakwe, Ben Domingue, Nick Haber, Sanmi Koyejo

机构 * Stanford University(斯坦福大学)

AI总结 本文提出了一种基于统计分析和LLM审查的系统性基准修订框架,通过识别无效问题提升AI基准的可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16675 2025-11-24 cs.LG cs.AI

Joint Design of Protein Surface and Structure Using a Diffusion Bridge Model

利用扩散桥模型联合设计蛋白质表面和结构

Guanlue Li, Xufeng Zhao, Fang Wu, Sören Laue

机构 * University of Hamburg(汉堡大学) Stanford University(斯坦福大学)

AI总结 PepBridge通过整合受体表面几何和生化性质,联合设计蛋白质表面和结构,提升蛋白质设计的物理合理性和结构可行性。

Comments 21 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04079 2025-11-24 cs.CL

Improving the Performance of Radiology Report De-identification with Large-Scale Training and Benchmarking Against Cloud Vendor Methods

通过大规模训练和与云服务提供商方法的基准测试来改进放射学报告去标识化性能

Eva Prakash, Maayane Attias, Pierre Chambon, Justin Xu, Steven Truong, Jean-Benoit Delbrouck, Tessa Cook, Curtis Langlotz

机构 * Stanford University(斯坦福大学) JP Morgan Chase & Co(摩根大通公司) Sorbonne University(索邦大学) University of Oxford(牛津大学) NVIDIA(英伟达) HOPPR University of Pennsylvania(宾夕法尼亚大学)

AI总结 本文提出了一种基于变压器的去标识化模型,通过大规模训练和与商业系统的基准测试,实现了在放射学报告中更高效的PHI检测性能。

Comments In submission to JAMIA

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24068 2025-11-24 cs.LG cs.AI

A Small Math Model: Recasting Strategy Choice Theory in an LLM-Inspired Architecture

一个小数学模型:在LLM启发式架构中重构策略选择理论

Roussel Rahman, Jeff Shrager

机构 * Stanford University(斯坦福大学) Bennu Climate, Inc.(Bennu气候公司)

AI总结 本文提出一个基于LLM架构的小数学模型,重构策略选择理论,扩展其功能并研究数学推理中的数字特征与关系。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12972 2025-11-24 cs.CV cs.AI

Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs Reasoning

对齐视觉与语言:无需注释的多模态知识图谱构建以增强大语言模型推理

Junming Liu, Siyuan Meng, Yanting Gao, Song Mao, Pinlong Cai, Guohang Yan, Yirong Chen, Zilin Bian, Ding Wang, Botian Shi

机构 * Tongji University(同济大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) East China Normal University(华东师范大学) Stanford University(斯坦福大学) New York University(纽约大学)

AI总结 本文提出 VaLiK 方法,通过跨模态信息补充构建无需注释的多模态知识图谱,提升大语言模型推理能力。

Comments 14 pages, 7 figures, 6 tables; Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16635 2025-11-21 cs.CV cs.CL

SurvAgent: Hierarchical CoT-Enhanced Case Banking and Dichotomy-Based Multi-Agent System for Multimodal Survival Prediction

SurvAgent: 基于层次化CoT增强的案例库与二元多智能体系统用于多模态生存预测

Guolin Huang, Wenting Chen, Jiaqi Yang, Xinheng Lyu, Xiaoling Luo, Sen Yang, Xiaohan Xing, Linlin Shen

机构 * Shenzhen University(深圳大学) Stanford University(斯坦福大学) University of Nottingham Ningbo China(诺丁汉大学宁波分校) Ant Group(蚂蚁集团)

AI总结 SurvAgent通过层次化CoT增强的多智能体系统,整合多模态数据,提升生存预测的可解释性与准确性。

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16043 2025-11-21 cs.LG

Agent0: Unleashing Self-Evolving Agents from Zero Data via Tool-Integrated Reasoning

Agent0: 通过工具集成推理从零数据中释放自进化代理

Peng Xia, Kaide Zeng, Jiaqi Liu, Can Qin, Fang Wu, Yiyang Zhou, Caiming Xiong, Huaxiu Yao

机构 * Stanford University(斯坦福大学) Salesforce Research(Salesforce研究院)

AI总结 Agent0通过多步骤共进化和工具集成,在无需外部数据的情况下自主提升代理性能,显著增强推理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15921 2025-11-21 cs.AI

Thinking, Faithful and Stable: Mitigating Hallucinations in LLMs

思考、忠实与稳定:减轻大语言模型幻觉的方法

Chelsea Zou, Yiheng Yao, Basant Khalil

机构 * Stanford University(斯坦福大学)

AI总结 本研究提出一种通过强化学习策略减少大语言模型幻觉的方法,通过置信度对齐和熵激增信号提升推理的连贯性和准确性。

Comments Originally released June 5, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15898 2025-11-21 cs.LG

Global Resolution: Optimal Multi-Draft Speculative Sampling via Convex Minimization

全局分辨率:通过凸优化实现最优多草稿投机采样

Rahul Krishna Thomas, Arka Pal

机构 * Stanford University(斯坦福大学)

AI总结 本文提出了一种基于凸优化的多草稿投机采样算法,通过将最优传输问题转化为最大流问题,实现了高效的高接受率和低开销的生成过程。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15619 2025-11-20 cs.LG stat.ML

CODE: A global approach to ODE dynamics learning

CODE:一种全局方法用于ODE动力学学习

Nils Wildt, Daniel M. Tartakovsky, Sergey Oladyshkin, Wolfgang Nowak

机构 * Institute for Modelling Hydraulic and Environmental Systems(流体与环境系统建模研究所) University of Stuttgart(斯图加特大学) Department of Energy Science and Engineering(能源科学与工程系) Stanford University(斯坦福大学) Cluster of Excellence SimTech(卓越中心SimTech)

AI总结 CODE通过多项式混沌展开方法,实现对ODE动态的全局建模,展现出在稀疏数据和噪声下的强外推能力,优于神经网络和核近似器方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14889 2025-11-20 cs.LG

Bringing Federated Learning to Space

Grace Kim, Filip Svoboda, Nicholas Lane

机构 * Stanford University(斯坦福大学) University of Cambridge(剑桥大学)

Comments 15 pages, 9 figures, 3 tables accepted to IEEE Aeroconf 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14780 2025-11-20 cs.AI

Ask WhAI:Probing Belief Formation in Role-Primed LLM Agents

Keith Moore, Jun W. Kim, David Lyu, Jeffrey Heo, Ehsan Adeli

机构 * Department of Biomedical Data Science, Stanford University(生物医学数据科学系,斯坦福大学)

Comments Preprint. Accepted for publication at AIAS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14778 2025-11-20 cs.AI

Learning Interestingness in Automated Mathematical Theory Formation

George Tsoukalas, Rahul Saha, Amitayush Thakur, Sabrina Reguyal, Swarat Chaudhuri

机构 * UT Austin(得克萨斯大学) Princeton University(普林斯顿大学) Stanford University(斯坦福大学)

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00261 2025-11-20 cs.CV cs.HC

Spot The Ball: A Benchmark for Visual Social Inference

Neha Balamurugan, Sarah Wu, Adam Chun, Gabe Gaw, Cristobal Eyzaguirre, Tobias Gerstenberg

机构 * Stanford University(斯坦福大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14721 2025-11-19 cs.LG math.OC

AdamHD: Decoupled Huber Decay Regularization for Language Model Pre-Training

Fu-Ming Guo, Yingfang Fan

机构 * Stanford University(斯坦福大学) Harvard Medical School(哈佛医学院)

Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: GPU-Accelerated and Scalable Optimization (ScaleOpt)

Journal ref 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: GPU-Accelerated and Scalable Optimization (ScaleOpt)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10683 2025-11-19 cs.RO cs.AI cs.CV

MotIF: Motion Instruction Fine-tuning

Minyoung Hwang, Joey Hejna, Dorsa Sadigh, Yonatan Bisk

机构 * Massachusetts Institute of Technology(麻省理工学院) Stanford University(斯坦福大学) Carnegie Mellon University(卡内基梅隆大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03336 2025-11-19 cs.RO

Benchmarking Population-Based Reinforcement Learning across Robotic Tasks with GPU-Accelerated Simulation

Asad Ali Shahid, Yashraj Narang, Vincenzo Petrone, Enrico Ferrentino, Ankur Handa, Dieter Fox, Marco Pavone, Loris Roveda

机构 * Dalle Molle Institute for Artificial Intelligence, IDSIA USI-SUPSI(达摩克利斯人工智能研究所,IDSIA USI-SUPSI) NVIDIA Corporation(NVIDIA公司) University of Salerno(萨勒诺大学) Stanford University(斯坦福大学)

Comments Accepted for publication at 2025 IEEE 21st International Conference on Automation Science and Engineering

Journal ref 2025 IEEE 21st International Conference on Automation Science and Engineering (CASE), Los Angeles, CA, USA, 2025, pp. 1231-1238

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.16324 2025-11-19 cs.LG math.OC

Achieving Instance-dependent Sample Complexity for Constrained Markov Decision Process

Jiashuo Jiang, Yinyu Ye

机构 * Department of Industrial Engineering & Decision Analytics, Hong Kong University of Science and Technology(香港科技大学工业工程与决策分析系) Institute for Computational and Mathematical Engineering, Stanford University(斯坦福大学计算与数学工程研究所) Department of Management Science & Engineering, Stanford University(斯坦福大学管理科学与工程系)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14871 2025-11-19 cs.LG cs.CV

Squeezed Diffusion Models

Jyotirmai Singh, Samar Khanna, James Burgess

机构 * Stanford University(斯坦福大学)

Comments 7 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08377 2025-11-19 cs.LG stat.ML

Proofs as Explanations: Short Certificates for Reliable Predictions

Avrim Blum, Steve Hanneke, Chirag Pabbaraju, Donya Saless

机构 * Toyota Technological Institute at Chicago(芝加哥技术学院) Purdue University(普渡大学) Stanford University(斯坦福大学)

Comments Fixed Crefs, added reference to open question on tolerance Carathéodory, other minor changes

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13145 2025-11-18 cs.CV cs.AI

Automated Road Distress Detection Using Vision Transformersand Generative Adversarial Networks

Cesar Portocarrero Rodriguez, Laura Vandeweyen, Yosuke Yamamoto

机构 * Stanford University(斯坦福大学)

详情

展开后加载摘要…

URL PDF HTML 收藏