ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents
ProvenanceGuard: 基于MCP的LLM智能体的源感知事实性验证
Ander Alvarez, Santhiya Rajan, Samuel Mugel, Román Orús
机构
*
Multiverse Computing
;
Parque Cientifico y Tecnológico de Gipuzkoa(吉普斯夸科技园)
;
Centre for Social Innovation(社会创新中心)
;
Donostia International Physics Center(多诺斯蒂亚国际物理中心)
;
Ikerbasque Foundation for Science(伊克尔巴斯克科学基金会)
机构
*
DeepWisdom(深智科技)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Renmin University of China(中国人民大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Agent Universe(智能体宇宙)
;
McGill University(麦吉尔大学)
;
Yale University(耶鲁大学)
Measuring AI Ability to Complete Long Software Tasks
衡量AI完成长期软件任务的能力
Thomas Kwa, Ben West, Joel Becker, Amy Deng, Katharyn Garcia, Max Hasin, Sami Jawhar, Megan Kinniment, Nate Rush, Sydney Von Arx, Ryan Bloom, Thomas Broadley, Haoxing Du, Brian Goodrich, Nikola Jurkovic, Luke Harold Miles, Seraphina Nix, Tao Lin, Chris Painter, Neev Parikh, David Rein, Lucas Jun Koba Sato, Hjalmar Wijk, Daniel M. Ziegler, Elizabeth Barnes, Lawrence Chan
机构
*
Model Evaluation & Threat Research (METR)(模型评估与威胁研究(METR))
;
Ohm Chip
;
Anthropic
Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-Based Retrieval Agents Improves Interpretable Claim Verification
混合事实核查:集成知识图谱、大语言模型和基于搜索的检索代理提高可解释的声明验证
Shaghayegh Kolli, Richard Rosenbaum, Timo Cavelius, Lasse Strothe, Andrii Lata, Jana Diesner
Comments16 pages, 3 figures, 7 tables. Substantially revised: reframed around an identifiability result for the counterbalanced concentration index, with access configuration presented as one generator of the failure rather than a separate confound. Code and data are included as ancillary files
AdaDINO: Context-Adaptive DINO-Distilled Vision Foundation Models for Efficient Open-Vocabulary Edge Inference
AdaVFM:通过LLM引导执行实现边缘智能的自适应视觉基础模型
Yiwei Zhao, Yi Zheng, Huapeng Su, Jieyu Lin, Stefano Ambrogio, Cijo Jose, Michael Ramamonjisoa, Patrick Labatut, Barbara De Salvo, Chiao Liu, Phillip B. Gibbons, Ziyun Li
机构
*
Institute for Clarity in Documentation(清晰文档研究所)
;
Inria Paris-Rocquencourt(巴黎-罗克Quantin 研究院)
;
Rajiv Gandhi University(拉朱·甘地大学)
;
Tsinghua University(清华大学)
;
Palmer Research Laboratories(帕勒尔研究实验室)
;
Shandong University(山东省大学)
;
City University of Hong Kong(香港城市大学)
CommentsWithdrawal due to some flaws in experimental methodology and unresolved ethical issues in data collection. We need to redesign the experiments and obtain proper ethical clearance before resubmission
Structure-Induced Information for Rerooting Levin Tree Search
结构信息用于重定根莱文树搜索
Jake Tuero, Michael Buro, Laurent Orseau, Levi H. S. Lelis
机构
*
Department of Computing Science, University of Alberta, Edmonton, Canada.
;
Alberta Machine Intelligence Institute (Amii), Edmonton, Canada.
;
Google DeepMind, London, United Kingdom.
Stochastic Resetting Accelerates Reinforcement Learning Beyond Random Search
随机重置加速强化学习中的策略收敛
Jello Zhou, David J. Schwab, Vudtiwat Ngampruetikorn
机构
*
Biophysics Program, Stanford University(斯坦福大学生物物理项目)
;
School of Physics, University of Sydney(悉尼大学物理学院)
;
National Institute for Theory and Mathematics in Biology, Northwestern University(生物理论与数学国家研究所,西北大学)
;
The University of Chicago(芝加哥大学)
;
Princeton-CUNY Center for the Physics of Biological Function, The Graduate Center, CUNY(普林斯顿-纽约大学生物物理功能中心,纽约大学研究生中心)
Towards a General Intelligence and Interface for Wearable Health Data
迈向可穿戴健康数据的通用智能与接口
Girish Narayanswamy, Maxwell A. Xu, A. Ali Heydari, Samy Abdel-Ghaffar, Marius Guerard, Kara Vaillancourt, Zhihan Zhang, Jake Garrison, Levi Albuquerque, Dimitris Spathis, Hong Yu, Hamid Palangi, Xuhai "Orson" Xu, David G. T. Barrett, Joseph Breda, Jed McGiffin, Yubin Kim, Yuwei Zhang, Naghmeh Rezaei, Samuel Solomon, Karan Ahuja, Tim Althoff, Jake Sunshine, Ming-Zher Poh, Benjamin Yetton, Ari Winbush, Nicholas B. Allen, James M. Rehg, Isaac Galatzer-Levy, Yun Liu, John Hernandez, Anupam Pathak, Conor Heneghan, Yuzhe Yang, Ahmed A. Metwally, Pushmeet Kohli, Mark Malhotra, Shwetak Patel, Xin Liu, Daniel McDuff
机构
*
Google Research(谷歌研究)
;
Google DeepMind(谷歌DeepMind)
;
University of Washington(华盛顿大学)
;
University of Oregon(俄勒冈大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)