arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 431 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 431 篇

2607.11175 2026-08-13 cs.AI 版本更新 70%

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy

自我进化临床系统之路:将医疗智能体从辅助扩展到自主

Chunzheng Zhu, Lei Tian, Bohan Tan, Ziqi Zhou, Yuxuan Sun, Yijun Wang, Chengchao Lv, Yilin Wen, Yijun He, Jinghao Lin, Yihang Chen, Chee Wei Tan, Qianshan Wei, Lei Zhao, Bin Pu, Kenli Li, Yuan Xue, Jianxin Lin

机构 * Hunan University(湖南大学) ByteDance(字节跳动) Duke University(杜克大学) Westlake University(西湖大学) The University of Hong Kong(香港大学) Nanyang Technological University(南洋理工大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) University of Macau(澳门大学) The Ohio State University(俄亥俄州立大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究探讨大语言模型等对医疗智能体的重塑,从临床部署出发,将其形式化为决策系统并给出自主性分类。沿统一框架扩展,强调临床环境扩展为关键方向,定位临床自我进化为前沿,还研究了多领域应用及挑战,提供医学成像系统路线图。

Comments Project page: https://github.com/zhcz328/Awesome-Medical-Agents

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24226 2026-08-12 cs.IR cs.LG 版本更新 70%

UniScale: Synergistic Entire Space Data and Model Scaling for Search Ranking

联合模型参数缩放与通用域数据集成用于电商搜索排序

Liren Yu, Caiyuan Li, Feiyi Dong, Tao Zhang, Zhixuan Zhang, Dan Ou, Haihong Tang, Bo Zheng

机构 * Taobao \& Tmall Group of Alibaba Hangzhou China Taobao \& Tmall Group of Alibaba Beijing China Taobao \& Tmall Group of Alibaba

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 提出UniScale框架,通过联合数据缩放(ES$^3$样本构建系统)和模型设计(HHSFT异构层次融合Transformer),解决工业搜索中仅增大模型参数或调整架构带来的性能瓶颈,在电商搜索平台上实现购买量提升1.70%、GMV提升2.04%。

Comments Accepted at CIKM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.07254 2026-08-11 cs.DL cs.AI 版本更新 70%

SCALE: Scientific Concept Aggregation via LLMs and Embeddings for Fine-Grained Taxonomy Extension

SCALE:基于大语言模型与嵌入的科学概念聚合,用于细粒度分类体系扩展

Daniele Raimondi, Feichi Lu, Oliver Grun, Mariia Eremina, Andrea Perlato

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 SCALE是结合LLMs与嵌入的框架,在OpenAlex分类体系主题下新增科学概念层,将语义相关关键词组织为概念单元,为细粒度学术分类等提供基础。

Comments 14 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.05563 2026-08-10 cs.CR cs.AI 版本更新 70%

When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems

当经验成为指令:自进化智能体技能系统中的轨迹中毒攻击

Jialuo Chen, Lingqi Jiang, Xinhao Deng, Xiaohu Du, Jianan Ma, Yunhao Feng, Yuqi Qing, Zhihao Yuan, Linkang Du, Jingyi Wang

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 本文提出PoisonedEvolution轨迹中毒攻击,针对自进化智能体技能系统的经验转指令过程,在SkillClaw等平台实现高成功率,证明了证据转化是这类系统的安全边界。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23397 2026-08-05 cs.AI cs.CR 版本更新 70%

A Unified Framework for Human AI Collaboration in Security Operations Centers with Trusted Autonomy

面向安全运营中心中人类与AI协作的统一框架:可信自主能力

Ahmad Mohsin, Helge Janicke, Ahmed Ibrahim, Iqbal H. Sarker, Seyit Camtepe

机构 * Centre for Securing Digital Futures, School of Science, Edith Cowan University(安全数字未来中心,科学学院,爱迪生·科文大学) CSIRO’s Data61(CSIRO的数据61)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 该研究提出一种整合AI自主、信任校准与人在回路决策的统一框架,适配SOC不同任务的复杂度与风险,通过AI-Avatar案例验证其可减少警报疲劳、增强响应协调,助力构建增强人类决策的下一代认知SOC。

Comments Accept

Journal ref ACM Transactions on Internet Technology July 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.18295 2026-08-04 cs.LG 版本更新 70%

On the Limits of Support-Preserving Alignment and Bounded Filtering

关于支持保持对齐和有界过滤的局限性

Aryan Dutt, Rui Mao, Anupam Chattopadhyay

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究在大语言模型中,重塑基础模型输出分布的对齐方案与有界安全过滤器结合能否消除有害行为。通过形式化设置并分析,发现有界过滤可能无法消除所有有害输出,实证评估显示有害输出率始终高于零。

Comments Withdrawn to allow a complete rewrite and substantial revision of the theoretical analysis and experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27860 2026-08-04 cs.AI 版本更新 70%

C-MIG: Multi-view Information Gain-based Retrieval-Augmented Generation for Clinical Diagnosis Reasoning

C-MIG:基于多视角信息增益的检索增强生成用于临床诊断推理

Yuwei Miao, Gen Li, Yunsheng Zeng, Xiandong Li, Yujin Wang, Siyu Chen, Luning Wang, Yunhao Qiao, Junfeng Wang, Jianwei Lv, Bo Yuan

机构 * Baidu Inc(百度公司)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出C-MIG框架,通过多视角信息增益和多重子查询检索增强策略,解决检索增强生成中奖励信号丢失和异构推理监督问题,在临床诊断任务上取得最优性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27269 2026-08-04 cs.AI 版本更新 70%

Unifying biomedical knowledge in a modern multimodal graph

OptimusKG:统一生物医学知识的现代多模态图

Lucas Vittor, Ayush Noori, Iñaki Arango, Joaquín Polonuer, Sam Rodriques, Andrew White, David A. Clifton, Marinka Zitnik

机构 * Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系) Department of Engineering Science, University of Oxford(牛津大学工程科学系) Edison Scientific Inc.(Edison科学公司) Oxford Suzhou Centre for Advanced Research, University of Oxford(牛津大学苏格兰研究中心) Broad Institute of MIT and Harvard(MIT与哈佛大学Broad研究所) Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(哈佛大学自然与人工智能研究所) Harvard Data Science Initiative(哈佛大学数据科学计划)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 OptimusKG通过整合结构化和半结构化资源,构建了多模态生物医学标记属性图,统一了分子、解剖、临床和环境领域的知识,验证了其在生物医学领域的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.11804 2026-08-04 cs.CV cs.LG 版本更新 70%

OSMDA: OpenStreetMap-based Domain Adaptation for Remote Sensing VLMs

基于OpenStreetMap的领域自适应方法:为遥感VLMs的领域适应

Stefan Maria Ailuro, Mario Markov, Mohammad Mahdi, Delyan Boychev, Luc Van Gool, Danda Pani Paudel

机构 * INSAIT, Sofia University ``St. Kliment Ohridski''(INSAIT,索菲亚大学『圣克莱门特·奥赫里德斯』)

专题命中 领域大模型 :language model(abstract);foundation model(abstract);分类 cs.LG

AI总结 本文提出OSMDA方法,利用OpenStreetMap数据生成自标注数据,无需人工标注或强外部模型,通过细调基础VLM实现遥感领域适应,经10个基准测试显示其在文本生成任务中达到SOTA,且训练成本更低。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10300 2026-07-30 cs.HC cs.AI 版本更新 70%

AI LEGO: Scaffolding Cross-Functional Collaboration in Industrial Responsible AI Practices during Early Design Stages

AI LEGO:在早期设计阶段为工业负责任的AI实践搭建跨职能协作框架

Muzhe Wu, Yanzhi Zhao, Shuyi Han, Michael Xieyang Liu, Hong Shen

机构 * Carnegie Mellon University(卡内基梅隆大学) Northwestern University(西北大学)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 针对工业跨职能团队早期设计阶段负责任AI实践的协作难题,开发基于边界对象理论的AI LEGO原型,可提升跨角色危害识别的数量与概率,优化协作效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14546 2026-07-29 cs.CV cs.AI 版本更新 70%

Leveraging ChatGPT's Multimodal Vision Capabilities to Rank Satellite Images by Poverty Level: Advancing Tools for Social Science Research

利用ChatGPT的多模态视觉能力按贫困水平对卫星图像进行排名:推进社会科学研究工具

Hamid Sarmadi, Ola Hall, Thorsteinn Rögnvaldsson, Mattias Ohlsson

机构 * Center for Applied Intelligent Systems Research (CAISR), Halmstad University, Sweden(应用智能系统研究中心(CAISR),哈马斯特德大学,瑞典) Department of Human Geography, Lund University, Sweden(人类地理系,吕勒奥大学,瑞典) Department of Earth and Environmental Sciences, Lund University, Sweden(地球与环境科学系,吕勒奥大学,瑞典)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究利用具视觉能力的大语言模型分析卫星图像预测村级贫困,通过成对比较方法证明ChatGPT能按贫困水平排名卫星图像,准确性与专家相当,凸显LLMs在社会经济研究中的前景与局限,为贫困监测等提供基础并引发对公共数据集可靠性的思考。

Comments A code and data availability statement, along with the repository address, has been added to the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13465 2026-07-27 cs.LG stat.ML 版本更新 70%

AdamNX: An Adam improvement algorithm based on a novel exponential decay mechanism for the second-order moment estimate

AdamNX:一种基于二阶矩估计的新型指数衰减机制的Adam改进算法

Meng Zhu, Quan Xiao, Weidong Min

机构 * School of Information Management and Mathematics, Jiangxi University of Finance and Economics(江西财经大学信息管理与数学学院) School of Mathematics and Computer Science, Nanchang University(南昌大学数学与计算机科学学院) Institute of Metaverse, Nanchang University(南昌大学元宇宙研究院) Jiangxi Provincial Key Laboratory of Virtual Reality(江西省虚拟现实重点实验室)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究Adam中二阶矩估计的指数衰减机制,提出AdamNX及新衰减率,使更新在训练平稳阶段接近动量-SGD行为,通过图像分类等任务实验报告结果,代码已开源。

Comments 29 pages, 6 figures, 12 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.21693 2026-07-21 cs.AI 版本更新 70%

Deterministic Hallucination Detection in Medical VQA via Confidence-Evidence Bayesian Gain

医学VQA中通过置信度-证据贝叶斯增益实现确定性幻觉检测

Mohammad Asadi, Tahoura Nedaee, Jack W. O'Sullivan, Euan Ashley, Ehsan Adeli

机构 * Department of Electrical Engineering, Stanford University, CA, USA(电气工程系,斯坦福大学) Department of Biology, Stanford University, CA, USA(生物学系,斯坦福大学) Division of Cardiology, Department of Medicine, Stanford University, CA, USA(心脏病学部,医学系,斯坦福大学) Department of Biomedical Data Science, Stanford University, CA, USA(生物医学数据科学系,斯坦福大学) Department of Computer Science, Stanford University, CA, USA(计算机科学系,斯坦福大学) Department of Psychiatry and Behavioral Sciences, Stanford University, CA, USA(精神病学与行为科学系,斯坦福大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出CEBaG方法,利用模型自身log概率中的不一致置信度和弱视觉证据敏感性,实现无需随机采样和外部模型的确定性幻觉检测,在医疗MLLM和VQA基准测试中取得最佳AUC表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05743 2026-07-21 cs.AI cs.CY 版本更新 70%

Artificially intelligent agents in the social and behavioral sciences: A history and outlook

社会与行为科学中的人工智能代理:历史与展望

Petter Holme, Milena Tsvetkova

机构 * Department of Computer Science, Aalto University(阿尔托大学计算机科学系) Center for Computational Social Science, Kobe University(大阪大学计算社会科学中心) Department of Methodology, London School of Economics and Political Science(伦敦政治经济学院方法论系)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 回顾人工智能代理在社会与行为科学中的历史发展与当前趋势,涵盖从可编程计算机到社会模拟、大语言模型实验等,强调其在科学过程中的作用及带来的变化,如社会系统科学兴起等,凸显与相关技术紧密相连。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.11722 2026-07-16 cs.CL 版本更新 70%

STEP: Career-Path Recommendation via Temporal and Educational Trajectory Modeling

STEP:通过时间和教育轨迹建模进行职业路径推荐

Iman Johary, Guillaume Bied, Alexandru C. Mara, Tijl De Bie

机构 * AIDA-IDLab, Ghent University(AIDA-ID实验室,根特大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究利用大语言模型从简历获取职业轨迹数据,提出STEP职业路径推荐系统,通过集成时间衰减GRU、FiLM及注意力序列池化预测下一份工作,引入ROUTE改进职业表示,在多数据集评估中优于基准,还公开了数据集和代码。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.09322 2026-07-14 cs.AI 版本更新 70%

LongMedBench: Benchmarking Medical Agents for Long-Horizon Clinical Decision-Making

LongMedBench:用于长期临床决策的医疗智能体基准测试

Zihan Xu, Yanzhen Chen, Xiaocheng Zhang, Zhiting Fan, Weiqi Zhai, Hongxia Xu, Zuozhu Liu

机构 * Zhejiang University(浙江大学) Alibaba Group(阿里巴巴集团) Transvascular Implantation Devices Research Institute(血管内植入装置研究所)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 介绍基于EHR的LongMedBench基准,用于长期临床决策。构建含多患者多事件数据集,提出评估分类法。实验表明大语言模型在隐式时间推理有挑战,RAG和智能体记忆系统对信息检索有帮助,决策任务性能依赖模型即时上下文。

Comments Submitted manuscript prior to peer review in MICCAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27905 2026-07-14 cs.CL 版本更新 70%

AI Research Agents Narrow Scientific Exploration

AI研究代理缩小科学探索范围

Yixuan Tang, Yi Yang

机构 * The Hong Kong University of Science and Technology(香港理工大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究通过四个AI研究代理框架和六个大语言模型生成37,802个科学想法,发现AI生成的想法比人类论文更集中、更接近起始文献,且与低引用论文相似,表明当前AI代理更适合局部细化而非拓宽科学探索。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03564 2026-07-14 cs.LG 版本更新 70%

CoGenCast: A Coupled Autoregressive-Flow Generative Framework for Time Series Forecasting

CoGenCast:用于时间序列预测的耦合自回归流生成框架

Mingyue Cheng, Yaguo Liu, Daoyu Wang, Xiaoyu Tao, Qi Liu

机构 * State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 针对时间序列预测需兼顾语义理解与随机建模的问题,提出CoGenCast框架,将预训练大语言模型与流匹配机制结合,通过修改注意力拓扑重配置模型,集成流匹配机制捕捉动态,实现多模态预测和跨域统一训练,实验显示性能具竞争力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12873 2026-07-13 cs.HC cs.AI 版本更新 70%

Knowledge-Based Design Requirements for Generative Social Robots in Higher Education

面向高等教育的生成社交机器人知识基础设计需求

Stephan Vonschallen, Dominique Oberle, Theresa Schmiedel, Friederike Eyssel

机构 * Zurich University of Applied Sciences(应用科学大学苏黎世) University of Applied Sciences and Arts Northwestern Switzerland(西北瑞士应用科学与艺术大学) Bielefeld University(比勒菲尔德大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文基于十二次半结构化访谈,识别出生成社交机器人在高等教育中需满足的十二项设计要求,涵盖自我知识、用户知识和情境知识,为设计负责任且有效的教学GSR提供结构化基础。

Comments This paper was accepted for the International Conference on Social Robotics 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04920 2026-07-10 cs.AI 版本更新 70%

Conversational AI for Rapid Scientific Prototyping: A Case Study on ESA's ELOPE Competition

用于快速科学原型设计的对话式人工智能:以欧空局的ELOPE竞赛为例

Nils Einecke

机构 * Honda Research Institute Europe GmbH(本田欧洲研究机构)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 以欧空局ELOPE竞赛为例,探讨用ChatGPT进行快速科学原型设计。介绍其在竞赛中的表现,包括提供代码、推理等,虽有局限但凸显人机协作潜力,分析优缺点,提出将大语言模型集成到科学工作流程可增强快速原型设计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25605 2026-07-09 cs.IR cs.AI cs.DB 版本更新 70%

Health System Scale Semantic Search Across Unstructured Clinical Notes

跨无结构临床笔记的健康系统规模语义搜索

Faith Wavinya Mutinda, Spandana Makeneni, Anna Lin, Shivaji Dutta, Irit R. Rasooly, Patrick Dibussolo, Shivani Kamath Belman, Hessam Shahriari, Kevin Murphy, Alex B. Ruan, Barbara H. Chaiyachati, Sanjay Chainani, Robert W. Grundmeier, Scott M. Haag, Jeffrey M. Miller, Heather M. Griffis, Ian M. Campbell

机构 * Department of Biomedical and Health Informatics, Children’s Hospital of Philadelphia(儿童医院哲学学院生物医学与健康信息学系) Google Cloud(谷歌云) Department of Pediatrics, University of Pennsylvania(宾夕法尼亚大学儿科系) Division of Neonatology, Children’s Hospital of Philadelphia(儿童医院哲学学院新生儿科) Division of Human Genetics, Children’s Hospital of Philadelphia(儿童医院哲学学院人类遗传学部)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 本文提出了一种在大规模健康系统中实现语义搜索的解决方案,通过优化嵌入模型和分块策略,实现了亚秒级查询延迟和高准确率的临床问答任务,同时展示了在临床实用性评估中减少任务完成时间的效果。

Comments For associated code, see https://github.com/Ian-Campbell-Lab/clinical-semantic-search

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.17437 2026-07-08 cs.SE cs.LG 版本更新 70%

A semantic mutation metric for metamorphic relation adequacy in scientific computing programs

一种用于科学计算程序元突变关系充分性的语义突变度量

Meng Li, Xiaohua Yang, Jie Liu, Shiyu Yan

机构 * School of Computing, University of South China(南华大学计算机学院) Hunan Engineering Research Center of Software Evaluation and Testing for Intellectual Equipment(湖南软件测评与智能设备工程研究中心) CNNC Key Laboratory on High Trusted Computing(中核集团高可信计算重点实验室)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.LG

AI总结 本文提出了一种基于领域语义操作符的语义突变度量(SMS),旨在解决传统突变度量在科学计算中忽略领域语义的问题,通过引入五个领域语义操作符,提高了对元突变关系充分性的评估能力。

Comments 93 pages in elsarticle review mode (12pt double-spaced, ~28-35 pp typeset), 3 figures. Replication package: https://doi.org/10.5281/zenodo.20250664. Corresponding author: Meng Li (mlemon@usc.edu.cn)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29580 2026-07-07 cs.CL 版本更新 70%

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

MAM-AI:面向桑给巴尔护士和助产士的设备端医疗检索增强生成系统

Yi Ren

机构 * École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院洛桑分校)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 针对撒哈拉以南非洲地区护士助产士难以获取权威指南的问题,提出完全运行于安卓设备上的医疗问答系统MAM-AI,采用300M嵌入模型检索87份指南文档,并用4B int4生成器离线生成带引用的回答,评估发现小生成器在安全性和有用性间存在权衡,通过优化提示词降低回避率。

Comments 38 pages. Video demo: https://www.youtube.com/watch?v=M_Kruluel28 ; browser demo, code, models, and benchmarks linked in the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29467 2026-07-07 cs.CL cs.IR 版本更新 70%

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

mamabench 和 mamaretrieval:评估孕产妇、新生儿和生殖健康领域医学检索增强生成的基准

Yi Ren

机构 * École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院洛桑校区)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 针对助产士咨询的孕产妇、新生儿和生殖健康问题,构建了包含25,949个问题的QA基准mamabench和基于3,185个查询的段落级相关性基准mamaretrieval,采用分级相关标注和标签质量审计。

Comments 13 pages, 3 tables. Datasets and construction code linked in the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11730 2026-07-07 cs.CV cs.HC cs.LG 版本更新 70%

Multimodal Ambivalence/Hesitancy Recognition in Videos for Personalized Digital Health Interventions

视频中矛盾/犹豫识别用于个性化数字健康干预

Manuela González-González, Soufiane Belharbi, Muhammad Osama Zeeshan, Masoumeh Sharafi, Muhammad Haseeb Aslam, Lorenzo Sia, Nicolas Richet, Marco Pedersoli, Alessandro Lameiras Koerich, Simon L Bacon, Eric Granger

机构 * LIVIA, Dept. of Systems Engineering, ETS Montreal, Canada(ETS蒙特利尔大学系统工程系LIVIA实验室) LIVIA, Dept. of Software and IT Engineering, ETS Montreal, Canada(ETS蒙特利尔大学软件与信息工程系LIVIA实验室) Dept. of Health, Kinesiology, & Applied Physiology, Concordia University, Montreal, Canada(康科迪亚大学健康、运动科学与应用生理学系) Montreal Behavioural Medicine Centre, CIUSSS Nord-de-l’Ile-de-Montréal, Canada(蒙特利尔行为医学中心,蒙特利尔北岛卫生与社会服务局)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文研究了通过深度学习模型在视频中进行矛盾/犹豫识别,以提升数字健康干预的个性化和成本效益,实验基于新的BAH视频数据集,发现需改进多模态模型以准确识别矛盾/犹豫。

Comments 11 pages, 4 figures, ACII 2026. arXiv admin note: substantial text overlap with arXiv:2505.19328

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07285 2026-07-07 cs.CL cs.CY 版本更新 70%

Why teaching resists automation in an AI-inundated era: Human judgment, non-modular work, and the limits of delegation

为何在人工智能泛滥的时代教学抵抗自动化:人类判断、非模块化工作与委托的局限

Songhee Han

机构 * Florida State University(佛罗里达州立大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文探讨了在人工智能普及背景下,教学工作难以自动化的原因,指出教学本质上具有解释性、关联性和专业判断,无法被完全自动化或委托给技术。

Comments Revised version; accepted for publication in TechTrends

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08029 2026-07-07 cs.LG cs.CV 版本更新 70%

CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space

CLARITY:用于通过在潜在空间中建模情境感知疾病轨迹来指导治疗决策的医学世界模型

Tianxingjian Ding, Yuanhao Zou, Chen Chen, Mubarak Shah, Yu Tian

机构 * Institute of Artificial Intelligence, University of Central Florida(中佛罗里达大学人工智能研究所)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 CLARITY通过在潜在空间中建模时间间隔和患者特定数据,预测疾病演变轨迹,生成个性化治疗方案,并在医学领域实现最先进的治疗规划性能。

Comments Accepted to ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.13191 2026-07-03 physics.comp-ph cond-mat.mtrl-sci cs.AI 版本更新 70%

From Experiments to Expertise: Scientific Knowledge Consolidation for AI-Driven Computational Physics

从实验到专长:面向AI驱动计算物理的科学知识巩固

Haonan Huang

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出QMatSuite平台,通过记录、检索和反思机制巩固计算材料科学中的知识,减少推理开销并提高模拟准确性。

Comments v2: camera-ready version, accepted at the ICML 2026 Workshop on AI for Physics (AI4Physics@ICML 2026). 20 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30296 2026-07-02 cs.AI 版本更新 70%

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

ManimAgent: 用于视觉教育的自进化多模态智能体

Wenjia Jiang, Zongyuan Cai, Yuanhang Shao, Chenru Wang, Boyan Han, Zhixue Song, Keyu Chen, Shengwei An, Xu Yang, Zhou Yang

机构 * University of Alberta(阿尔伯塔大学) Southeast University(东南大学) Virginia Tech(弗吉尼亚理工学院) Xidian University(西安电子科技大学) Vivavia Inc(Vivavia公司)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出ManimAgent,通过双通道情节记忆库跨任务传递反思经验,无需权重更新或人工种子,在代码生成任务中提升通过率并减少反思轮次。

Comments Project page: https://manimagent.github.io/. Code: https://github.com/jwj1342/Paper2Manim

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17406 2026-07-02 cs.AI 版本更新 70%

EvoMaster: A Foundational Evolving Agent Framework for Agentic Science at Scale

EvoMaster:一种用于大规模代理科学的基础进化代理框架

Xinyu Zhu, Yuzhu Cai, Zexi Liu, Cheng Wang, Fengyang Li, Wenkai Jin, Wanxu Liu, Zehao Bing, Bingyang Zheng, Jingyi Chai, Shuo Tang, Rui Ye, Yuwen Du, Xianghe Pang, Yaxin Du, Tingjia Miao, Yuzhi Zhang, Ruoxue Liao, Zhaohan Ding, Linfeng Zhang, Yanfeng Wang, Weinan E, Siheng Chen

机构 * School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) SciLand DP Technology(DP技术)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 EvoMaster通过持续自我进化机制,使代理能迭代优化假设并积累知识,实现跨学科的高效科学发现,其易用性与性能在多个基准测试中均表现优异。

Comments 44 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏