arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Carnegie Mellon University(卡内基梅隆大学)

共收录 2342
2511.10874 2026-02-18 cs.RO cs.MA

Collaborative Multi-Robot Non-Prehensile Manipulation via Flow-Matching Co-Generation

通过流匹配共生成实现协作多机器人非抓取操作

Yorai Shaoul, Zhe Chen, Mohamed Naveed Gul Mohamed, Federico Pecora, Maxim Likhachev, Jiaoyang Li

机构 * Carnegie Mellon University(卡内基梅隆大学) Amazon Robotics(亚马逊机器人)

AI总结 本文提出一种统一框架,通过流匹配共生成与匿名多机器人运动规划,实现协作多机器人非抓取操作,提升复杂多物体环境下的操作效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04407 2026-02-18 cs.GT cs.LG

Scale-Invariant Regret Matching and Online Learning with Optimal Convergence: Bridging Theory and Practice in Zero-Sum Games

尺度不变的遗憾匹配与最优收敛的在线学习:在零和博弈中弥合理论与实践的鸿沟

Brian Hu Zhang, Ioannis Anagnostides, Tuomas Sandholm

机构 * Massachusetts Institute of Technology(麻省理工学院) Carnegie Mellon University(卡内基梅隆大学)

AI总结 本文提出了一种新的尺度不变且参数无关的PRM+变体IREG-PRM+,实现了最优收敛并优于现有方法,同时扩展了零和博弈到更广泛的变分不等式问题。

Comments Compared to the previous version, this version includes new results on harmonic games and extensive-form games. Abstract abridged due to arXiv length constraints

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02649 2026-02-18 cs.AI

From Prompts to Protection: Large Language Model-Enabled In-Context Learning for Smart Public Safety UAV

从提示到保护:基于大语言模型的上下文学习用于智能公共安全无人机

Yousef Emami, Hao Zhou, Miguel Gutierrez Gaitan, Kai Li, Luis Almeida, Zhu Han

机构 * Real-Time and Embedded Computing Systems Research Centre (CISTER)(实时嵌入式计算系统研究中心) Carnegie Mellon University(卡内基梅隆大学) Pontificia Universidad Católica de Chile(天主教大学) School of Computer Science, McGill University(麦吉尔大学计算机科学学院) Instituto de Telecomunicações, Faculdade de Engenharia, Universidade do Porto(电信研究所,工程学院,葡萄牙里斯本大学) University of Houston(休斯顿大学)

AI总结 本文提出利用大语言模型的上下文学习提升公共安全无人机的自主决策能力,通过自然语言提示和示例指导实现路径规划和速度控制,减少数据丢失并缓解安全漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15368 2026-02-18 cs.CV cs.AI cs.LG eess.IV

GMAIL: Generative Modality Alignment for generated Image Learning

GMAIL: 生成模态对齐用于生成图像学习

Shentong Mo, Sukmin Yun

机构 * Department of Machine Learning, CMU, USA(卡内基梅隆大学机器学习系) Department of Machine Learning, MBZUAI, UAE(马斯克大学人工智能研究所) Department of Artificial Intelligence, Hanyang University ERICA, South Korea(翰阳大学ERICA人工智能系)

AI总结 GMAIL通过多模态学习方法对齐生成图像与真实图像,提升视觉-语言任务中的生成图像学习效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15252 2026-02-18 cs.GT cs.AI cs.LG

Decision Making under Imperfect Recall: Algorithms and Benchmarks

不完美回忆下的决策:算法与基准测试

Emanuel Tewolde, Brian Hu Zhang, Ioannis Anagnostides, Tuomas Sandholm, Vincent Conitzer

机构 * Computer Science Dept., Carnegie Mellon University, Pittsburgh, USA(卡内基梅隆大学计算机科学系) Foundations of Cooperative AI Lab (FOCAL)(协作人工智能基础实验室(FOCAL)) Strategy Robot, Inc.(策略机器人公司) Strategic Machine, Inc.(战略机器公司) Optimized Markets, Inc.(优化市场公司)

AI总结 本文提出首个不完美回忆决策问题的基准测试集,并引入后悔匹配算法家族,在大规模约束优化中表现优异。

Comments 39 pages, 71 figures, 4 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03262 2026-02-18 cs.SE cs.CL

Is Vibe Coding Safe? Benchmarking Vulnerability of Agent-Generated Code in Real-World Tasks

Vibe 编程是否安全?在现实任务中对代理生成代码的漏洞基准测试

Songwen Zhao, Danqing Wang, Kexun Zhang, Jiaxuan Luo, Zhuo Li, Lei Li

机构 * Carnegie Mellon University Language Technologies Institute(卡内基梅隆大学语言技术研究所) Columbia University(哥伦比亚大学) Johns Hopkins University(约翰霍普金斯大学) HydroX AI

AI总结 本文通过基准测试发现,Vibe 编程在现实任务中生成的代码安全性不足,尽管功能正确但存在安全漏洞,警示其在安全敏感领域应用的风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01143 2026-02-18 cs.AI cs.LG

Generalized Parallel Scaling with Interdependent Generations

基于相互依赖世代的通用并行扩展

Harry Dong, David Brandfonbrener, Eryk Helenowski, Yun He, Mrinal Kumar, Han Fang, Yuejie Chi, Karthik Abinav Sankararaman

机构 * Meta Carnegie Mellon University(卡内基梅隆大学) Yale University(耶鲁大学)

AI总结 Bridge通过将批量LLM隐藏状态视为整体张量,实现相互依赖的并行生成,提升响应准确率和一致性,适用于任意生成宽度和聚合技术。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06134 2026-02-18 cs.AI

OpenAgentSafety: A Comprehensive Framework for Evaluating Real-World AI Agent Safety

OpenAgentSafety: 一个全面评估现实世界AI代理安全性的框架

Sanidhya Vijayvargiya, Aditya Bharat Soni, Xuhui Zhou, Zora Zhiruo Wang, Nouha Dziri, Graham Neubig, Maarten Sap

机构 * Language Technologies Institute, Carnegie Mellon University(语言技术研究所,卡内基梅隆大学) Allen Institute for Artificial Intelligence(人工智能研究院)

AI总结 OpenAgentSafety提出一个全面评估AI代理安全性的框架,通过真实工具和多任务测试揭示代理在现实世界中的安全漏洞,强调需要加强安全防护。

Comments 26 pages, 10 figures, Accepted at ICLR 2026 and IASEAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14989 2026-02-17 cs.CV cs.AI cs.LG

ThermEval: A Structured Benchmark for Evaluation of Vision-Language Models on Thermal Imagery

ThermEval: 一种用于评估视觉语言模型在热成像上的性能的结构化基准

Ayush Shrivastava, Kirtan Gangani, Laksh Jain, Mayank Goel, Nipun Batra

机构 * Indian Institute of Technology, Gandhinagar, India(印度理工学院加尔各答分校) Carnegie Mellon University, Pittsburgh, USA(卡内基梅隆大学)

AI总结 ThermEval通过结构化基准评估视觉语言模型在热成像上的性能,揭示其在温度推理和颜色映射转换上的不足,推动热视觉语言模型的发展。

Comments 8 Pages with 2 figures of main content. 2 pages of References. 10 pages of appendix with 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14874 2026-02-17 cs.RO

Affordance Transfer Across Object Instances via Semantically Anchored Functional Map

通过语义锚定的功能映射实现跨物体实例的 affordance 转移

Xiaoxiang Dong, Weiming Zhi

机构 * Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所) School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) Australian Centre for Robotics, The University of Sydney(悉尼大学机器人中心) College of Connected Computing, Vanderbilt University(范德比大学连接计算学院)

AI总结 本文提出语义锚定的功能映射方法,通过单个视觉示范实现跨不同几何物体的 affordance 转移,提升机器人感知与行动的实用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14432 2026-02-17 cs.LG cs.AI stat.ML

S2D: Selective Spectral Decay for Quantization-Friendly Conditioning of Neural Activations

S2D:选择性谱衰减用于神经激活的量化友好条件化

Arnav Chavan, Nahush Lele, Udbhav Bamba, Sankalp Dayal, Aditi Raghunathan, Deepak Gupta

机构 * Amazon(亚马逊公司) Carnegie Mellon University(卡内基梅隆大学)

AI总结 S2D通过选择性谱衰减减少神经激活异常值,提升模型在量化过程中的准确性和部署效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14224 2026-02-17 cs.SD cs.CL cs.MM

The Interspeech 2026 Audio Reasoning Challenge: Evaluating Reasoning Process Quality for Audio Reasoning Models and Agents

Interspeech 2026音频推理挑战:评估音频推理模型和代理的推理过程质量

Ziyang Ma, Ruiyang Xu, Yinghao Ma, Chao-Han Huck Yang, Bohan Li, Jaeyeon Kim, Jin Xu, Jinyu Li, Carlos Busso, Kai Yu, Eng Siong Chng, Xie Chen

机构 * Shanghai Jiao Tong University(上海交通大学) Nanyang Technological University(南洋理工大学) Queen Mary University of London(伦敦大学Queen Mary) NVIDIA(NVIDIA公司) Carnegie Mellon University(卡内基梅隆大学) Qwen Team, Alibaba Group(通义实验室,阿里巴巴集团) Microsoft Corporation(微软公司)

AI总结 Interspeech 2026音频推理挑战通过评估推理过程质量,探讨了音频推理模型和代理在事实性和逻辑性方面的表现及改进方向。

Comments The official website of the Audio Reasoning Challenge: https://audio-reasoning-challenge.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13928 2026-02-17 cs.SD cs.LG

voice2mode: Phonation Mode Classification in Singing using Self-Supervised Speech Models

voice2mode:基于自监督语音模型的歌唱发声模式分类

Aju Ani Justus, Ruchit Agrawal, Sudarsana Reddy Kadiri, Shrikanth Narayanan

机构 * University of Birmingham, School of Computer Science, Birmingham, UK(伯明翰大学计算机科学学院) Carnegie Mellon University, Information Systems, Doha, Qatar(卡内基梅隆大学信息系统) University of Southern California, Department of Electrical(南加州大学电气工程系)

AI总结 voice2mode利用自监督语音模型提取的嵌入对四种歌唱发声模式进行分类,实验表明基础模型特征在准确率上显著优于传统方法。

Comments Accepted to the Speech, Music and Mind (SMM26) workshop at the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2026). This is the preprint version of the paper to appear in the proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25260 2026-02-17 cs.AI cs.CL cs.LG

Internal Planning in Language Models: Characterizing Horizon and Branch Awareness

语言模型中的内部规划:刻画视野与分支意识

Muhammed Ustaomeroglu, Baris Askin, Gauri Joshi, Carlee Joe-Wong, Guannan Qu

机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) Carnegie Mellon University(卡内基梅隆大学)

AI总结 研究通过分析语言模型内部计算结构,揭示规划视野与分支意识的特性,为理解模型内部动态提供通用工具。

Comments Accepted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18053 2026-02-17 cs.RO

V2V-GoT: Vehicle-to-Vehicle Cooperative Autonomous Driving with Multimodal Large Language Models and Graph-of-Thoughts

V2V-GoT: 基于多模态大语言模型和思维图的车对车协同自动驾驶

Hsu-kuang Chiu, Ryo Hachiuma, Chien-Yi Wang, Yu-Chiang Frank Wang, Min-Hung Chen, Stephen F. Smith

机构 * NVIDIA Carnegie Mellon University(卡内基梅隆大学)

AI总结 本文提出V2V-GoT框架,结合多模态大语言模型和图-思维方法,提升车对车协同自动驾驶的感知、预测和规划能力。

Comments Accepted by ICRA 2026 (IEEE International Conference on Robotics and Automation). Project: https://eddyhkchiu.github.io/v2vgot.github.io/ Code: https://github.com/eddyhkchiu/V2V-GoT Dataset: https://huggingface.co/datasets/eddyhkchiu/V2V-GoT-QA

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15990 2026-02-17 cs.RO cs.CV

GelSLAM: A Real-time, High-Fidelity, and Robust 3D Tactile SLAM System

GelSLAM:一种实时、高保真度和鲁棒的3D触觉SLAM系统

Hung-Jui Huang, Mohammad Amin Mirzaee, Michael Kaess, Wenzhen Yuan

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

AI总结 GelSLAM通过触觉传感实现实时高保真3D SLAM,以高精度重建物体形状并提升手部操作任务的鲁棒性。

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02873 2026-02-17 cs.AI

It's the Thought that Counts: Evaluating the Attempts of Frontier LLMs to Persuade on Harmful Topics

想法才是关键:评估前沿大语言模型在有害话题上的说服尝试

Matthew Kowal, Jasper Timm, Jean-Francois Godbout, Thomas Costello, Antonio A. Arechar, Gordon Pennycook, David Rand, Adam Gleave, Kellin Pelrine

机构 * Université de Montréal, MILA(蒙特利尔大学,MILA) Carnegie Mellon University(卡内基梅隆大学) MIT, Center for Research and Teaching in Economics(麻省理工学院,经济研究与教学中心) Cornell University, University of Regina(康奈尔大学, Regina大学) Cornell University, MIT(康奈尔大学,麻省理工学院)

AI总结 本文提出APE基准测试,评估前沿大语言模型在有害话题上的说服意愿,揭示模型在有害情境下尝试说服的倾向及风险。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07861 2026-02-17 cs.CL cs.AI cs.LG

Scalable LLM Reasoning Acceleration with Low-rank Distillation

可扩展的大语言模型推理加速与低秩蒸馏

Harry Dong, Bilge Acun, Beidi Chen, Yuejie Chi

机构 * CMU Department of Electrical and Computer Engineering(卡内基梅隆大学电气与计算机工程系) Carnegie Mellon University(卡内基梅隆大学) Meta FAIR at Meta(Meta FAIR) CMU(卡内基梅隆大学)

AI总结 Caprese通过低秩蒸馏方法恢复高效推理方法中丢失的数学推理能力,同时减少参数数量和延迟,提升响应效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09980 2026-02-17 cs.CV cs.RO

V2V-LLM: Vehicle-to-Vehicle Cooperative Autonomous Driving with Multimodal Large Language Models

V2V-LLM:基于多模态大语言模型的车与车协同自动驾驶

Hsu-kuang Chiu, Ryo Hachiuma, Chien-Yi Wang, Stephen F. Smith, Yu-Chiang Frank Wang, Min-Hung Chen

机构 * NVIDIA Carnegie Mellon University(卡内基梅隆大学)

AI总结 本文提出基于多模态大语言模型的V2V-LLM,通过车与车协同感知提升自动驾驶安全性和性能。

Comments Accepted by ICRA 2026 (IEEE International Conference on Robotics and Automation). Project: https://eddyhkchiu.github.io/v2vllm.github.io/ Code: https://github.com/eddyhkchiu/V2V-LLM Dataset: https://huggingface.co/datasets/eddyhkchiu/V2V-GoT-QA

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13576 2026-02-17 cs.CR cs.AI cs.CL

Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges

评分标准作为攻击面:LLM裁判中的隐蔽偏好漂移

Ruomeng Ding, Yifei Pang, He Sun, Yizhong Wang, Zhiwei Steven Wu, Zhun Deng

机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) Carnegie Mellon University(卡内基梅隆大学) Yale University(耶鲁大学) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

AI总结 研究揭示了基于评分标准的LLM裁判中隐蔽的偏好漂移问题,通过评分标准攻击可系统性降低目标领域准确性,影响模型对齐流程。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13418 2026-02-17 cs.LG

Text Has Curvature

文本具有曲率

Karish Grover, Hanqing Zeng, Yinglong Xia, Christos Faloutsos, Geoffrey J. Gordon

机构 * Carnegie Mellon University(卡内基梅隆大学) Meta

AI总结 本文提出Texture,一种文本原生的词级曲率信号,通过定义和应用曲率检测与利用,建立文本曲率范式,提升长上下文推断和生成性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13398 2026-02-17 cs.LG q-bio.QM

Accelerated Discovery of Cryoprotectant Cocktails via Multi-Objective Bayesian Optimization

通过多目标贝叶斯优化加速冻保护剂混合物的发现

Daniel Emerson, Nora Gaby-Biegel, Purva Joshi, Yoed Rabin, Rebecca D. Sandlin, Levent Burak Kara

机构 * Mechanical Engineering Department, Carnegie Mellon University(卡内基梅隆大学机械工程系) Center for Engineering in Medicine & Surgery, Department of Surgery, Massachusetts General Hospital, Harvard Medical School, and Shriners Children’s(医学与手术工程中心,外科部,麻省总医院,哈佛医学院,以及谢尔曼儿童医院)

AI总结 通过多目标贝叶斯优化结合高通量筛选,高效发现同时具有高CPA浓度和高细胞存活率的冻保护剂混合物。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13346 2026-02-17 q-bio.GN cs.AI cs.CV

CellMaster: Collaborative Cell Type Annotation in Single-Cell Analysis

CellMaster: 单细胞分析中的协作细胞类型注释

Zhen Wang, Yiming Gao, Jieyuan Liu, Enze Ma, Jefferson Chen, Mark Antkowiak, Mengzhou Hu, JungHo Kong, Dexter Pratt, Zhiting Hu, Wei Wang, Trey Ideker, Eric P. Xing

机构 * Halicioglu Data Science Institute, University of California, San Diego, CA, USA(哈利奇奥格卢数据科学研究所,加州大学圣地亚哥分校) Department of Electrical & Computer Engineering, Texas A&M University, College Station, TX, USA(电气与计算机工程系,德克萨斯A&M大学) Department of Medicine, University of California, San Diego, CA, USA(医学系,加州大学圣地亚哥分校) Department of Chemistry and Biochemistry, University of California San Diego, La Jolla, CA, USA(化学与生物化学系,加州大学圣地亚哥分校) Moores Cancer Center, University of California, San Diego, La Jolla, CA, USA(摩尔癌症中心,加州大学圣地亚哥分校) Mohamed bin Zayed University of AI, Abu Dhabi, UAE(穆罕默德·本·扎耶德人工智能大学) School of Computer Science, Carnegie Mellon University, Pittsburgh, PA, USA(计算机科学学院,卡内基梅隆大学)

AI总结 CellMaster通过利用LLM编码知识实现零样本细胞类型注释,提升单细胞分析的准确性和可解释性。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13324 2026-02-17 cs.CV cs.AI cs.RO

Synthesizing the Kill Chain: A Zero-Shot Framework for Target Verification and Tactical Reasoning on the Edge

构建杀伤链:一种零样本框架,用于边缘节点上的目标验证和战术推理

Jesse Barkley, Abraham George, Amir Barati Farimani

机构 * With the Department of Mechanical Engineering, Carnegie Mellon University(卡内基梅隆大学机械工程系)

AI总结 本文提出了一种零样本框架,通过边缘设备实现目标验证和战术推理,验证了分层架构在动态军事环境中的有效性。

Comments 8 Pages, 3 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13212 2026-02-17 cs.RO cs.MA cs.SY eess.SY

UAVGENT: A Language-Guided Distributed Control Framework

UAVGENT: 一种语言引导的分布式控制框架

Ziyi Zhang, Xiyu Deng, Guannan Qu, Yorie Nakahira

机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) Carnegie Mellon University(卡内基梅隆大学)

AI总结 UAVGENT通过结合语言引导的任务推理与分布式反馈控制,实现了多无人机系统在复杂任务中的鲁棒性和稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13177 2026-02-16 math.OC cs.DS cs.LG

Improved Regret Guarantees for Online Mirror Descent using a Portfolio of Mirror Maps

使用镜像映射组合改进在线镜像下降的后悔保证

Swati Gupta, Jai Moondra, Mohit Singh

机构 * Massachusetts Institute of Technology(麻省理工学院) Carnegie Mellon University(卡内基梅隆大学) Georgia Institute of Technology(佐治亚理工学院)

AI总结 本文提出通过组合镜像映射改进在线镜像下降的后悔保证,展示基于块范数的镜像映射在稀疏损失函数中取得多项式级别的后悔改进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12763 2026-02-16 cs.HC cs.AI

"Not Human, Funnier": How Machine Identity Shapes Humor Perception in Online AI Stand-up Comedy

不是人类,更有趣:机器身份如何塑造在线AI单口喜剧的幽默感知

Xuehan Huang, Canwen Wang, Yifei Hao, Daijin Yang, Ray LC

机构 * The University of Hong Kong Hong Kong, SAR China Carnegie Mellon University\ -Computer Interaction Institute Pittsburgh United States East China Normal University Shanghai China Northeastern University\ of Art, Media City University of Hong Kong\ for Narrative Spaces Hong Kong, SAR China The University of Hong Kong Carnegie Mellon University\ -Computer Interaction Institute East China Normal University City University of Hong Kong\ for Narrative Spaces

AI总结 本研究探讨了AI身份如何影响幽默感知,通过设计基于机器身份的代理,发现其在单口喜剧表演中比基线GPT代理更有趣,提出人机集成系统应明确利用AI的独特身份。

Comments 27 pages, 5 figures. Conditionally Accepted to CHI '26

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07675 2026-02-16 cs.LG

Semantic Caching for Low-Cost LLM Serving: From Offline Learning to Online Adaptation

语义缓存用于低成本LLM服务:从离线学习到在线适应

Xutong Liu, Baran Atalar, Xiangxiang Dai, Jinhang Zuo, Siwei Wang, John C. S. Lui, Wei Chen, Carlee Joe-Wong

机构 * University of Washington(华盛顿大学) Carnegie Mellon University(卡内基梅隆大学) The Chinese University of Hong Kong(香港中文大学) City University of Hong Kong(香港城市大学) Microsoft Research(微软研究院)

AI总结 本文提出了一种基于学习的语义缓存框架,解决LLM服务中的缓存淘汰问题,通过离线优化和在线学习方法,在未知查询和成本分布下实现高效缓存管理。

Comments Accepted to INFOCOM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12714 2026-02-16 cs.LG

ADEPT: RL-Aligned Agentic Decoding of Emotion via Evidence Probing Tools -- From Consensus Learning to Ambiguity-Driven Emotion Reasoning

ADEPT: 通过证据探测工具实现情感的代理解码 -- 从共识学习到由模糊性驱动的情感推理

Esther Sun, Bo-Hao Su, Abinay Reddy Naini, Shinji Watanabe, Carlos Busso

机构 * cmu(卡内基梅隆大学)

AI总结 ADEPT通过证据探测工具实现情感的代理解码,从共识学习转向由模糊性驱动的情感推理,提升情感识别的准确性和可解释性。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12405 2026-02-16 cs.RO cs.LG

Self-Refining Vision Language Model for Robotic Failure Detection and Reasoning

自适应多任务视觉语言模型用于机器人故障检测与推理

Carl Qi, Xiaojie Wang, Silong Yong, Stephen Sheng, Huitan Mao, Sriram Srinivasan, Manikantan Nambi, Amy Zhang, Yesh Dattatreya

机构 * UT Austin(德克萨斯大学) Amazon Robotics(亚马逊机器人技术) Carnegie Mellon University(卡内基梅隆大学)

AI总结 ARMOR是一种自适应多任务视觉语言模型,通过多任务自 refinement 过程提升机器人故障检测与推理性能,实现故障检测率提升30%和推理表现提升100%。

详情

展开后加载摘要…

URL PDF HTML 收藏