arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Carnegie Mellon University(卡内基梅隆大学)

共收录 2349
2505.23799 2025-11-25 cs.CL cs.AI cs.HC cs.LG

Estimating LLM Consistency: A User Baseline vs Surrogate Metrics

估计LLM一致性:用户基准与替代指标

Xiaoyuan Wu, Weiran Lin, Omer Akgul, Lujo Bauer

机构 * Carnegie Mellon University(卡内基梅隆大学) RSAC Labs(RSAC实验室)

AI总结 本文提出了一种基于logits的集成方法,用于估计LLM一致性,并展示了其在匹配人类评分方面与现有最佳指标相当。

Comments Published as a main conference paper at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16466 2025-11-25 cs.CR cs.AI

Incalmo: An Autonomous LLM-assisted System for Red Teaming Multi-Host Networks

Incalmo:一种自主的LLM辅助系统用于多主机网络红队测试

Brian Singer, Keane Lucas, Lakshmi Adiga, Meghna Jain, Lujo Bauer, Vyas Sekar

机构 * Carnegie Mellon University(卡内基梅隆大学)

AI总结 Incalmo是一种基于LLM的自主红队测试系统,通过高层次任务规划和专用任务代理实现多主机网络攻击,显著提高了红队测试的效率和成功率。

Comments 18 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05986 2025-11-24 cs.CL

Fine-Grained Reward Optimization for Machine Translation using Error Severity Mappings

细粒度奖励优化用于机器翻译的误差严重性映射

Miguel Moura Ramos, Tomás Almeida, Daniel Vareta, Filipe Azevedo, Sweta Agrawal, Patrick Fernandes, André F. T. Martins

机构 * Instituto Superior Técnico, Universidade de Lisboa (ELLIS Unit Lisbon)(里斯本大学理工学院(ELLIS单位里斯本)) Instituto de Telecomunicações(电信研究所) Carnegie Mellon University(卡内基梅隆大学) TransPerfect Core contributor(TransPerfect核心贡献者)

AI总结 本文提出利用细粒度词级奖励和误差严重性映射优化,提升机器翻译质量与训练稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01478 2025-11-24 cs.CL cs.AI cs.LG

SePer: Measure Retrieval Utility Through The Lens Of Semantic Perplexity Reduction

SePer:通过语义困惑度降低的视角衡量检索效用

Lu Dai, Yijie Xu, Jinhui Ye, Hao Liu, Hui Xiong

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) The Hong Kong University of Science and Technology(香港科学与技术大学) Carnegie Mellon University(卡内基梅隆大学)

AI总结 SePer通过语义困惑度降低衡量检索效用,提供更精确的RAG评估方法

Comments ICLR 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14945 2025-11-21 cs.CV

Unsupervised Discovery of Long-Term Spatiotemporal Periodic Workflows in Human Activities

无监督发现人类活动中的长期时空周期性工作流

Fan Yang, Quanting Xie, Atsunori Moteki, Shoichi Masui, Shan Jiang, Kanji Uchino, Yonatan Bisk, Graham Neubig

机构 * Fujitsu Research of America, USA(美国富士通研究机构) Fujitsu Limited, Japan(日本富士通有限公司) Carnegie Mellon University, USA(美国卡内基梅隆大学)

AI总结 本文提出首个包含长期周期性工作流的基准,通过轻量级基线实现无监督检测和异常检测,显著优于现有方法并具备实际部署优势。

Comments accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10848 2025-11-21 cs.LG cs.AI

STAMP: Spatial-Temporal Adapter with Multi-Head Pooling

STAMP:带有多头池化的空间-时间适配器

Brad Shook, Abby Turner, Jieshi Chen, Michał Wiliński, Mononito Goswami, Jonathan Elmer, Artur Dubrawski

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Pittsburgh School of Medicine(匹兹堡大学医学学院)

AI总结 STAMP通过多头池化机制,利用通用TSFMs生成的单变量嵌入,隐式建模EEG数据的空间-时间特性,实现与现有EEGFMs相当的性能。

Comments Accepted as a Proceedings paper at Machine Learning for Health (ML4H) 2025, invited presentation at the Time Series for Health (TS4H) Workshop, NeurIPS 2025. v2: Updated author affiliation and corrected a duplicated word in the text. No other changes

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00911 2025-11-21 cs.LG stat.ML

Distance-Preserving Representations for Genomic Spatial Reconstruction

用于基因组空间重建的距离保持表示

Wenbin Zhou, Jin-Hong Du

机构 * Heinz College of Information Systems and Public Policy and Machine Learning Department, Carnegie Mellon University(信息系统与公共政策学院及机器学习系,卡内基梅隆大学) Musketeers Foundation Institute of Data Science and the Department of Statistics and Actuarial Science, University of Hong Kong(数据科学研究所及统计与精算科学系,香港大学)

AI总结 本文提出dp-VAE框架,通过距离保持正则化器重建基因组空间坐标,提升单细胞数据的空间分析能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07279 2025-11-21 cs.CL cs.AI

MAQuA: Adaptive Question-Asking for Multidimensional Mental Health Screening using Item Response Theory

MAQuA:基于项目反应理论的多维心理健康筛查自适应提问框架

Vasudha Varadarajan, Hui Xu, Rebecca Astrid Boehme, Mariam Marlan Mirstrom, Sverker Sikstrom, H. Andrew Schwartz

机构 * Language Technologies Institute, Carnegie Mellon University(卡内基梅隆大学语言技术研究所) Stony Brook University(石溪大学) Aarhus University(奥胡斯大学) Department of Psychology, Lund University(吕勒欧大学心理学系) College of Connected Computing, Vanderbilt University(范德比尔特大学连接计算学院)

AI总结 MAQuA通过结合IRT和因子分析,实现多维心理健康筛查的自适应提问,有效减少评估问题数量,提升诊断效率和患者体验。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09109 2025-11-21 cs.CV cs.CL

CAIRe: Cultural Attribution of Images by Retrieval-Augmented Evaluation

CAIRe:通过检索增强评估进行图像文化归因

Arnav Yayavaram, Siddharth Yayavaram, Simran Khanuja, Michael Saxon, Graham Neubig

机构 * BITS Pilani(比斯·皮兰大学) Carnegie Mellon University(卡内基梅隆大学) University of California, Santa Barbara(加州大学圣巴巴拉分校)

AI总结 CAIRe通过检索增强评估方法,评估图像在不同文化标签下的相关性,有效衡量文化偏见,提升跨文化公平性。

Comments Preprint, under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.01144 2025-11-21 eess.IV cs.AI cs.CV cs.LG

LEARNER: Contrastive Pretraining for Learning Fine-Grained Patient Progression from Coarse Inter-Patient Labels

LEARNER: 通过对比学习从粗粒度患者间标签中学习细粒度患者进展

Jana Armouti, Nikhil Madaan, Rohan Panda, Tom Fox, Laura Hutchins, Amita Krishnan, Ricardo Rodriguez, Bennett DeBoisblanc, Deva Ramanan, John Galeotti, Gautam Gare

机构 * Carnegie Mellon University(卡内基梅隆大学) LSUHSC Internal Medicine(LSUHSC内科) Cosmetic Surgery Facility LLC(美容外科诊所有限公司)

AI总结 LEARNER通过对比学习利用患者间粗粒度标签,学习细粒度患者内变化,提升个性化医学中的治疗响应预测性能。

Comments Under review at ISBI 2026 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14876 2025-11-20 cs.CR cs.CV cs.LG cs.RO

Attacking Autonomous Driving Agents with Adversarial Machine Learning: A Holistic Evaluation with the CARLA Leaderboard

Henry Wong, Clement Fung, Weiran Lin, Karen Li, Stanley Chen, Lujo Bauer

机构 * Carnegie Mellon University(卡内基梅隆大学)

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14868 2025-11-20 cs.CL cs.LG

Hierarchical Token Prepending: Enhancing Information Flow in Decoder-based LLM Embeddings

Xueying Ding, Xingyue Huang, Mingxuan Ju, Liam Collins, Yozen Liu, Leman Akoglu, Neil Shah, Tong Zhao

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Oxford(牛津大学) Snap Inc(Snap公司)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12001 2025-11-20 cs.CL cs.HC

Critical or Compliant? The Double-Edged Sword of Reasoning in Chain-of-Thought Explanations

Eunkyu Park, Wesley Hanwen Deng, Vasudha Varadarajan, Mingxi Yan, Gunhee Kim, Maarten Sap, Motahhare Eslami

机构 * Seoul National University(首尔国立大学) Language Technologies Institute, Carnegie Mellon University(语言技术研究所,卡内基梅隆大学) Human-Computer Interaction Institute, Carnegie Mellon University(人机交互研究所,卡内基梅隆大学)

Comments Under review; 16 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01293 2025-11-19 cs.CV cs.AI

GMAT: Grounded Multi-Agent Clinical Description Generation for Text Encoder in Vision-Language MIL for Whole Slide Image Classification

Ngoc Bui Lam Quang, Nam Le Nguyen Binh, Thanh-Huy Nguyen, Le Thien Phuc Nguyen, Quan Nguyen, Ulas Bagci

机构 * AI VIETNAM(AI越南) Carnegie Mellon University(卡内基梅隆大学) University of Wisconsin-Madison(威斯康星大学麦迪逊分校) PTIT Northwestern University(西北大学)

Comments Acccepted in MICCAI Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17805 2025-11-19 cs.CR cs.LG

AdRo-FL: Informed and Secure Client Selection for Federated Learning in the Presence of Adversarial Aggregator

Md. Kamrul Hossain, Walid Aljoby, Anis Elgabli, Ahmed M. Abdelmoniem, Khaled A. Harras

机构 * College of Computing and Mathematics, King Fahd University of Petroleum & Minerals(计算机与数学学院,国王法赫德石油与矿物大学) School of Electronic Engineering and Computer Science, Queen Mary University of London(电子工程与计算机科学学院,女王玛丽大学) Department of Computer Science, Carnegie Mellon University(计算机科学系,卡内基梅隆大学)

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10683 2025-11-19 cs.RO cs.AI cs.CV

MotIF: Motion Instruction Fine-tuning

Minyoung Hwang, Joey Hejna, Dorsa Sadigh, Yonatan Bisk

机构 * Massachusetts Institute of Technology(麻省理工学院) Stanford University(斯坦福大学) Carnegie Mellon University(卡内基梅隆大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13785 2025-11-19 eess.SY cs.AI cs.SY

Quantifying Distribution Shift in Traffic Signal Control with Histogram-Based GEH Distance

Federico Taschin, Ozan K. Tonguz

机构 * KTH Royal Institute of Technology(皇家理工学院) Carnegie Mellon University(卡内基梅隆大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13772 2025-11-19 cs.MM cs.AI cs.CV cs.CY

Can LLMs Create Legally Relevant Summaries and Analyses of Videos?

Lyra Hoeben-Kuil, Gijs van Dijck, Jaromir Savelka, Johanna Gunawan, Konrad Kollnig, Marta Kolacz, Mindy Duffourc, Shashank Chakravarthy, Hannes Westermann

机构 * Maastricht Law and Tech Lab, Faculty of Law, Maastricht University, The Netherlands(马斯特里赫特大学法学院与科技实验室,荷兰) Brightlands Institute for Smart Society, Maastricht University, The Netherlands(智慧社会Brightlands研究所,马斯特里赫特大学,荷兰) Computer Science Department, Carnegie Mellon University, Pittsburg, USA(计算机科学系,卡内基梅隆大学,美国匹兹堡)

Comments Accepted for publication at JURIX 2025 Torino, Italy. This is the preprint version. Code and data available at: https://github.com/maastrichtlawtech/jurix2025_LLM_video_analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05817 2025-11-19 cs.HC cs.MM cs.SD

TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech

Weiyan Shi, Sunaya Upadhyay, Geraldine Quek, Kenny Tsu Wei Choo

机构 * Singapore University of Technology and Design(新加坡科技设计大学) Carnegie Mellon University(卡内基梅隆大学)

Comments Accepted at AAAI 2026 Workshop on Creative AI for Live Interactive Performances (CLIP). To be published in Springer CCIS series

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12179 2025-11-19 cs.AI cs.MA

Co-Alignment: Rethinking Alignment as Bidirectional Human-AI Cognitive Adaptation

Yubo Li, Weiyi Song

机构 * Carnegie Mellon University(卡内基梅隆大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13378 2025-11-19 cs.CV

Governance-Ready Small Language Models for Medical Imaging: Prompting, Abstention, and PACS Integration

Yiting Wang, Ziwei Wang, Di Zhu, Jiachen Zhong, Weiyi Li

机构 * Department of Data Science, University of Southern California(数据科学系,南加州大学) Department of Electrical and Computer Engineering, Carnegie Mellon University(电气与计算机工程系,卡内基梅隆大学) Department of Computer Science and Engineering, Santa Clara University(计算机科学与工程系,圣克拉拉大学) Department of Applied Mathematics, University of Washington(应用数学系,华盛顿大学) School of Computer Science, Georgia Institute of Technology(计算机科学学院,佐治亚理工学院)

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16029 2025-11-19 cs.CL cs.AI cs.LG

EvoLM: In Search of Lost Language Model Training Dynamics

Zhenting Qi, Fan Nie, Alexandre Alahi, James Zou, Himabindu Lakkaraju, Yilun Du, Eric Xing, Sham Kakade, Hanlin Zhang

机构 * Harvard(哈佛大学) Stanford(斯坦福大学) EPFL(苏黎世联邦理工学院) CMU(卡内基梅隆大学)

Comments NeurIPS 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00827 2025-11-19 cs.AI

MIMIC-\RNum{4}-Ext-22MCTS: A 22 Millions-Event Temporal Clinical Time-Series Dataset with Relative Timestamp for Risk Prediction

Jing Wang, Xing Niu, Tong Zhang, Jie Shen, Juyong Kim, Jeremy C. Weiss

机构 * National Library of Medicine(国家医学图书馆) AWS AI Labs(AWS人工智能实验室) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Stevens Institute of Technology(史蒂文斯理工学院) Carnegie Mellon University(卡内基梅隆大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17708 2025-11-18 cs.LG

The Third Pillar of Causal Analysis? A Measurement Perspective on Causal Representations

Dingling Yao, Shimeng Huang, Riccardo Cadei, Kun Zhang, Francesco Locatello

机构 * Institute of Science and Technology Austria(奥地利科学技术研究院) Carnegie Mellon University(卡内基梅隆大学) Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(马尔代夫穆罕默德·本·扎耶德人工智能大学)

Comments Camera-ready version for NeurIPS2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12754 2025-11-18 cs.AI cs.LG cs.MA

Adaptively Coordinating with Novel Partners via Learned Latent Strategies

Benjamin Li, Shuyang Shi, Lucia Romero, Huao Li, Yaqi Xie, Woojun Kim, Stefanos Nikolaidis, Michael Lewis, Katia Sycara, Simon Stepputtis

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Pittsburgh(匹兹堡大学) University of Southern California(南加州大学) Virginia Tech(弗吉尼亚理工学院)

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12691 2025-11-18 cs.CV cs.AI

R$^{2}$Seg: Training-Free OOD Medical Tumor Segmentation via Anatomical Reasoning and Statistical Rejection

Shuaike Shen, Ke Liu, Jiaqing Xie, Shangde Gao, Chunhua Shen, Ge Liu, Mireia Crispin-Ortuzar, Shangqi Gao

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Cambridge(剑桥大学) Zhejiang University(浙江大学) ETH Zurich(苏黎世联邦理工学院) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12429 2025-11-18 cs.LG

Tailored Primitive Initialization is the Secret Key to Reinforcement Learning

Yihang Yao, Guangtao Zeng, Raina Wu, Yang Zhang, Ding Zhao, Zhang-Wei Hong, Chuang Gan

机构 * Carnegie Mellon University(卡内基梅隆大学) Massachusetts Institute of Technology(麻省理工学院) MIT-IBM Watson AI Lab(MIT-IBM沃森人工智能实验室)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06776 2025-11-18 cs.RO

FALCON: Learning Force-Adaptive Humanoid Loco-Manipulation

Yuanhang Zhang, Yifu Yuan, Prajwal Gurunath, Ishita Gupta, Shayegan Omidshafiei, Ali-akbar Agha-mohammadi, Marcell Vazquez-Chanlatte, Liam Pedersen, Tairan He, Guanya Shi

机构 * Carnegie Mellon University(卡内基梅隆大学) Field AI Nissan USA(日产美国)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10931 2025-11-18 eess.IV cs.CV

Towards Collective Intelligence: Uncertainty-aware SAM Adaptation for Ambiguous Medical Image Segmentation

Mingzhou Jiang, Jiaying Zhou, Junde Wu, Tianyang Wang, Yueming Jin, Min Xu

机构 * Department of Computer Science, The University of Alabama at Birmingham(阿拉巴马大学伯明翰分校计算机科学系) Doctoral Training Centre, University of Oxford(牛津大学博士培训中心) Department of Biomedical Engineering and Department of Electrical and Computer Engineering, National University of Singapore(新加坡国立大学生物医学工程系和电气与计算机工程系) Computational Biology Department, Carnegie Mellon University(卡内基梅隆大学计算生物学系) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12181 2025-11-18 cs.CV cs.LG

MixAR: Mixture Autoregressive Image Generation

Jinyuan Hu, Jiayou Zhang, Shaobo Cui, Kun Zhang, Guangyi Chen

机构 * Tsinghua University(清华大学) Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(Mohamed bin Zayed人工智能大学) École Polytechnique Fédérale de Lausanne (EPFL)(洛桑联邦理工学院) Carnegie Mellon University(卡内基梅隆大学)

详情

展开后加载摘要…

URL PDF HTML 收藏