arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 686 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 隐私与版权 686 篇

2504.09862 2025-11-18 cs.LG 57%

RadarLLM: Empowering Large Language Models to Understand Human Motion from Millimeter-Wave Point Cloud Sequence

Zengyuan Lai, Jiarui Yang, Songpengcheng Xia, Lizhou Lin, Lan Sun, Renwen Wang, Jianran Liu, Qi Wu, Ling Pei

专题命中 隐私与版权 :alignment(abstract);分类 cs.LG

Comments Accepted by AAAI 2026 (extended version with supplementary materials)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02573 2025-11-14 cs.CL 57%

Guess or Recall? Training CNNs to Classify and Localize Memorization in LLMs

Jérémie Dentan, Davide Buscaldi, Sonia Vanier

专题命中 隐私与版权 :alignment(abstract);分类 cs.CL

Comments This paper has been accepted for publication at AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14567 2025-11-12 cs.OS cs.AI 57%

Integrating Artificial Intelligence into Operating Systems: A Survey on Techniques, Applications, and Future Directions

Yifan Zhang, Xinkui Zhao, Ziying Li, Guanjie Cheng, Jianwei Yin, Lufei Zhang, Zuoning Chen

机构 * School of Software Technology, Zhejiang University(浙江大学软件学院) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) State Key Laboratory of Mathematical Engineering and Advanced Computing(数学工程与先进计算国家重点实验室) Chinese Academy of Engineering(中国工程院) College of Cyber Security, Jinan University(暨南大学网络安全学院)

专题命中 隐私与版权 :safety(abstract);分类 cs.AI

Comments 68 pages,9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07577 2025-11-12 cs.CR cs.CL cs.IR 57%

A Decentralized Retrieval Augmented Generation System with Source Reliabilities Secured on Blockchain

Yining Lu, Wenyi Tang, Max Johnson, Taeho Jung, Meng Jiang

专题命中 隐私与版权 :trustworthy(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01284 2025-11-04 cs.CV cs.AI 57%

Adaptation of Foundation Models for Medical Image Analysis: Strategies, Challenges, and Future Directions

Karma Phuntsho, Abdullah, Kyungmi Lee, Ickjai Lee, Euijoon Ahn

机构 * James Cook University(詹姆斯库克大学)

专题命中 隐私与版权 :trustworthy(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23019 2025-10-28 cs.LG cs.DC 57%

Sentinel: Dynamic Knowledge Distillation for Personalized Federated Intrusion Detection in Heterogeneous IoT Networks

Gurpreet Singh, Keshav Sood, P. Rajalakshmi, Yong Xiang

专题命中 隐私与版权 :alignment(abstract);分类 cs.LG

Comments This is a preprint version of a paper currently under review for possible publication in IEEE TDSC

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19036 2025-10-23 cs.CL 57%

From Memorization to Generalization: Fine-Tuning Large Language Models for Biomedical Term-to-Identifier Normalization

Suswitha Pericharla, Daniel B. Hier, Tayo Obafemi-Ajayi

专题命中 隐私与版权 :alignment(abstract);分类 cs.CL

Comments Submitted for publication to BMC BioData Mining

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18674 2025-10-22 cs.CR cs.AI 57%

Exploring Membership Inference Vulnerabilities in Clinical Large Language Models

Alexander Nemecek, Zebin Yun, Zahra Rahmani, Yaniv Harel, Vipin Chaudhary, Mahmood Sharif, Erman Ayday

机构 * Case Western Reserve University(凯斯西储大学) Tel Aviv University(特拉维夫大学)

专题命中 隐私与版权 :alignment(abstract);分类 cs.AI

Comments Accepted at the 1st IEEE Workshop on Healthcare and Medical Device Security, Privacy, Resilience, and Trust (IEEE HMD-SPiRiT)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18493 2025-10-22 cs.CR cs.AI cs.HC 57%

One Size Fits All? A Modular Adaptive Sanitization Kit (MASK) for Customizable Privacy-Preserving Phone Scam Detection

Kangzhong Wang, Zitong Shen, Youqian Zhang, Michael MK Cheung, Xiapu Luo, Grace Ngai, Eugene Yujun Fu

机构 * The Hong Kong Polytechnic University(香港理工大学) The Education University of Hong Kong(香港教育大学)

专题命中 隐私与版权 :safety(abstract);分类 cs.AI

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13606 2025-10-16 cs.LG 57%

Towards Robust Knowledge Removal in Federated Learning with High Data Heterogeneity

Riccardo Santi, Riccardo Salami, Simone Calderara

机构 * AImageLab, University of Modena and Reggio Emilia(AImageLab,摩德纳和雷吉奥艾米利亚大学)

专题命中 隐私与版权 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10805 2025-10-14 cs.HC cs.AI 57%

Therapeutic AI and the Hidden Risks of Over-Disclosure: An Embedded AI-Literacy Framework for Mental Health Privacy

Soraya S. Anvari, Rina R. Wehbe

机构 * Dalhousie University(达尔豪西大学)

专题命中 隐私与版权 :safety(abstract);分类 cs.AI

Comments Accepted to SMASH 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09434 2025-10-13 cs.CL 57%

Domain-Adapted Pre-trained Language Models for Implicit Information Extraction in Crash Narratives

Xixi Wang, Jordanka Kovaceva, Miguel Costa, Shuai Wang, Francisco Camara Pereira, Robert Thomson

机构 * Department of Technology, Management and Economics, Technical University of Denmark(技术、管理与经济系,丹麦技术大学) Department of Mechanics and Maritime Sciences, Chalmers University of Technology(机械与航海科学系,查尔姆斯理工大学) Department of Computer Science and Engineering, Chalmers University of Technology(计算机科学与工程系,查尔姆斯理工大学)

专题命中 隐私与版权 :safety(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06223 2025-10-10 cs.HC cs.AI 57%

A Multimodal GUI Architecture for Interfacing with LLM-Based Conversational Assistants

Hans G. W. van Dam

专题命中 隐私与版权 :alignment(abstract);分类 cs.AI

Comments 24 pages, 19 figures, code available at https://github.com/hansvdam/langbar

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07762 2025-10-10 cs.AI 57%

From Noisy to Native: LLM-driven Graph Restoration for Test-Time Graph Domain Adaptation

Xiangwei Lv, JinLuan Yang, Wang Lin, Jingyuan Chen, Beishui Liao

机构 * Zhejiang University(浙江大学)

专题命中 隐私与版权 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.01344 2025-10-07 cs.HC cs.AI cs.CR 57%

Privacy Leakage Overshadowed by Views of AI: A Study on Human Oversight of Privacy in Language Model Agent

Zhiping Zhang, Bingcan Guo, Tianshi Li

机构 * Northeastern University(东北大学) University of Washington(华盛顿大学)

专题命中 隐私与版权 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02357 2025-10-06 cs.CR cs.AI 57%

Privacy in the Age of AI: A Taxonomy of Data Risks

Grace Billiris, Asif Gill, Madhushi Bandara

专题命中 隐私与版权 :trustworthy(abstract);分类 cs.AI

Comments 12 pages, 2 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01967 2025-10-03 cs.CR cs.AI cs.CV 57%

ZK-WAGON: Imperceptible Watermark for Image Generation Models using ZK-SNARKs

Aadarsh Anantha Ramakrishnan, Shubham Agarwal, Selvanayagam S, Kunwar Singh

专题命中 隐私与版权 :trustworthy(abstract);分类 cs.AI

Comments Accepted at AI-ML Systems 2025, Bangalore, India, https://www.aimlsystems.org/2025/

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14234 2025-10-02 cond-mat.mtrl-sci cs.LG 57%

OBELiX: A Curated Dataset of Crystal Structures and Experimentally Measured Ionic Conductivities for Lithium Solid-State Electrolytes

Félix Therrien, Jamal Abou Haibeh, Divya Sharma, Rhiannon Hendley, Leah Wairimu Mungai, Sun Sun, Alain Tchagang, Jiang Su, Samuel Huberman, Yoshua Bengio, Hongyu Guo, Alex Hernández-García, Homin Shin

机构 * Mila McGill University(麦吉尔大学) University of Ottawa(Ottawa大学) Technical University of Kenya(肯尼亚技术大学) National Research Council Canada(加拿大国家研究委员会) Université de Montréal(蒙特利尔大学)

专题命中 隐私与版权 :safety(abstract);分类 cs.LG

Comments 10 pages, 4 figures and 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04117 2025-09-30 cs.CL 57%

Unveiling Over-Memorization in Finetuning LLMs for Reasoning Tasks

Zhiwen Ruan, Yun Chen, Yutao Hou, Peng Li, Yang Liu, Guanhua Chen

机构 * Southern University of Science and Technology(南方科技大学) Tsinghua University(清华大学) Shanghai University of Finance and Economics(上海财经大学)

专题命中 隐私与版权 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21712 2025-09-29 cs.CR cs.AI 57%

Not My Agent, Not My Boundary? Elicitation of Personal Privacy Boundaries in AI-Delegated Information Sharing

Bingcan Guo, Eryue Xu, Zhiping Zhang, Tianshi Li

机构 * Department of Human Centered Design & Engineering, University of Washington(华盛顿大学人中心设计与工程系) School of Information Sciences, University of Illinois(伊利诺伊大学香槟分校信息科学学院) Khoury College of Computer Sciences, Northeastern University(东北大学计算机科学学院)

专题命中 隐私与版权 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18400 2025-09-24 cs.AI 57%

ATLAS: Benchmarking and Adapting LLMs for Global Trade via Harmonized Tariff Code Classification

Pritish Yuvraj, Siva Devarakonda

专题命中 隐私与版权 :alignment(abstract);分类 cs.AI

Journal ref Paper in Review For ICLR 2026 (Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00073 2025-09-05 cs.LG 57%

Mitigating Clinician Information Overload: Generative AI for Integrated EHR and RPM Data Analysis

Ankit Shetgaonkar, Dipen Pradhan, Lakshit Arora, Sanjay Surendranath Girija, Shashank Kapoor, Aman Raj

专题命中 隐私与版权 :safety(abstract);分类 cs.LG

Comments Accepted at IEEE COMPSAC 2025

Journal ref 2025 IEEE 49th Annual Computers, Software, and Applications Conference (COMPSAC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02411 2025-09-03 cs.CR cs.AI 57%

A Survey: Towards Privacy and Security in Mobile Large Language Models

Honghui Xu, Kaiyang Li, Wei Chen, Danyang Zheng, Zhiyuan Li, Zhipeng Cai

机构 * Department of Information Technology, Kennesaw State University(肯纳邦大学信息科技系) Department of Computer Science, Georgia State University(佐治亚州立大学计算机科学系) Nexa AI School of Computing and Artificial Intelligence, Southwest Jiaotong University(西南交通大学计算机与人工智能学院)

专题命中 隐私与版权 :trustworthy(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19338 2025-09-03 cs.LG cs.CR 57%

Membership Inference Attacks on Large-Scale Models: A Survey

Hengyu Wu, Yang Cao

机构 * Institute of Science Tokyo(东京科学研究所)

专题命中 隐私与版权 :alignment(abstract);分类 cs.LG

Comments Preprint. Submitted for peer review. The final version may differ

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21815 2025-09-01 cs.LG 57%

Achieving Hilbert-Schmidt Independence Under Rényi Differential Privacy for Fair and Private Data Generation

Tobias Hyrup, Emmanouil Panagiotou, Arjun Roy, Arthur Zimek, Eirini Ntoutsi, Peter Schneider-Kamp

机构 * Department of Mathematics and Computer Science University of Southern Denmark(丹麦南部大学数学与计算机科学系) Department of Mathematics and Computer Science Freie Universität Berlin(柏林自由大学数学与计算机科学系) Faculty for Informatik Universität der Bundeswehr München(联邦国防军大学慕尼黑信息学院)

专题命中 隐私与版权 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19495 2025-08-28 cs.DC cs.LG eess.SP 57%

Towards 6G Intelligence: The Role of Generative AI in Future Wireless Networks

Muhammad Ahmed Mohsin, Junaid Ahmad, Muhammad Hamza Nawaz, Muhammad Ali Jamshed

机构 * School of Electrical Engineering, Stanford University(斯坦福大学电气工程学院) College of Science and Engineering, University of Glasgow(格拉斯哥大学科学与工程学院)

专题命中 隐私与版权 :trustworthy(abstract);分类 cs.LG

Comments Submitted as a chapter to the book Ambient Intelligence for 6G

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19115 2025-08-27 cs.CR cs.AI 57%

SecureV2X: An Efficient and Privacy-Preserving System for Vehicle-to-Everything (V2X) Applications

Joshua Lee, Ali Arastehfard, Weiran Liu, Xuegang Ban, Yuan Hong

机构 * University of California Santa Barbara(加州大学圣芭芭拉分校) University of Connecticut(康涅狄格大学) Alibaba Group(阿里巴巴集团) University of Washington(华盛顿大学)

专题命中 隐私与版权 :safety(abstract);分类 cs.AI

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16713 2025-08-26 cs.SE cs.AI hep-ex 57%

CelloAI: Leveraging Large Language Models for HPC Software Development in High Energy Physics

Mohammad Atif, Kriti Chopra, Ozgur Kilic, Tianle Wang, Zhihua Dong, Charles Leggett, Meifeng Lin, Paolo Calafiura, Salman Habib

机构 * Brookhaven National Laboratory(布鲁克海文国家实验室) Lawrence Berkeley National Laboratory(伯克利国家实验室) Argonne National Laboratory(阿贡国家实验室)

专题命中 隐私与版权 :safety(abstract);分类 cs.AI

Comments 12 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04519 2025-08-21 eess.AS cs.LG 57%

GenVC: Self-Supervised Zero-Shot Voice Conversion

Zexin Cai, Henry Li Xinyuan, Ashi Garg, Leibny Paola García-Perera, Kevin Duh, Sanjeev Khudanpur, Matthew Wiesner, Nicholas Andrews

机构 * Human Language Technology Center of Excellence Johns Hopkins University(人类语言技术中心杰出成就约翰霍普金斯大学)

专题命中 隐私与版权 :alignment(abstract);分类 cs.LG

Comments accepted by 2025 IEEE ASRU

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11579 2025-08-18 cs.CY 57%

Intergenerational Support for Deepfake Scams Targeting Older Adults

Karina LaRubbio, Alyssa Lanter, Seihyun Lee, Mahima Ramesh, Diana Freed

专题命中 隐私与版权 :safety(abstract);分类 cs.CY

Comments 3 pages, poster at the Twenty-First Symposium on Usable Privacy and Security (SOUPS) at https://www.usenix.org/conference/soups2025/presentation/larubbio-poster

详情

展开后加载摘要…

URL PDF HTML 收藏