Security and Privacy in Retrieval-Augmented Generation: Architectures, Threats, Defenses, and Future Directions for Building Trustworthy Systems
检索增强生成中的安全与隐私:构建可信系统的架构、威胁、防御与未来方向
Balamurugan Palanisamy, G S S Chalapathi, Vikas Hassija, Rajkumar Buyya
机构
*
Department of Electrical and Electronics Engineering(电子与电气工程系)
;
Birla Institute of Technology and Science, Pilani, Pilani Campus(比拉理工学院和科学学院,比拉校区)
;
Department of Computer Engineering, KIIT University(计算机工程系,KIIT大学)
;
Quantum Cloud Computing and Distributed Systems (qCLOUDS) Laboratory(量子云计算与分布式系统(qCLOUDS)实验室)
;
Department of Computing and Information Systems(计算与信息系统系)
;
The University of Melbourne(墨尔本大学)
How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations
OCR推理有多鲁棒?评估视觉语言模型在视觉扰动下的OCR推理鲁棒性
Yuxing Cheng, Yuan Wu, Yi Chang
机构
*
School of Artificial Intelligence, Jilin University(吉林大学人工智能学院)
;
Engineering Research Center of Knowledge-Driven Human-Machine Intelligence, MOE, China(教育部知识驱动人机智能工程研究中心)
;
International Center of Future Science, Jilin University(吉林大学未来科学国际合作中心)
机构
*
Beihang University(北京航空航天大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Beijing Jiaotong University(北京交通大学)
;
ATeam
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling
面向多模态大语言模型的长视频快速有效理解:自适应准高斯采样
Kun Zhang, Chenxin Fang, Tao Chen, Baiyang Song, Yunhang Shen, Yiyi Zhou, Rongrong Ji
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(厦门大学多媒体可信感知与高效计算教育部重点实验室)
专题命中
幻觉与鲁棒性
:multimodal large language model(title,abstract);分类 cs.CV
机构
*
Research Institute of Trustworthy Autonomous Systems and Department of Computer Science and Engineering, Southern University of Science and Technology(南方科技大学可信自主系统研究院与计算机科学与工程系)
;
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
专题命中
VLM训练与架构
:MLLM(title,title_cn);LLaVA(summary_cn,abstract);multimodal large language model(abstract);分类 cs.CV、cs.LG
Laurens Samson, Nimrod Barazani, Sennay Ghebreab, Yuki M. Asano
机构
*
Socially-Intelligent Artificial Systems Group, University of Amsterdam(智能社会人工智能系统组,阿姆斯特丹大学)
;
University of Amsterdam(阿姆斯特丹大学)
;
Fundamental AI Lab, University of Technology Nuremberg(基础人工智能实验室,纽伦堡技术大学)
专题命中
VLM训练与架构
:visual language model(title,abstract);VLM(abstract_cn);分类 cs.CV
MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models
MedLayBench-V:面向医学视觉语言模型中专家与普通人语义对齐的大规模基准
Han Jang, Junhyeok Lee, Heeseong Eum, Kyu Sung Choi
机构
*
Seoul National University(首尔国立大学)
;
Seoul National University College of Medicine(首尔国立大学医学院)
;
Department of Radiology, Seoul National University Hospital(首尔国立大学医院放射科)
;
Healthcare AI Research Institute, Seoul National University Hospital(首尔国立大学医院健康人工智能研究所)
;
The Advanced Imaging and Computational Neuroimaging (AICON) Laboratory(先进影像与计算神经影像实验室)
专题命中
VLM训练与架构
:vision language model(title);vision-language model(abstract)
ASSCG: Just-Right Gating over Chattering for Fast-Slow LLM Planning in Autonomous Driving
ASSCG:自动驾驶中快慢LLM规划的恰到好处门控
Sining Ang, Yuan Chen, Liu Haiyan, Xuanyao Mao, Jason Bao, Xuliang, Bingchuan Sun, Yan Wang
机构
*
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
Department of Automation, University of Science and Technology of China(中国科学技术大学自动化系)
;
Beijing University of Aeronautics and Astronautics(北京航空航天大学)
;
Lenovo Group Limited(联想集团)