LLaVAShield: Safeguarding Multimodal Multi-Turn Dialogues in Vision-Language Models
LLaVAShield: 保障视觉语言模型中的多模态多轮对话安全
Guolei Huang, Qinzhi Peng, Gan Xu, Yao Huang, Yuxuan Lu, Yongjun Shen
机构
*
Southeast University(东南大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
Zhejiang University of Technology(浙江工业大学)
;
Tsinghua University(清华大学)
;
RealAI
机构
*
School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院)
;
School of Biomedical Engineering, Shanghai Jiao Tong University(上海交通大学生物医学工程学院)
;
Longwood Valley MedTech
Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
机器去学习并不如你所想:生成式AI政策与研究的启示
A. Feder Cooper, Christopher A. Choquette-Choo, Miranda Bogen, Kevin Klyman, Matthew Jagielski, Katja Filippova, Ken Liu, Alexandra Chouldechova, Jamie Hayes, Yangsibo Huang, Eleni Triantafillou, Peter Kairouz, Nicole Elyse Mitchell, Niloofar Mireshghallah, Abigail Z. Jacobs, James Grimmelmann, Vitaly Shmatikov, Christopher De Sa, Ilia Shumailov, Andreas Terzis, Solon Barocas, Jennifer Wortman Vaughan, danah boyd, Yejin Choi, Sanmi Koyejo, Fernando Delgado, Percy Liang, Daniel E. Ho, Pamela Samuelson, Miles Brundage, David Bau, Seth Neel, Hanna Wallach, Amy B. Cyphert, Mark A. Lemley, Nicolas Papernot, Katherine Lee
机构
*
The GenLaw Center(GenLaw中心)
;
Microsoft Research(微软研究院)
;
Stanford University(斯坦福大学)
;
Google DeepMind(谷歌DeepMind)
;
Center for Democracy & Technology(民主与科技中心)
;
Princeton(普林斯顿)
;
Google(谷歌)
;
University of Washington(华盛顿大学)
;
University of Michigan(密歇根大学)
;
Cornell Tech(康奈尔科技)
;
Cornell Law School(康奈尔法学院)
;
Cornell University(康奈尔大学)
;
Lighthouse
;
Stanford Law School(斯坦福法学院)
;
UC Berkeley(伯克利大学)
;
Independent(独立研究者)
;
Northeastern University(东北大学)
;
Harvard Business School(哈佛商学院)
;
W. Virginia University College of Law(维珍尼亚大学法学院)
机构
*
Department of Computer Science and Engineering(计算机科学与工程系)
;
Bangladesh University of Engineering and Technology (BUET)(孟加拉工程与技术大学)
;
Faculty of Information Technology(信息技术学院)
;
Monash University(莫纳什大学)
;
Qatar Computing Research Institute (QCRI)(卡塔尔计算研究所)
GTPO: Stabilizing Group Relative Policy Optimization via Gradient and Entropy Control
GTPO:通过梯度和熵控制稳定群体相对策略优化
Marco Simoni, Aleksandar Fontana, Giulio Rossolini, Andrea Saracino, Paolo Mori
机构
*
National Doctorate on Artificial Intelligence, Sapienza Università di Roma(人工智能国家博士学院,罗马萨皮恩扎大学)
;
Department of Excellence in Robotics and AI, TeCIP, Scuola Superiore Sant’Anna, Pisa(机器人与人工智能卓越部门,TeCIP,圣安娜高等学院,比萨)
;
Institute of Informatics and Telematics, National Research Council of Italy, Pisa(信息与电信研究所,意大利国家研究理事会,比萨)
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
SafeHumanoid: 通过VLM-RAG驱动的人形机器人上半身阻抗控制
Yara Mahmoud, Jeffrin Sam, Nguyen Khang, Marcelino Fernando, Issatay Tokmurziyev, Miguel Altamirano Cabrera, Muhammad Haris Khan, Artem Lykov, Dzmitry Tsetserukou
机构
*
Skolkovo Institute of Science and Technology(斯克洛尔沃科学与技术研究所)