Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients
揭示视觉语言大模型中的训练数据暴露:基于参数梯度的方法
Zhihao Zhu, Hongyi Tang, Yi Yang, Ahmed Abbasi
机构
*
Department of Information Systems, Business Statistics and Operations Management (ISOM), Hong Kong University of Science and Technology, Hong Kong, China(信息系统、商业统计与运营管理系(ISOM),香港科技大学,香港,中国)
;
Department of IT, Analytics, and Operations, University of Notre Dame, Notre Dame, Indiana, USA(信息技术、分析与运营系,诺丁汉大学,诺丁汉,印第安纳州,美国)
UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving
UniDrive: 面向自动驾驶可解释风险理解的统一视觉-语言与定位框架
Xiaowei Gao, Pengxiang Li, Yitai Cheng, Ruihan Xu, James Haworth, Stephen Law, Yun Ye
机构
*
organization= Department of Earth Science \& Engineering, Imperial College London , city= London , postcode= SW7 2AZ , country= United Kingdom
;
organization= SpaceTimeLab, Department of Civil, Environmental
;
Geomatic Engineering, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Department of Computing, The Hong Kong Polytechnic University , city= Hong Kong , country= China
;
organization= Trinity College, University of Oxford , city= Oxford , postcode= OX1 3BH , country= United Kingdom
;
organization= Department of Geography, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Centre for Global Infrastructure Resilience, The Bartlett School of Sustainable Construction, University College London , city= London , postcode= WC1E 7HB , country= United Kingdom
Page image classifier fine-tuned on century-spanning archives of scanned documents for further content-specific processing
基于百年跨度扫描文档档案微调的页面图像分类器,用于进一步的内容特定处理
Kateryna Lutsai, Dana Křivánková, Pavel Straňák, David Novák
机构
*
Institute of Formal and Applied Linguistics, Charles University MFF(查尔斯大学数学与物理学院形式与应用语言学研究所)
;
Institute of Archaeology, Czech Academy of Sciences(捷克科学院考古研究所)
机构
*
College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院)
;
Department of Bioengineering and Imperial-X, Imperial College London(帝国理工学院伦敦校区生物工程系)
;
Department of Pathology, Xiangtan Maternal and Child Health Hospital(湘潭 maternal and child health hospital pathology department)
;
Department of Pathology, The First People’s Hospital of Xiangtan City(湘潭市第一人民医院病理科)
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
State Key Lab of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China(中国科学院人工智能安全国家重点实验室,计算技术研究所,北京,中国)
;
Harbin Institute of Technology (Weihai)(哈尔滨工业大学(威海))
HANCLIP: A Family of Hyperbolic Angular Negation Vision Language Models
HANCLIP:双曲角否定视觉语言模型系列
Hoang-Bao Le, Aiden Durrant, Thai Son Mai, Binh T. Nguyen, Liting Zhou, Cathal Gurrin
机构
*
ADAPT Centre Dublin City University, Ireland(爱尔兰都柏林城市大学ADAPT中心)
;
University of East Anglia Norwich, UK(英国东英吉利大学)
;
Queen’s University Belfast Belfast, UK(英国贝尔法斯特女王大学)
;
University of Science Vietnam National University Ho Chi Minh City, Vietnam(越南胡志明市国家大学理科大学)
An Approach to Simultaneous Acquisition of Real-Time MRI Video, EEG, and Surface EMG for Articulatory, Brain, and Muscle Activity During Speech Production
一种用于言语产生过程中发音、大脑和肌肉活动的实时MRI视频、脑电图和表面肌电图同步采集方法
Jihwan Lee, Parsa Razmara, Kevin Huang, Sean Foley, Aditya Kommineni, Haley Hsu, Woojae Jeong, Prakash Kumar, Xuan Shi, Yoonjeong Lee, Tiantian Feng, Takfarinas Medani, Ye Tian, Sudarsana Reddy Kadiri, Krishna S. Nayak, Dani Byrd, Louis Goldstein, Richard M. Leahy, Shrikanth Narayanan
机构
*
Signal Analysis and Interpretation Laboratory, University of Southern California(南加州大学信号分析与解释实验室)
;
Ming Hsieh Dept. of Electrical and Computer Engineering, University of Southern California(南加州大学明希斯电气与计算机工程系)
;
Dept. of Linguistics, University of Southern California(南加州大学语言学系)
M^2C-EvDet: Multi-Domain Multi-Order Cross-Modal Knowledge Distillation for Event-based Object Detection
M^2C-EvDet:面向事件目标检测的多域多阶跨模态知识蒸馏
Wei Bao, Siqi Li, Shouan Pan, Yi Xie, Yue Gao
机构
*
BNRist, THUIBCS, BLBCI, School of Software, Tsinghua University(清华大学软件学院、北京信息科学与技术国家研究中心、清华-英特尔先进计算与智能技术联合研究中心、北京国家区块链与物联网技术研究中心)
;
Yangtze Delta Region Institute, Tsinghua University(清华大学长三角研究院)
;
School of Economics and Management, Beijing Forestry University(北京林业大学经济管理学院)
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴达摩院)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Monash University(莫纳什大学)
;
TRE, Alibaba Group(阿里巴巴TRE)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Xiamen University(厦门大学)
;
Kling Team, Kuaishou Technology(快手科技Kling团队)
;
National University of Singapore(新加坡国立大学)
;
Southern University of Science and Technology(南方科技大学)
机构
*
School of Software Engineering, Xi’an Jiaotong University(西安交通大学软件学院)
;
School of Information and Communication Engineering, Xi’an Jiaotong University(西安交通大学信息与通信工程学院)
;
SIGS, Tsinghua University(清华大学深圳国际研究生院)
机构
*
ByteDance Intelligent Creation(字节跳动智能创作)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of California, Merced(加州大学默塞德分校)
;
University of California, Los Angeles(加州大学洛杉矶分校)
Interpretable Material Spatial Intelligence for Discovery of Governing Microstructural Features
可解释的材料空间智能用于发现主导微观结构特征
Mathieu Calvat, Gregory Sparks, Dhruv Anjaria, Chris Bean, Haoren Wang, Paul Gradl, Timothy M. Smith, Allison M. Beese, Gabriel Demeneghi, Kenneth Vecchio, Morad Behandish, J. C. Stinville
MM-TRELLIS: Point-Cloud Guided Multi-Modal 3D Vehicle Generation in Autonomous Driving
MM-TRELLIS: 自动驾驶中基于点云引导的多模态3D车辆生成
Hongli Xiao, Youjian Zhang, Yucai Bai, Chaoyue Wang, Yaohui Jin, Xiaoguang Ren, Wenjing Yang, Long Lan
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院教育部人工智能重点实验室)
;
Academy of Military Science(军事科学院)
;
Bosch innovation software development (Wuxi) Co., Ltd.(博世创新软件开发(无锡)有限公司)
;
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机学院)
;
Shopee Pte. Ltd.(Shopee私人有限公司)