Revealing Training Data Exposure in Vision Language Large Models via Parameter Gradients
揭示视觉语言大模型中的训练数据暴露:基于参数梯度的方法
Zhihao Zhu, Hongyi Tang, Yi Yang, Ahmed Abbasi
机构
*
Department of Information Systems, Business Statistics and Operations Management (ISOM), Hong Kong University of Science and Technology, Hong Kong, China(信息系统、商业统计与运营管理系(ISOM),香港科技大学,香港,中国)
;
Department of IT, Analytics, and Operations, University of Notre Dame, Notre Dame, Indiana, USA(信息技术、分析与运营系,诺丁汉大学,诺丁汉,印第安纳州,美国)
UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving
UniDrive: 面向自动驾驶可解释风险理解的统一视觉-语言与定位框架
Xiaowei Gao, Pengxiang Li, Yitai Cheng, Ruihan Xu, James Haworth, Stephen Law, Yun Ye
机构
*
organization= Department of Earth Science \& Engineering, Imperial College London , city= London , postcode= SW7 2AZ , country= United Kingdom
;
organization= SpaceTimeLab, Department of Civil, Environmental
;
Geomatic Engineering, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Department of Computing, The Hong Kong Polytechnic University , city= Hong Kong , country= China
;
organization= Trinity College, University of Oxford , city= Oxford , postcode= OX1 3BH , country= United Kingdom
;
organization= Department of Geography, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Centre for Global Infrastructure Resilience, The Bartlett School of Sustainable Construction, University College London , city= London , postcode= WC1E 7HB , country= United Kingdom
Page image classifier fine-tuned on century-spanning archives of scanned documents for further content-specific processing
基于百年跨度扫描文档档案微调的页面图像分类器,用于进一步的内容特定处理
Kateryna Lutsai, Dana Křivánková, Pavel Straňák, David Novák
机构
*
Institute of Formal and Applied Linguistics, Charles University MFF(查尔斯大学数学与物理学院形式与应用语言学研究所)
;
Institute of Archaeology, Czech Academy of Sciences(捷克科学院考古研究所)
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
State Key Lab of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China(中国科学院人工智能安全国家重点实验室,计算技术研究所,北京,中国)
;
Harbin Institute of Technology (Weihai)(哈尔滨工业大学(威海))
HANCLIP: A Family of Hyperbolic Angular Negation Vision Language Models
HANCLIP:双曲角否定视觉语言模型系列
Hoang-Bao Le, Aiden Durrant, Thai Son Mai, Binh T. Nguyen, Liting Zhou, Cathal Gurrin
机构
*
ADAPT Centre Dublin City University, Ireland(爱尔兰都柏林城市大学ADAPT中心)
;
University of East Anglia Norwich, UK(英国东英吉利大学)
;
Queen’s University Belfast Belfast, UK(英国贝尔法斯特女王大学)
;
University of Science Vietnam National University Ho Chi Minh City, Vietnam(越南胡志明市国家大学理科大学)
M^2C-EvDet: Multi-Domain Multi-Order Cross-Modal Knowledge Distillation for Event-based Object Detection
M^2C-EvDet:面向事件目标检测的多域多阶跨模态知识蒸馏
Wei Bao, Siqi Li, Shouan Pan, Yi Xie, Yue Gao
机构
*
BNRist, THUIBCS, BLBCI, School of Software, Tsinghua University(清华大学软件学院、北京信息科学与技术国家研究中心、清华-英特尔先进计算与智能技术联合研究中心、北京国家区块链与物联网技术研究中心)
;
Yangtze Delta Region Institute, Tsinghua University(清华大学长三角研究院)
;
School of Economics and Management, Beijing Forestry University(北京林业大学经济管理学院)
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴达摩院)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Monash University(莫纳什大学)
;
TRE, Alibaba Group(阿里巴巴TRE)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Xiamen University(厦门大学)
;
Kling Team, Kuaishou Technology(快手科技Kling团队)
;
National University of Singapore(新加坡国立大学)
;
Southern University of Science and Technology(南方科技大学)
机构
*
School of Software Engineering, Xi’an Jiaotong University(西安交通大学软件学院)
;
School of Information and Communication Engineering, Xi’an Jiaotong University(西安交通大学信息与通信工程学院)
;
SIGS, Tsinghua University(清华大学深圳国际研究生院)
Interpretable Material Spatial Intelligence for Discovery of Governing Microstructural Features
可解释的材料空间智能用于发现主导微观结构特征
Mathieu Calvat, Gregory Sparks, Dhruv Anjaria, Chris Bean, Haoren Wang, Paul Gradl, Timothy M. Smith, Allison M. Beese, Gabriel Demeneghi, Kenneth Vecchio, Morad Behandish, J. C. Stinville
MM-TRELLIS: Point-Cloud Guided Multi-Modal 3D Vehicle Generation in Autonomous Driving
MM-TRELLIS: 自动驾驶中基于点云引导的多模态3D车辆生成
Hongli Xiao, Youjian Zhang, Yucai Bai, Chaoyue Wang, Yaohui Jin, Xiaoguang Ren, Wenjing Yang, Long Lan
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院教育部人工智能重点实验室)
;
Academy of Military Science(军事科学院)
;
Bosch innovation software development (Wuxi) Co., Ltd.(博世创新软件开发(无锡)有限公司)
;
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机学院)
;
Shopee Pte. Ltd.(Shopee私人有限公司)
Prob-BBDM: a Probabilistic Brownian Bridge Diffusion Model for MRI sequence image-to-image translation
Prob-BBDM:用于MRI序列图像到图像翻译的概率布朗桥扩散模型
Martin Valls, Pascal Bourdon, Christine Fernandez-Maloigne, Guillaume Herpe, David Helbert
机构
*
University of Poitiers, CNRS, XLIM, France(波尔多大学,法国国家科学研究中心,XLIM,法国)
;
University of Poitiers, CNRS, Laboratory of Applied Mathematics, France(波尔多大学,法国国家科学研究中心,应用数学实验室,法国)
;
Poitiers University Hospital(波尔多大学医院)
;
I3M common laboratory CNRS-Siemens Healthinners, Poitiers University Hospital and University of Poitiers, France(I3M共同实验室,法国国家科学研究中心-西门子医疗,波尔多大学医院和波尔多大学,法国)
机构
*
The Hong Kong University of Technology and Science (Guangzhou)(香港科技与应用科技大学(广州))
;
East China Normal University(华东师范大学)
;
Shanghai Qi Zhi Institute(上海启智研究院)
;
Wuhan University(武汉大学)
;
Institute of Deep Perception Technology, JITRI(视觉感知技术研究院,JITRI)
;
CAIR, Hong Kong Institute of Science and Innovation (HKISI)(创新科技研究院,香港科学与创新研究院(HKISI))
;
MAIS, Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院)