InfSplign: Inference-Time Spatial Alignment of Text-to-Image Diffusion Models
InfSplign: 文本到图像扩散模型推理时的空间对齐
Sarah Rastegar, Violeta Chatalbasheva, Sieger Falkena, Anuj Singh, Yanbo Wang, Tejas Gokhale, Hamid Palangi, Hadi Jamali-Rad
机构
*
Delft University of Technology(代尔夫特理工大学)
;
University of Maryland Baltimore County(马里兰大学巴尔的摩县分校)
;
Shell Information Technology International(壳牌信息科技国际)
;
Google(谷歌)
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
文本到图像扩散模型中后门检测的动态注意力分析
Zhongqi Wang, Jie Zhang, Shiguang Shan, Xilin Chen
机构
*
Key Laboratory of AI Safety of CAS, Institute of Computing Technology (ICT), Chinese Academy of Sciences (CAS), Beijing 100190, China, and also with the University of Chinese Academy of Sciences (UCAS), Beijing 100049, China(中国科学院人工智能安全重点实验室,计算技术研究所(ICT),中国科学院(CAS),北京100190,中国,以及中国科学院大学(UCAS),北京100049,中国)
SP-Guard: Selective Prompt-adaptive Guidance for Safe Text-to-Image Generation
Sumin Yu, Taesup Moon
机构
*
Department of Electrical and Computer Engineering, Seoul National University, Seoul, South Korea(电气与计算机工程系,首尔国立大学,首尔,韩国)
;
IPAI / ASRI / INMC, Seoul National University, Seoul, South Korea(IPAI/ASRI/INMC,首尔国立大学,首尔,韩国)
Hawk: Leveraging Spatial Context for Faster Autoregressive Text-to-Image Generation
Zhi-Kai Chen, Jun-Peng Jiang, Han-Jia Ye, De-Chuan Zhan
机构
*
School of Artificial Intelligence, Nanjing University, China(南京大学人工智能学院)
;
National Key Laboratory for Novel Software Technology, Nanjing University, China(南京大学新型软件技术国家重点实验室)
机构
*
Institute of Artificial Intelligence, Beihang University(北航人工智能研究院)
;
College of AI, Tsinghua University(清华人工智能学院)
;
Security Group, Alibaba Group(阿里集团安全组)
机构
*
College of Computer Science and Technology, Zhejiang University, China(浙江大学计算机科学与技术学院)
;
Mohamed bin Zayed University of Artificial Intelligence, Abu Dhabi(阿布扎赫德莫罕默德·本·扎耶德人工智能大学)
;
Department of Information Engineering and Computer Science, University of Trento, Italy(意大利特伦托大学信息工程与计算机科学系)
Free Lunch Alignment of Text-to-Image Diffusion Models without Preference Image Pairs
Jia Jun Cheng Xian, Muchen Li, Haotian Yang, Xin Tao, Pengfei Wan, Leonid Sigal, Renjie Liao
机构
*
University of British Columbia(不列颠哥伦比亚大学)
;
Vector Institute for AI(人工智能向量研究所)
;
Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)
;
Kling Team, Kuaishou Technology(快手科技 Kling 团队)
;
NSERC CRC Chair(加拿大NSERC CRC主席)
Taming the Tri-Space Tension: ARC-Guided Hallucination Modeling and Control for Text-to-Image Generation
Jianjiang Yang, Ziyan Huang, Yanshu li, Da Peng, Huaiyuan Yao
机构
*
University of Bristol(布里斯托大学)
;
South China University of Technology(华南理工大学)
;
Brown University(布朗大学)
;
Xi’an Jiaotong University(西安交通大学)
;
Arizona State University(亚利桑那州立大学)
FairCoT: Enhancing Fairness in Text-to-Image Generation via Chain of Thought Reasoning with Multimodal Large Language Models
Zahraa Al Sahili, Ioannis Patras, Matthew Purver
机构
*
School of Electronic Engineering and Computer Science, Queen Mary University of London(伦敦女王学院电子工程与计算机科学学院)
;
Department of Knowledge Technologies, Jožef Stefan Institute(Jožef Stefan研究所知识技术系)