CommentsPublished in ICACIn 2024. Appears in Advances on Intelligent Computing and Data Science II, Lecture Notes on Data Engineering and Communications Technologies, vol. 254, Springer, 2025
Journal refAdvances on Intelligent Computing and Data Science II (ICACIn 2024), Lecture Notes on Data Engineering and Communications Technologies, vol. 254, Springer, Cham, 2025
机构
*
The University of Tokyo(东京大学)
;
Nara Institute of Science and Technology(奈良科学技术研究所)
;
Chungnam National University(忠南国立大学)
;
Institute of Science Tokyo(东京科学大学)
PRISM: Prompt Refinement via Image-grounded Self-rewarding Mechanism for Text-to-Image Generation
PRISM:通过基于图像的自我奖励机制进行文本到图像生成的提示优化
Guo Tang, HongJie Luo, Tianxu Wang, Ying Zhang, Hao Wang
机构
*
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Sun Yat-sen University(中山大学)
;
South China University of Technology(华南理工大学)
;
Guangdong University of Technology(广东工业大学)
机构
*
Institute of Artificial Intelligence, State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(北京航空航天大学虚拟现实技术与系统国家重点实验室人工智能研究院)
;
College of Artificial Intelligence, Tsinghua University(清华大学人工智能学院)
;
Security Department, Alibaba Group(阿里巴巴集团安全部)
;
School of Cyber Science and Technology, Shenzhen Campus of Sun Yat-Sen University(中山大学深圳校区网络科学与技术学院)
Haodong Lei, Hongsong Wang, Bingxuan Dai, Pan Zhou
机构
*
College of Software Engineering, Southeast University(东南大学软件工程学院)
;
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
School of Cyber Science and Engineering, Southeast University(东南大学网络空间安全学院)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算机与信息系统学院)
On the Limitations of Cross-Lingual Consistency in Multilingual Text-to-image Generation
多语言文本到图像生成中跨语言一致性的局限性研究
Sicheng Zhang, Zhonghao Yan, Binzhu Xie, Shi Qiu, Muzammal Naseer, Naveed Akhtar, Mubarak Shah
机构
*
Khalifa University(哈利法大学)
;
Queen Mary University of London(伦敦玛丽女王大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The University of Western Australia(西澳大学)
;
The University of Melbourne(墨尔本大学)
;
University of Central Florida(中佛罗里达大学)
AEGIS: A Mechanism-Guided Defense against Visual Synonym Jailbreaks in Text-to-Image Models
AEGIS:一种针对文本到图像模型中视觉同义词越狱的机制引导防御
Yuanmin Huang, Zhenfei Zhang, Mi Zhang, Geng Hong, Qinqin He, Jialing Tao, Hui Xue, Min Yang
机构
*
Fudan University(复旦大学)
;
Alibaba Group(阿里巴巴集团)
;
Shanghai Pudong Research Institute of Cryptology(上海浦东密码研究所)
;
Engineering Research Center of Cyber Security Auditing and Monitoring, Ministry of Education(教育部网络安全审计与监测工程研究中心)
Decoupled Guidance: Disentangling Subject and Context Pathways in Text-to-Image Personalization
解耦引导:文本到图像个性化中的主体与上下文路径分离
Seongmin Kim, Kyucheol Shin, Heesun Jung, Jinseo Kim, Sungyong Baik
机构
*
Dept. of Artificial Intelligence, Hanyang University, South Korea(韩国汉阳大学人工智能系)
;
Dept. of Data Science, Hanyang University, South Korea(韩国汉阳大学数据科学系)