HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments Published in International Conference on Learning Representations (ICLR) 2025