arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 762 信号源:cs.CV, cs.GR, cs.MM

1. 其他图像生成 762 篇

1711.03180 2026-06-04 math.NA cs.NA cs.NE math.AP 50%

Deep D-bar: Real time Electrical Impedance Tomography Imaging with Deep Neural Networks

深度D-bar:基于深度神经网络的实时电阻抗断层成像

Sarah Jane Hamilton, Andreas Hauptmann

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出利用深度神经网络对电阻抗断层成像进行后处理,以提升图像清晰度和可靠性,通过模拟数据训练网络并应用于实验数据,实现高非线性逆问题的有效解决。

Comments 11 pages, 13 figures

Journal ref IEEE Transactions on Medical Imaging, 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.22902 2026-05-25 cs.LG cs.AI cs.CL 50%

Transcoders Trace Visual Grounding and Hallucinations in Vision-Language Models

Transcoders 追踪视觉语言模型中的视觉基础与幻觉

Dimitrios Damianos, Leon Voukoutis, Georgios Skyrianos, Vassilis Katsouros, Georgios Paraskevopoulos

机构 * Institute of Language and Speech Processing(语言与语音处理研究所) Athena Research Center(雅典研究中心)

专题命中 其他图像生成 :generative vision(abstract)

AI总结 采用基于Transcoders的功能中心框架分解视觉语言模型的计算路径,揭示视觉输入如何影响文本生成,并通过反事实分析和图结构特征预测幻觉。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.19227 2026-05-20 cs.CR cs.AI 50%

Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models

逐token被入侵:统一自回归模型中的后门漏洞

Tobias Braun, Jonas Henry Grebe, Hossein Shakibania, Anna Rohrbach, Marcus Rohrbach

机构 * TU Darmstadt(图宾根大学)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文研究了统一自回归模型中的后门漏洞问题,提出了一种名为Token by Token Backdoor Attack (ToBAC)的攻击方法,展示了如何通过数据和模型污染策略在多模态生成中引发有害行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07473 2026-05-13 quant-ph cond-mat.stat-mech cs.AI cs.ET cs.LG 50%

Breaking QAOA's Fixed Target Hamiltonian Barrier: A Fully Connected Quantum Boltzmann Machine via Bilevel Optimization

突破QAOA固定目标哈密顿量障碍:通过双层优化的全连接量子玻尔兹曼机

Jun Liu

机构 * School of Economics, Hunan University of Finance and Economics(湖南财经大学经济学院)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出全连接量子玻尔兹曼机,通过双层优化架构改进传统QAOA电路,展示模型在单一层数下具有高测量精度,并在噪声环境下表现出强鲁棒性,同时在图像生成中也表现出稳定性。

Comments 34 pages, 8 figures, 3 tables, 1 algorithm

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15641 2026-05-01 cs.CR 50%

ComMark: Covert and Robust Black-Box Model Watermarking with Compressed Samples

ComMark:基于压缩样本的隐蔽且鲁棒的黑盒模型水印技术

Yunfei Yang, Xiaojun Chen, Zhendong Zhao, Yu Zhou, Xiaoyan Gu, Juan Cao

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出ComMark,一种利用频域变换生成隐蔽且抗攻击的压缩水印样本的黑盒模型水印框架,通过过滤高频信息提升隐蔽性和鲁棒性,适用于图像识别、语音识别等多种任务。

Comments Extended version of the paper accepted by ICMR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11567 2026-04-29 cs.CY 50%

Consumer Law for AI Agents

人工智能代理的消费者法

Christoph Busch

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文探讨欧盟现行消费者法是否能适应人工智能代理带来的变革,分析其对电子商务和人类中心消费法的挑战,并提出未来兼顾人类与机器的消费者法框架。

Journal ref German Law Journal 26 (2025) 1367-1382

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.19431 2026-04-22 cs.LO cs.AI 50%

Counting Worlds Branching Time Semantics for post-hoc Bias Mitigation in generative AI

基于分支时间语义的计数世界语义用于生成AI的事后偏差缓解

Alessandro G. Buda, Giuseppe Primiero, Leonardo Ceragioli, Melissa Antonelli

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出CTLF分支时间逻辑,用于分析生成AI输出序列中的偏差,通过计数世界语义验证输出序列是否符合保护属性的概率分布,预测生成新输出时的偏差范围,以及确定需移除的输出数量以恢复公平性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09583 2026-04-14 cs.HC 50%

Trace-Aware Workflows for Co-Creating Branded Content with Generative AI

具有痕迹意识的工作流:利用生成式AI共同创建品牌内容

Taehyun Yang, Eunhye Kim, Zhongzheng Xu, Fumeng Yang

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文探讨小型企业主在利用生成式AI进行品牌内容创作时面临的挑战,并提出一个原型系统,通过痕迹记录支持迭代流程中的反馈探索与内容优化。

Comments Accepted to CHI 2026 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18388 2026-04-07 cs.HC cs.AI 50%

Exploration vs. Fixation: Scaffolding Divergent and Convergent Thinking for Human-AI Co-Creation with Generative Models

探索与固定:基于生成模型的人机协同创作中发散与聚合思维的支架设计

Chao Wen, Tung Phung, Pronita Mehrotra, Sumit Gulwani, Roger E. Beaty, Tomohiro Nagashima, Adish Singla

机构 * Max Planck Institute for Software Systems(马克斯·普朗克软件系统研究所) Microsoft(微软) Pennsylvania State University(宾夕法尼亚州立大学) Saarland Informatics Campus, Saarland University(萨尔大学萨尔信息学园区)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出HAICo系统,通过发散与聚合两种模式引导人机协同创作,提升创意和可用性。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06755 2026-03-17 cs.LG quant-ph 50%

Implementation of Quantum Implicit Neural Representation in Deterministic and Probabilistic Autoencoders for Image Reconstruction/Generation Tasks

在确定性和概率性自编码器中实现量子隐式神经表示用于图像重建/生成任务

Saadet Müzehher Eren

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出基于量子隐式神经表示(QINR)的自编码器和变分自编码器,用于图像重建和生成任务,通过在潜在空间中转换信息生成丰富特征,并展示QINR-VAE在图像生成中比量子生成对抗网络更稳定。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19743 2026-03-12 cs.SE cs.AI 50%

What Makes Code Generation Ethically Sourced?

什么使代码生成具有伦理来源?

Zhuolin Xu, Chenglin Li, Qiushi Li, Shin Hwei Tan

机构 * Concordia University(康科迪亚大学)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出伦理来源代码生成(ES-CodeGen)的概念,通过文献综述和调查,确定了11个关键维度,强调代码质量和伦理问题的重要性。

Journal ref Proc. 48th International Conference on Software Engineering (ICSE 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04245 2026-03-05 cs.SE cs.AI cs.HC 50%

LikeThis! Empowering App Users to Submit UI Improvement Suggestions Instead of Complaints

LikeThis! 使应用用户能够提交UI改进建议而不是投诉

Jialiang Wei, Ali Ebrahimi Pourasad, Walid Maalej

机构 * University of Hamburg(汉堡大学)

专题命中 其他图像生成 :image generation(abstract)

AI总结 LikeThis! 通过生成式AI帮助用户提交具体的UI改进建议,提升用户与开发者的协作效率。

Comments Accepted at 2026 IEEE/ACM 48th International Conference on Software Engineering (ICSE '26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09844 2026-03-05 quant-ph cs.LG 50%

On the Generalization Limits of Quantum Generative Adversarial Networks with Pure State Generators

量子生成对抗网络在纯态生成器中的泛化极限研究

Jasmin Frkatovic, Akash Malemath, Ivan Kankeu, Yannick Werner, Matthias Tschöpe, Vitor Fortes Rey, Sungho Suh, Paul Lukowicz, Nikolaos Palaiodimopoulos, Maximilian Kiefer-Emmanouilidis

机构 * Department of Computer Science and Research Initiative QC-AI(计算机科学系和QC-AI研究计划) RPTU Kaiserslautern-Landau(凯撒斯劳滕-兰道大学) Embedded Intelligence(嵌入式智能) German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心) Department of Artificial Intelligence(人工智能系) Korea University(韩国大学) Department of Physics(物理系)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本研究探讨了量子生成对抗网络在纯态生成器中的泛化极限,推导出判别器质量的理论下限,揭示了现有量子生成模型在泛化能力上的根本挑战。

Comments 20 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20101 2026-02-25 astro-ph.SR astro-ph.IM 50%

HelioSpectrotron 5000: An interactive multi-resolution solar spectral atlas

HelioSpectrotron 5000:一个交互式多分辨率太阳光谱图集

A. G. M. Pietrow

专题命中 其他图像生成 :image generation(abstract)

AI总结 HS5000通过多分辨率光谱数据和交互式功能,实现了高分辨率参考光谱与地面观测的高效对比与分析。

Comments Published in the Open Journal of Astrophysics, 5 pages, 3 figures

Journal ref Open Journal of Astrophysics, 9 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01527 2026-02-23 cs.HC 50%

PlayFutures: Imagining Civic Futures with AI and Puppets

PlayFutures: 用人工智能和木偶想象公民的未来

Supratim Pait, Sumita Sharma, Ashley Frith, Michael Nitsche, Noura Howell

专题命中 其他图像生成 :image generation(abstract)

AI总结 PlayFutures通过结合人工智能和木偶制作,探讨如何利用新技术重新想象公民空间的未来,并通过儿童参与的游戏和表演促进公民参与。

Comments This version is a position paper presented at the "CHI 2024 Workshop on Child-centred AI Design, May 11, 2024, Honolulu, HI, USA." Updated version of paper accepted In. Proc. ACM SIGCSE TS 2026 V.2

Journal ref PlayFutures: Imagining Civic Futures with AI and Puppets. In Proceedings of the 57th ACM Technical Symposium on Computer Science Education V.2 (SIGCSE TS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22093 2026-01-30 cs.CY cs.AI 50%

Investigating Associational Biases in Inter-Model Communication of Large Generative Models

研究大生成模型跨模型通信中的关联偏差

Fethiye Irmak Dogan, Yuval Weiss, Kajal Patel, Jiaee Cheong, Hatice Gunes

机构 * University of Cambridge(剑桥大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Harvard University(哈佛大学)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本研究探讨大生成模型跨模型通信中的关联偏差,通过数据集分析发现人口分布漂移,并提出缓解策略以减少不平等影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20462 2026-01-29 cs.CE physics.comp-ph 50%

CM-GAI: Continuum Mechanistic Generative Artificial Intelligence Theory for Data Dynamics

CM-GAI:持续力学生成人工智能理论用于数据动态

Shan Tang, Ziwei Cao, Zhenling Yang, Jiachen Guo, Yicheng Lu, Wing Kam Liu, Xu Guo

专题命中 其他图像生成 :image generation(abstract)

AI总结 CM-GAI通过连续力学理论框架,利用少量数据生成材料、结构和系统层面的动态响应,为工程应用提供新工具。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17761 2026-01-27 cs.LG cs.AI cs.CL 50%

AR-Omni: A Unified Autoregressive Model for Any-to-Any Generation

AR-Omni:一种统一的自回归模型用于任意到任意生成

Dongjie Cheng, Ruifeng Yuan, Yongqi Li, Runyang You, Wenjie Wang, Liqiang Nie, Lei Zhang, Wenjie Li

机构 * The Hong Kong Polytechnic University(香港理工大学) University of Science and Technology of China(中国科学技术大学) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))

专题命中 其他图像生成 :image generation(abstract)

AI总结 AR-Omni提出一种无需专家解码器的统一自回归模型,实现多模态任意到任意生成,解决模态不平衡、视觉保真度和稳定性与创造力平衡问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17096 2026-01-27 cs.CY cs.AI cs.CL 50%

Beyond Instrumental and Substitutive Paradigms: Introducing Machine Culture as an Emergent Phenomenon in Large Language Models

超越工具性和替代性范式:引入机器文化作为大型语言模型中的涌现现象

Yueqing Hu, Xinyang Peng, Yukun Zhao, Lin Qiu, Ka-lai Hung, Kaiping Peng

专题命中 其他图像生成 :image generation(abstract)

AI总结 本研究提出机器文化作为大型语言模型中的一种新兴现象,挑战传统工具性和替代性范式,揭示模型在文化表现上的独特特性。

Comments 16 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13901 2026-01-21 astro-ph.IM 50%

Prospecting MeerKAT Continuum Data for Enigmatic Radio Sources with Unsupervised Vector-Quantised Variational Autoencoders

利用无监督向量量化变分自编码器探测恩igmatic射电源的MeerKAT连续数据

Fernando L. Ventura, Kshitij Thorat, Anna Bosman, Roger Deane, Christopher Cleghorn

专题命中 其他图像生成 :image generation(abstract)

AI总结 利用无监督向量量化变分自编码器探测MeerKAT连续数据中的异常射电源,提升大规模数据搜索效率。

Comments 12 pages, 14 figures, submitted to RAS Techniques and Instruments

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12979 2026-01-19 cs.LG cs.AI 50%

A Distributed Generative AI Approach for Heterogeneous Multi-Domain Environments under Data Sharing constraints

在数据共享约束下用于异构多域环境的分布式生成AI方法

Youssef Tawfilis, Hossam Amer, Minar El-Aasser, Tallal Elshabrawy

专题命中 其他图像生成 :image generation(abstract)

AI总结 本研究提出了一种在数据共享约束下用于异构多域环境的分布式生成AI方法,结合KLD加权聚类联邦学习和异构U型分割学习,提升多域生成模型的训练效率和性能。

Comments Accepted and published in Transactions on Machine Learning Research (TMLR), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.01097 2026-01-06 stat.ML cs.LG 50%

Neural Networks on Symmetric Spaces of Noncompact Type

在非紧致型对称空间上神经网络

Xuan Son Nguyen, Shuo Yang, Aymeric Histace

机构 * ETIS, UMR 8051, CY Cergy Paris University, ENSEA, CNRS, France(ETIS研究所、法国巴黎中央理工学院、ENSEA、CNRS)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出了一种在非紧致型对称空间上设计神经网络的新方法,通过统一的距离公式实现完全连接层和注意力机制,验证了其在图像分类等任务中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08365 2025-12-10 cs.DC cs.LG 50%

Magneton: Optimizing Energy Efficiency of ML Systems via Differential Energy Debugging

Magneton: 通过微分能量调试优化机器学习系统的能效

Yi Pan, Wenbo Qian, Dedong Xie, Ruiyan Hu, Yigong Hu, Baris Kasikci

机构 * University of Washington(华盛顿大学) Boston University(波士顿大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他图像生成 :image generation(abstract)

AI总结 Magneton 通过微分能量调试技术,识别并诊断 ML 系统中的能耗低效问题,发现多个已知和未知的能耗浪费案例。

Comments 12 pages, 10 fi

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04109 2025-12-05 cs.CY 50%

ChatGPT-5 in Secondary Education: A Mixed-Methods Analysis of Student Attitudes, AI Anxiety, and Hallucination-Aware Use

ChatGPT-5在中学教育中的应用:学生态度、AI焦虑与幻觉意识使用混合方法分析

Tryfon Sivenas

专题命中 其他图像生成 :image generation(abstract)

AI总结 研究通过混合方法分析ChatGPT-5在中学教育中的应用,探讨学生态度、AI焦虑及幻觉意识使用,揭示教学优势与限制,提出知识保障策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22113 2025-12-02 cs.LG 50%

Countering adversarial evasion in regression analysis

对抗回归分析中的对抗逃逸

David Benfield, Phan Tu Vuong, Alain Zemkoho

机构 * School of Mathematical Sciences University of Southampton(数学科学学院 英国南安普顿大学)

专题命中 其他图像生成 :image generation(abstract)

AI总结 本文提出了一种适用于回归场景的悲观双层优化方法,以应对对抗逃逸问题,无需假设对手解的凸性或唯一性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00346 2025-11-04 cs.CR cs.AI cs.LG 50%

Exploiting Latent Space Discontinuities for Building Universal LLM Jailbreaks and Data Extraction Attacks

Kayua Oleques Paim, Rodrigo Brandao Mansilha, Diego Kreutz, Muriel Figueredo Franco, Weverton Cordeiro

专题命中 其他图像生成 :image generation(abstract)

Comments 10 pages, 5 figures, 4 tables, Published at the Brazilian Symposium on Cybersecurity (SBSeg 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04925 2025-11-04 cs.SE cs.AI 50%

Why Attention Fails: A Taxonomy of Faults in Attention-Based Neural Networks

Sigma Jahan, Saurabh Singh Rajput, Tushar Sharma, Mohammad Masudur Rahman

机构 * Dalhousie University(达尔豪斯大学)

专题命中 其他图像生成 :image generation(abstract)

Journal ref IEEE/ACM 48th International Conference on Software Engineering (ICSE) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15827 2025-11-03 cs.LG cs.AI 50%

DualOptim: Enhancing Efficacy and Stability in Machine Unlearning with Dual Optimizers

Xuyang Zhong, Haochen Luo, Chen Liu

专题命中 其他图像生成 :image generation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18156 2025-10-10 cs.CY cs.AI 50%

Adoption of Watermarking for Generative AI Systems in Practice and Implications under the new EU AI Act

Bram Rijsbosch, Gijs van Dijck, Konrad Kollnig

机构 * Maastricht University - Law and Tech Lab(马斯特里赫特大学-法律与科技实验室) Maastricht University - Law(马斯特里赫特大学-法律) Tech Lab The Netherlands(科技实验室荷兰) Tech Lab(科技实验室)

专题命中 其他图像生成 :image generation(abstract)

Comments Note that this work has not yet been published in a peer review journal, it is therefore potentially still subject to change. Update October 2025: extended and slightly revised arguments in legal analysis, results unchanged

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03254 2025-10-07 cs.LG cs.CR 50%

Adversarial training with restricted data manipulation

David Benfield, Stefano Coniglio, Phan Tu Vuong, Alain Zemkoho

机构 * School of Mathematical Sciences University of Southampton(南安普顿大学数学科学学院) Department of Economic Sciences University of Bergamo(伯尔戈米大学经济科学学院)

专题命中 其他图像生成 :image generation(abstract)

Comments 21 page, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏