arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 2542 信号源:cs.CV, cs.GR, cs.MM

1. 效率与蒸馏 2542 篇

2405.18132 2024-05-29 cs.CV 57%

EG4D: Explicit Generation of 4D Object without Score Distillation

Qi Sun, Zhiyang Guo, Ziyu Wan, Jing Nathan Yan, Shengming Yin, Wengang Zhou, Jing Liao, Houqiang Li

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16570 2024-05-29 cs.CV cs.AI 57%

ID-to-3D: Expressive ID-guided 3D Heads via Score Distillation Sampling

Francesca Babiloni, Alexandros Lattas, Jiankang Deng, Stefanos Zafeiriou

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Explore our 3D results at: https://idto3d.github.io ; fixed broken url to project page

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17082 2024-05-21 cs.CV stat.ML 57%

DreamPropeller: Supercharge Text-to-3D Generation with Parallel Sampling

Linqi Zhou, Andy Shih, Chenlin Meng, Stefano Ermon

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Github repo: https://github.com/alexzhou907/DreamPropeller; Project page: https://alexzhou907.github.io/dreampropeller_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13573 2024-05-09 cs.CV cs.AI cs.LG 57%

ToDo: Token Downsampling for Efficient Generation of High-Resolution Images

Ethan Smith, Nayan Saxena, Aninda Saha

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Journal ref 2024, Proceedings of the Thirty-Third International Joint Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.11305 2024-05-08 cs.CV 57%

On Good Practices for Task-Specific Distillation of Large Pretrained Visual Models

Juliette Marrie, Michael Arbel, Julien Mairal, Diane Larlus

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Journal ref Published in Transactions on Machine Learning Research (TMLR), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.13745 2024-04-23 cs.CV cs.AI 57%

A Nasal Cytology Dataset for Object Detection and Deep Learning

Mauro Camporeale, Giovanni Dimauro, Matteo Gelardi, Giorgia Iacobellis, Mattia Sebastiano Ladisa, Sergio Latrofa, Nunzia Lomonte

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Pre Print almost ready to be submitted

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.11416 2024-04-18 cs.CV 57%

Neural Shrödinger Bridge Matching for Pansharpening

Zihan Cao, Xiao Wu, Liang-Jian Deng

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.10177 2024-04-18 cs.CV cs.AI 57%

TCJA-SNN: Temporal-Channel Joint Attention for Spiking Neural Networks

Rui-Jie Zhu, Malu Zhang, Qihang Zhao, Haoyu Deng, Yule Duan, Liang-Jian Deng

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Neural Networks and Learning Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.07150 2024-04-16 cs.RO cs.AI cs.CL cs.CV 57%

Embodied Agents for Efficient Exploration and Smart Scene Description

Roberto Bigazzi, Marcella Cornia, Silvia Cascianelli, Lorenzo Baraldi, Rita Cucchiara

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Accepted by IEEE International Conference on Robotics and Automation (ICRA 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07191 2024-04-16 cs.CV 57%

InstantMesh: Efficient 3D Mesh Generation from a Single Image with Sparse-view Large Reconstruction Models

Jiale Xu, Weihao Cheng, Yiming Gao, Xintao Wang, Shenghua Gao, Ying Shan

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Technical report. Project: https://github.com/TencentARC/InstantMesh

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10144 2024-04-11 cs.LG cs.AI cs.CV 57%

Data-Efficient Multimodal Fusion on a Single GPU

Noël Vouitsis, Zhaoyan Liu, Satya Krishna Gorti, Valentin Villecroze, Jesse C. Cresswell, Guangwei Yu, Gabriel Loaiza-Ganem, Maksims Volkovs

专题命中 效率与蒸馏 :text-to-image(abstract);分类 cs.CV

Comments CVPR 2024 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06353 2024-04-10 cs.LG cs.AI cs.CV 57%

High Noise Scheduling is a Must

Mahmut S. Gokmen, Cody Bumgardner, Jie Zhang, Ge Wang, Jin Chen

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06091 2024-04-10 cs.CV 57%

Hash3D: Training-free Acceleration for 3D Generation

Xingyi Yang, Xinchao Wang

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments https://adamdad.github.io/hash3D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00921 2024-04-02 cs.CV 57%

Towards Label-Efficient Human Matting: A Simple Baseline for Weakly Semi-Supervised Trimap-Free Human Matting

Beomyoung Kim, Myeong Yeon Yi, Joonsang Yu, Young Joon Yoo, Sung Ju Hwang

专题命中 效率与蒸馏 :image synthesis(abstract);分类 cs.CV

Comments Preprint, 15 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18575 2024-03-28 cs.CV 57%

HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions

Hao Xu, Haipeng Li, Yinqiao Wang, Shuaicheng Liu, Chi-Wing Fu

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14621 2024-03-22 cs.CV 57%

GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation

Yinghao Xu, Zifan Shi, Wang Yifan, Hansheng Chen, Ceyuan Yang, Sida Peng, Yujun Shen, Gordon Wetzstein

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Project page: https://justimyhxu.github.io/projects/grm/ Code: https://github.com/justimyhxu/GRM

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11887 2024-03-19 cs.CV cs.AI cs.LG 57%

SuperLoRA: Parameter-Efficient Unified Adaptation of Multi-Layer Attention Modules

Xiangyu Chen, Jing Liu, Ye Wang, Pu Perry Wang, Matthew Brand, Guanghui Wang, Toshiaki Koike-Akino

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments 33 pages, 29 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06213 2024-03-12 cs.CV cs.AI 57%

$V_kD:$ Improving Knowledge Distillation using Orthogonal Projections

Roy Miles, Ismail Elezi, Jiankang Deng

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

Comments CVPR 2024. Code available at https://github.com/roymiles/vkd

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05325 2024-03-11 cs.CV 57%

Fine-tuning a Multiple Instance Learning Feature Extractor with Masked Context Modelling and Knowledge Distillation

Juan I. Pisula, Katarzyna Bozek

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.06461 2024-02-16 cs.LG cs.CV stat.ML 57%

Sequential Flow Straightening for Generative Modeling

Jongmin Yoon, Juho Lee

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments 21 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09460 2024-01-17 stat.CO cs.CV cs.NA math.NA stat.ML 57%

Accelerated Bayesian imaging by relaxed proximal-point Langevin sampling

Teresa Klatzer, Paul Dobson, Yoann Altmann, Marcelo Pereyra, Jesús María Sanz-Serna, Konstantinos C. Zygalakis

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments 34 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16666 2024-01-11 cs.CV eess.IV 57%

SC-VAE: Sparse Coding-based Variational Autoencoder with Learned ISTA

Pan Xiao, Peijie Qiu, Sungmin Ha, Abdalla Bani, Shuang Zhou, Aristeidis Sotiras

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

Comments 21 pages, 23 figures, and 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10461 2023-12-21 cs.CV 57%

Rethinking the Up-Sampling Operations in CNN-based Generative Network for Generalizable Deepfake Detection

Chuangchuang Tan, Huan Liu, Yao Zhao, Shikui Wei, Guanghua Gu, Ping Liu, Yunchao Wei

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09510 2023-12-18 cs.CV cs.AI 57%

Fast Sampling generative model for Ultrasound image reconstruction

Hengrong Lan, Zhiqiang Li, Qiong He, Jianwen Luo

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments submitted to ISBI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.03934 2023-12-07 eess.IV cs.CV 57%

Inflating 2D Convolution Weights for Efficient Generation of 3D Medical Images

Yanbin Liu, Girish Dwivedi, Farid Boussaid, Frank Sanfilippo, Makoto Yamada, Mohammed Bennamoun

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

Comments Published at Computer Methods and Programs in Biomedicine (CMPB) 2023

Journal ref Computer Methods and Programs in Biomedicine (2023): 107685

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02189 2023-12-06 cs.CV cs.AI 57%

StableDreamer: Taming Noisy Score Distillation Sampling for Text-to-3D

Pengsheng Guo, Hans Hao, Adam Caccavale, Zhongzheng Ren, Edward Zhang, Qi Shan, Aditya Sankar, Alexander G. Schwing, Alex Colburn, Fangchang Ma

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16565 2023-12-05 cs.CV cs.SD eess.AS 57%

DiffusionTalker: Personalization and Acceleration for Speech-Driven 3D Face Diffuser

Peng Chen, Xiaobao Wei, Ming Lu, Yitong Zhu, Naiming Yao, Xingyu Xiao, Hui Chen

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.20082 2023-12-01 cs.CV 57%

Control4D: Efficient 4D Portrait Editing with Text

Ruizhi Shao, Jingxiang Sun, Cheng Peng, Zerong Zheng, Boyao Zhou, Hongwen Zhang, Yebin Liu

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments The link to our project website is https://control4darxiv.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.10123 2023-11-20 cs.CV 57%

MetaDreamer: Efficient Text-to-3D Creation With Disentangling Geometry and Texture

Lincong Feng, Muyu Wang, Maoyu Wang, Kuo Xu, Xiaoli Liu

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:2306.17843, arXiv:2209.14988 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.07372 2023-11-07 cs.CV eess.IV 57%

Image Compression with Product Quantized Masked Image Modeling

Alaaeldin El-Nouby, Matthew J. Muckley, Karen Ullrich, Ivan Laptev, Jakob Verbeek, Hervé Jégou

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏