arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

视频大模型

视频理解、视频生成、视频语言模型和时序视觉推理。

共收录 2143 信号源:cs.CV, eess.IV, cs.MM

1. 视频生成 2143 篇

2403.16112 2024-03-26 cs.CV cs.AI cs.LG 57%

Opportunities and challenges in the application of large artificial intelligence models in radiology

Liangrui Pan, Zhenyu Zhao, Ying Lu, Kewei Tang, Liyong Fu, Qingchun Liang, Shaoliang Peng

专题命中 视频生成 :video generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09334 2024-03-26 cs.CV 57%

Video Editing via Factorized Diffusion Distillation

Uriel Singer, Amit Zohar, Yuval Kirstain, Shelly Sheynin, Adam Polyak, Devi Parikh, Yaniv Taigman

专题命中 视频生成 :video generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14186 2024-03-22 cs.CV cs.AI cs.GR 57%

StyleCineGAN: Landscape Cinemagraph Generation using a Pre-trained StyleGAN

Jongwoo Choi, Kwanggyoon Seo, Amirsaman Ashtari, Junyong Noh

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Project website: https://jeolpyeoni.github.io/stylecinegan_project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12035 2024-03-19 cs.CV 57%

CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility

Bojia Zi, Shihao Zhao, Xianbiao Qi, Jianan Wang, Yukai Shi, Qianyu Chen, Bin Liang, Kam-Fai Wong, Lei Zhang

专题命中 视频生成 :video generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08268 2024-03-14 cs.CV 57%

Follow-Your-Click: Open-domain Regional Image Animation via Short Prompts

Yue Ma, Yingqing He, Hongfa Wang, Andong Wang, Chenyang Qi, Chengfei Cai, Xiu Li, Zhifeng Li, Heung-Yeung Shum, Wei Liu, Qifeng Chen

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Project Page: https://follow-your-click.github.io/ Github Page: https://github.com/mayuelala/FollowYourClick

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07198 2024-03-13 cs.CV 57%

Action Reimagined: Text-to-Pose Video Editing for Dynamic Human Actions

Lan Wang, Vishnu Boddeti, Sernam Lim

专题命中 视频生成 :video understanding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17403 2024-02-28 cs.CV 57%

Sora Generates Videos with Stunning Geometrical Consistency

Xuanyi Li, Daquan Zhou, Chenxu Zhang, Shaodong Wei, Qibin Hou, Ming-Ming Cheng

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments 5 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17139 2024-02-28 cs.CV cs.AI 57%

Video as the New Language for Real-World Decision Making

Sherry Yang, Jacob Walker, Jack Parker-Holder, Yilun Du, Jake Bruce, Andre Barreto, Pieter Abbeel, Dale Schuurmans

专题命中 视频生成 :video generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.01107 2024-02-27 cs.CV cs.AI cs.LG stat.ML 57%

Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models

Hyeonho Jeong, Jong Chul Ye

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments Accepted to ICLR 2024, Project Page: http://ground-a-video.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.05957 2024-02-09 cs.CV cs.LG 57%

DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles

Tal Daniel, Aviv Tamar

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments TMLR 2024. Project site: https://taldatech.github.io/ddlp-web

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.17599 2024-01-05 cs.CV 57%

Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models

Wen Wang, Yan Jiang, Kangyang Xie, Zide Liu, Hao Chen, Yue Cao, Xinlong Wang, Chunhua Shen

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments Add customized video editing. Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03775 2023-12-22 cs.CV 57%

FAAC: Facial Animation Generation with Anchor Frame and Conditional Control for Superior Fidelity and Editability

Linze Li, Sunqi Fan, Hengjun Pu, Zhaodong Bing, Yao Tang, Tianzhu Ye, Tong Yang, Liangyu Chen, Jiajun Liang

专题命中 视频生成 :video generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10656 2023-12-20 cs.CV 57%

VidToMe: Video Token Merging for Zero-Shot Video Editing

Xirui Li, Chao Ma, Xiaokang Yang, Ming-Hsuan Yang

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Project page: https://vidtome-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02503 2023-12-06 cs.CV 57%

SAVE: Protagonist Diversification with Structure Agnostic Video Editing

Yeji Song, Wonsik Shin, Junsoo Lee, Jeesoo Kim, Nojun Kwak

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments Project website: https://ldynx.github.io/SAVE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.18670 2023-12-04 cs.CV 57%

SAVE: Spectral-Shift-Aware Adaptation of Image Diffusion Models for Text-driven Video Editing

Nazmul Karim, Umar Khalid, Mohsen Joneidi, Chen Chen, Nazanin Rahnavard

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments 11 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17098 2023-11-29 cs.CV 57%

ControlVideo: Conditional Control for One-shot Text-driven Video Editing and Beyond

Min Zhao, Rongzhen Wang, Fan Bao, Chongxuan Li, Jun Zhu

专题命中 视频生成 :long video(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.08289 2023-11-29 cs.CV cs.GR 57%

Continuously Controllable Facial Expression Editing in Talking Face Videos

Zhiyao Sun, Yu-Hui Wen, Tian Lv, Yanan Sun, Ziyang Zhang, Yaoyuan Wang, Yong-Jin Liu

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Affective Computing (DOI: 10.1109/TAFFC.2023.3334511). Demo video: https://youtu.be/WD-bNVya6kM . Project page: https://raineggplant.github.io/FEE4TV

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15769 2023-11-28 cs.CV 57%

Side4Video: Spatial-Temporal Side Network for Memory-Efficient Image-to-Video Transfer Learning

Huanjin Yao, Wenhao Wu, Zhiheng Li

专题命中 视频生成 :video understanding(abstract);分类 cs.CV

Comments Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09777 2023-11-28 cs.CV 57%

DriveDreamer: Towards Real-world-driven World Models for Autonomous Driving

Xiaofeng Wang, Zheng Zhu, Guan Huang, Xinze Chen, Jiagang Zhu, Jiwen Lu

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Project Page: https://drivedreamer.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.07945 2023-11-20 cs.CV 57%

Edit-A-Video: Single Video Editing with Object-Aware Consistency

Chaehun Shin, Heeseung Kim, Che Hyun Lee, Sang-gil Lee, Sungroh Yoon

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments ACML 2023 Best Paper Award

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09535 2023-10-12 cs.CV 57%

FateZero: Fusing Attentions for Zero-shot Text-based Video Editing

Chenyang Qi, Xiaodong Cun, Yong Zhang, Chenyang Lei, Xintao Wang, Ying Shan, Qifeng Chen

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments Accepted to ICCV 2023 as an Oral Presentation. Project page: https://fate-zero-edit.github.io ; GitHub repository: https://github.com/ChenyangQiQi/FateZero

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.06309 2023-10-11 cs.MM cs.HC 57%

Encoding and Decoding Narratives: Datafication and Alternative Access Models for Audiovisual Archives

Yuchen Yang

专题命中 视频生成 :text-to-video(abstract);分类 cs.MM

Comments arXiv admin note: substantial text overlap with arXiv:2310.05825

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05835 2023-10-10 cs.MM cs.HC 57%

Latent Wander: an Alternative Interface for Interactive and Serendipitous Discovery of Large AV Archives

Yuchen Yang, Linyida Zhang

专题命中 视频生成 :text-to-video(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16148 2023-09-29 cs.CV 57%

OSM-Net: One-to-Many One-shot Talking Head Generation with Spontaneous Head Motions

Jin Liu, Xi Wang, Xiaomeng Fu, Yesheng Chai, Cai Yu, Jiao Dai, Jizhong Han

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Paper Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.10091 2023-09-20 cs.CV cs.AI cs.CL cs.LG 57%

Unified Coarse-to-Fine Alignment for Video-Text Retrieval

Ziyang Wang, Yi-Lin Sung, Feng Cheng, Gedas Bertasius, Mohit Bansal

专题命中 视频生成 :text-to-video(abstract);分类 cs.CV

Comments ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.00810 2023-09-06 cs.CV cs.AI 57%

RenAIssance: A Survey into AI Text-to-Image Generation in the Era of Large Model

Fengxiang Bie, Yibo Yang, Zhongzhu Zhou, Adam Ghanem, Minjia Zhang, Zhewei Yao, Xiaoxia Wu, Connor Holmes, Pareesa Golnari, David A. Clifton, Yuxiong He, Dacheng Tao, Shuaiwen Leon Song

专题命中 视频生成 :video generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.15472 2023-08-30 cs.CV 57%

Learning Modulated Transformation in GANs

Ceyuan Yang, Qihang Zhang, Yinghao Xu, Jiapeng Zhu, Yujun Shen, Bo Dai

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14748 2023-08-29 cs.GR cs.CV 57%

MagicAvatar: Multimodal Avatar Generation and Animation

Jianfeng Zhang, Hanshu Yan, Zhongcong Xu, Jiashi Feng, Jun Hao Liew

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Project page: https://magic-avatar.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14392 2023-08-29 cs.CV 57%

1st Place Solution for the 5th LSVOS Challenge: Video Instance Segmentation

Tao Zhang, Xingye Tian, Yikang Zhou, Yu Wu, Shunping Ji, Cilin Yan, Xuebo Wang, Xin Tao, Yuan Zhang, Pengfei Wan

专题命中 视频生成 :long video(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.03504 2023-08-03 cs.CV cs.SD eess.AS 57%

Ada-TTA: Towards Adaptive High-Quality Text-to-Talking Avatar Synthesis

Zhenhui Ye, Ziyue Jiang, Yi Ren, Jinglin Liu, Chen Zhang, Xiang Yin, Zejun Ma, Zhou Zhao

专题命中 视频生成 :video generation(abstract);分类 cs.CV

Comments Accepted by ICML 2023 Workshop, 6 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏