arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

视频大模型

视频理解、视频生成、视频语言模型和时序视觉推理。

共收录 470 信号源:cs.CV, eess.IV, cs.MM

1. 动作与事件理解 470 篇

2308.12962 2023-08-25 cs.CV 57%

Motion-Guided Masking for Spatiotemporal Representation Learning

David Fan, Jue Wang, Shuai Liao, Yi Zhu, Vimal Bhat, Hector Santos-Villalobos, Rohith MV, Xinyu Li

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Accepted to ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16803 2023-08-01 cs.CV 57%

DPMix: Mixture of Depth and Point Cloud Video Experts for 4D Action Segmentation

Yue Zhang, Hehe Fan, Yi Yang, Mohan Kankanhalli

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.07754 2023-07-19 cs.CV cs.AI 57%

Bidirectionally Deformable Motion Modulation For Video-based Human Pose Transfer

Wing-Yin Yu, Lai-Man Po, Ray C. C. Cheung, Yuzhi Zhao, Yu Xue, Kun Li

专题命中 动作与事件理解 :video generation(abstract);分类 cs.CV

Comments ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.08351 2023-05-31 physics.optics eess.IV physics.bio-ph 57%

Parallelized computational 3D video microscopy of freely moving organisms at multiple gigapixels per second

Kevin C. Zhou, Mark Harfouche, Colin L. Cooke, Jaehee Park, Pavan C. Konda, Lucas Kreiss, Kanghyun Kim, Joakim Jönsson, Jed Doman, Paul Reamey, Veton Saliu, Clare B. Cook, Maxwell Zheng, Jack P. Bechtel, Aurélien Bègue, Matthew McCarroll, Jennifer Bagwell, Gregor Horstmeyer, Michel Bagnat, Roarke Horstmeyer

专题命中 动作与事件理解 :long video(abstract);分类 eess.IV

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11975 2023-04-25 cs.CV 57%

MRSN: Multi-Relation Support Network for Video Action Detection

Yin-Dong Zheng, Guo Chen, Minglei Yuan, Tong Lu

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11082 2023-04-25 cs.CV 57%

DynIBaR: Neural Dynamic Image-Based Rendering

Zhengqi Li, Qianqian Wang, Forrester Cole, Richard Tucker, Noah Snavely

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

Comments Award Candidate, CVPR 2023 Project page: dynibar.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09448 2023-04-20 cs.LG cs.CL cs.CV 57%

EC^2: Emergent Communication for Embodied Control

Yao Mu, Shunyu Yao, Mingyu Ding, Ping Luo, Chuang Gan

专题命中 动作与事件理解 :video-language(abstract);分类 cs.CV

Comments Published in CVPR2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.02717 2023-04-12 cs.CV 57%

BasicTAD: an Astounding RGB-Only Baseline for Temporal Action Detection

Min Yang, Guo Chen, Yin-Dong Zheng, Tong Lu, Limin Wang

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Accepted by CVIU

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.00254 2023-04-05 cs.CV 57%

DOAD: Decoupled One Stage Action Detection Network

Shuning Chang, Pichao Wang, Fan Wang, Jiashi Feng, Mike Zheng Show

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments 11 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.08063 2023-03-21 cs.CV 57%

MINOTAUR: Multi-task Video Grounding From Multimodal Queries

Raghav Goyal, Effrosyni Mavroudi, Xitong Yang, Sainbayar Sukhbaatar, Leonid Sigal, Matt Feiszli, Lorenzo Torresani, Du Tran

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments 22 pages, 8 figures and 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09114 2023-03-17 cs.CV 57%

AU-aware graph convolutional network for Macro- and Micro-expression spotting

Shukang Yin, Shiwei Wu, Tong Xu, Shifeng Liu, Sirui Zhao, Enhong Chen

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

Comments Accepted by ICME-2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09055 2023-03-17 cs.CV 57%

TemporalMaxer: Maximize Temporal Context with only Max Pooling for Temporal Action Localization

Tuan N. Tang, Kwonyoung Kim, Kwanghoon Sohn

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.03684 2023-03-17 cs.CV 57%

MOSO: Decomposing MOtion, Scene and Object for Video Prediction

Mingzhen Sun, Weining Wang, Xinxin Zhu, Jing Liu

专题命中 动作与事件理解 :video generation(abstract);分类 cs.CV

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.02906 2023-03-07 cs.CV 57%

MotionVideoGAN: A Novel Video Generator Based on the Motion Space Learned from Image Pairs

Jingyuan Zhu, Huimin Ma, Jiansheng Chen, Jian Yuan

专题命中 动作与事件理解 :video generation(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Multimedia as a regular paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.03125 2023-03-03 cs.CV 57%

Self-supervised and Weakly Supervised Contrastive Learning for Frame-wise Action Representations

Minghao Chen, Renbo Tu, Chenxi Huang, Yuqi Lin, Boxi Wu, Deng Cai

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

Comments author conflicts

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.07134 2022-11-29 cs.CV 57%

ETAD: Training Action Detection End to End on a Laptop

Shuming Liu, Mengmeng Xu, Chen Zhao, Xu Zhao, Bernard Ghanem

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.13425 2022-11-23 cs.CV 57%

Do we really need temporal convolutions in action segmentation?

Dazhao Du, Bing Su, Yu Li, Zhongang Qi, Lingyu Si, Ying Shan

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.03064 2022-10-28 cs.CV 57%

A Simple and Efficient Pipeline to Build an End-to-End Spatial-Temporal Action Detector

Lin Sui, Chen-Lin Zhang, Lixin Gu, Feng Han

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Accepted By WACV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.10474 2022-10-12 q-bio.QM cs.CV 57%

Rapid detection and recognition of whole brain activity in a freely behaving Caenorhabditis elegans

Yuxiang Wu, Shang Wu, Xin Wang, Chengtian Lang, Quanshi Zhang, Quan Wen, Tianqi Xu

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

Journal ref PLOS Computational Biology 18(10): e1010594, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04889 2022-10-11 cs.CV 57%

Turbo Training with Token Dropout

Tengda Han, Weidi Xie, Andrew Zisserman

专题命中 动作与事件理解 :video-language(abstract);分类 cs.CV

Comments BMVC2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04331 2022-10-11 cs.CV 57%

Students taught by multimodal teachers are superior action recognizers

Gorjan Radevski, Dusan Grujicic, Matthew Blaschko, Marie-Francine Moens, Tinne Tuytelaars

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Extended abstract accepted at the 2nd Ego4D Workshop @ ECCV 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.14757 2022-09-30 cs.CV 57%

Speeding Up Action Recognition Using Dynamic Accumulation of Residuals in Compressed Domain

Ali Abdari, Pouria Amirjan, Azadeh Mansouri

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.11908 2022-09-16 cs.CV 57%

Adaptive Perception Transformer for Temporal Action Localization

Yizheng Ouyang, Tianjin Zhang, Weibo Gu, Hongfa Wang

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.12673 2022-08-29 cs.CV cs.LG 57%

Enabling Weakly-Supervised Temporal Action Localization from On-Device Learning of the Video Stream

Yue Tang, Yawen Wu, Peipei Zhou, Jingtong Hu

专题命中 动作与事件理解 :long video(abstract);分类 cs.CV

Comments Manuscript received April 07, 2022; revised June 11, 2022; accepted July 05, 2022. This article was presented in the International Conference on 2022 and appears as part of the ESWEEK-TCAD special issue

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.10271 2022-08-12 cs.CV 57%

End-to-end Temporal Action Detection with Transformer

Xiaolong Liu, Qimeng Wang, Yao Hu, Xu Tang, Shiwei Zhang, Song Bai, Xiang Bai

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Image Processing (TIP). Code: https://github.com/xlliu7/TadTR

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.01159 2022-08-09 cs.CV 57%

BATMAN: Bilateral Attention Transformer in Motion-Appearance Neighboring Space for Video Object Segmentation

Ye Yu, Jialin Yuan, Gaurav Mittal, Li Fuxin, Mei Chen

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Accepted by ECCV 2022 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.10448 2022-07-22 cs.CV 57%

An Efficient Spatio-Temporal Pyramid Transformer for Action Detection

Yuetian Weng, Zizheng Pan, Mingfei Han, Xiaojun Chang, Bohan Zhuang

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments Accepted to ECCV 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.10213 2022-07-22 cs.CV 57%

Spotting Temporally Precise, Fine-Grained Events in Video

James Hong, Haotian Zhang, Michaël Gharbi, Matthew Fisher, Kayvon Fatahalian

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments ECCV 2022; Website URL: https://jhong93.github.io/projects/spot.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.12374 2022-07-13 cs.CV 57%

MM-Pyramid: Multimodal Pyramid Attentional Network for Audio-Visual Event Localization and Video Parsing

Jiashuo Yu, Ying Cheng, Rui-Wei Zhao, Rui Feng, Yuejie Zhang

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.15268 2022-07-01 cs.CV 57%

Submission to Generic Event Boundary Detection Challenge@CVPR 2022: Local Context Modeling and Global Boundary Decoding Approach

Jiaqi Tang, Zhaoyang Liu, Jing Tan, Chen Qian, Wayne Wu, Limin Wang

专题命中 动作与事件理解 :video understanding(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:2112.04771

详情

展开后加载摘要…

URL PDF HTML 收藏