arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
1903.03299 2021-10-26 cs.CV

You Only Recognize Once: Towards Fast Video Text Spotting

Zhanzhan Cheng, Jing Lu, Yi Niu, Shiliang Pu, Fei Wu, Shuigeng Zhou

Comments Accepted by ACM Multimedia 2019. Code is available at https://davar-lab.github.io/publication.html or https://github.com/hikopensource/DAVAR-Lab-OCR

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.08061 2021-10-25 cs.CV

Invertible Frowns: Video-to-Video Facial Emotion Translation

Ian Magnusson, Aruna Sankaranarayanan, Andrew Lippman

Comments 9 pages, 2 figures, 4 tables, accepted at ADGD @ ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.13385 2021-10-22 eess.IV cs.MM

A Complete End-To-End Open Source Toolchain for the Versatile Video Coding (VVC) Standard

Adam Wieckowski, Christian Lehmann, Benjamin Bross, Detlev Marpe, Thibaud Biatek, Mickael Raulet, Jean Le Feuvre

Comments 4 pages, 2 figures, accepted to ACM International Conference on Multimedia (MM'21)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09819 2021-10-20 cs.CV

LSTC: Boosting Atomic Action Detection with Long-Short-Term Context

Yuxi Li, Boshen Zhang, Jian Li, Yabiao Wang, Weiyao Lin, Chengjie Wang, Jilin Li, Feiyue Huang

Comments ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09756 2021-10-20 cs.CV cs.CL cs.MM

A Picture is Worth a Thousand Words: A Unified System for Diverse Captions and Rich Images Generation

Yupan Huang, Bei Liu, Jianlong Fu, Yutong Lu

Comments ACM MM 2021 (Video and Demo Track). Code: https://github.com/researchmm/generate-it

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09753 2021-10-20 cs.CV cs.CL cs.MM

Unifying Multimodal Transformer for Bi-directional Image and Text Generation

Yupan Huang, Hongwei Xue, Bei Liu, Yutong Lu

Comments ACM MM 2021 (Industrial Track). Code: https://github.com/researchmm/generate-it

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09699 2021-10-20 cs.CV eess.IV

Image Quality Assessment in the Modern Age

Kede Ma, Yuming Fang

Comments ACM Multimedia 2021 Tutorial

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.09109 2021-10-19 cs.CV cs.MM eess.IV

Patch-Based Deep Autoencoder for Point Cloud Geometry Compression

Kang You, Pan Gao

Comments Accepted to ACM Multimedia Asia (MMAsia '21)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.08818 2021-10-19 cs.CV cs.GR cs.MM

MeronymNet: A Hierarchical Approach for Unified and Controllable Multi-Category Object Generation

Rishabh Baghel, Abhishek Trivedi, Tejas Ravichandran, Ravi Kiran Sarvadevabhatla

Comments Accepted at ACM Multimedia (ACMMM) 2021 [ORAL] . Website : https://meronymnet.github.io/. arXiv admin note: text overlap with arXiv:2006.00190

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.08521 2021-10-19 eess.IV cs.CV

Locally Adaptive Structure and Texture Similarity for Image Quality Assessment

Keyan Ding, Yi Liu, Xueyi Zou, Shiqi Wang, Kede Ma

Journal ref Proceedings of the 29th ACM International Conference on Multimedia, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.00522 2021-10-18 cs.CV cs.LG cs.MM

Conditional Extreme Value Theory for Open Set Video Domain Adaptation

Zhuoxiao Chen, Yadan Luo, Mahsa Baktashmotlagh

Comments Camera-ready. Accepted by ACM International Conference on Multimedia in Asia 2021 (MMAsia 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.06827 2021-10-14 cs.MM cs.CV cs.LG

NoisyActions2M: A Multimedia Dataset for Video Understanding from Noisy Labels

Mohit Sharma, Raj Patra, Harshal Desai, Shruti Vyas, Yogesh Rawat, Rajiv Ratn Shah

Comments Accepted at ACM Multimedia Asia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.06805 2021-10-14 cs.MM

Assisting News Media Editors with Cohesive Visual Storylines

Gonçalo Marcelino, David Semedo, André Mourão, Saverio Blasi, Marta Mrak, João Magalhães

Comments Accepted at ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.06484 2021-10-14 cs.CV

Domain Adaptive Semantic Segmentation without Source Data

Fuming You, Jingjing Li, Lei Zhu, Ke Lu, Zhi Chen, Zi Huang

Comments Accepted by ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.01860 2021-10-12 cs.CV

Spatiotemporal Inconsistency Learning for DeepFake Video Detection

Zhihao Gu, Yang Chen, Taiping Yao, Shouhong Ding, Jilin Li, Feiyue Huang, Lizhuang Ma

Comments To appear in ACM MM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.04492 2021-10-12 cs.CV

Weight Evolution: Improving Deep Neural Networks Training through Evolving Inferior Weight Values

Zhenquan Lin, Kailing Guo, Xiaofen Xing, Xiangmin Xu

Comments This paper is accepted by ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.04435 2021-10-12 cs.CV cs.CL

Two-stage Visual Cues Enhancement Network for Referring Image Segmentation

Yang Jiao, Zequn Jie, Weixin Luo, Jingjing Chen, Yu-Gang Jiang, Xiaolin Wei, Lin Ma

Comments Accepted by ACM MM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.02048 2021-10-06 cs.LG

Graph Coloring: Comparing Cluster Graphs to Factor Graphs

Simon Streicher, Johan du Preez

Journal ref SAWACMMM '17: Proceedings of the ACM Multimedia 2017 Workshop on South African Academic Participation; October 2017; Pages 35-42

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12651 2021-09-28 cs.IR cs.AI cs.MM

Why Do We Click: Visual Impression-aware News Recommendation

Jiahao Xun, Shengyu Zhang, Zhou Zhao, Jieming Zhu, Qi Zhang, Jingjie Li, Xiuqiang He, Xiaofei He, Tat-Seng Chua, Fei Wu

Comments Accepted by ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12307 2021-09-28 cs.CV cs.AI cs.MM

Multi-Modal Multi-Instance Learning for Retinal Disease Recognition

Xirong Li, Yang Zhou, Jie Wang, Hailan Lin, Jianchun Zhao, Dayong Ding, Weihong Yu, Youxin Chen

Comments Accepted by ACM Multimedia 2021 (Main Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09869 2021-09-28 cs.CR

FakeTagger: Robust Safeguards against DeepFake Dissemination via Provenance Tracking

Run Wang, Felix Juefei-Xu, Meng Luo, Yang Liu, Lina Wang

Comments Accepted to ACM Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.11778 2021-09-27 cs.CV cs.CL

Dense Contrastive Visual-Linguistic Pretraining

Lei Shi, Kai Shuang, Shijie Geng, Peng Gao, Zuohui Fu, Gerard de Melo, Yunpeng Chen, Sen Su

Comments Accepted by ACM Multimedia 2021. arXiv admin note: text overlap with arXiv:2007.13135

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.11243 2021-09-24 cs.CV

Pairwise Emotional Relationship Recognition in Drama Videos: Dataset and Benchmark

Xun Gao, Yin Zhao, Jie Zhang, Longjun Cai

Journal ref ACM MM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05778 2021-09-21 cs.CL

Text is NOT Enough: Integrating Visual Impressions into Open-domain Dialogue Generation

Lei Shen, Haolan Zhan, Xin Shen, Yonghao Song, Xiaofang Zhao

Comments Accepted by ACM MultiMedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.07951 2021-09-17 cs.CV

Overview of Tencent Multi-modal Ads Video Understanding Challenge

Zhenzhi Wang, Liyu Wu, Zhimin Li, Jiangfeng Xiong, Qinglin Lu

Comments 8-page extended version of our challenge paper in ACM MM 2021. It presents the overview of grand challenge "Multi-modal Ads Video Understanding" in ACM MM 2021. Our grand challenge is also the Tencent Advertising Algorithm Competition (TAAC) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.07756 2021-09-17 cs.CV

Dense Semantic Contrast for Self-Supervised Visual Representation Learning

Xiaoni Li, Yu Zhou, Yifei Zhang, Aoting Zhang, Wei Wang, Ning Jiang, Haiying Wu, Weiping Wang

Comments ACM MM 2021 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05534 2021-09-14 cs.CV

DSSL: Deep Surroundings-person Separation Learning for Text-based Person Retrieval

Aichun Zhu, Zijie Wang, Yifeng Li, Xili Wan, Jing Jin, Tian Wang, Fangqiang Hu, Gang Hua

Comments Accepted by ACM MM'21

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.02216 2021-09-14 cs.CV

Learning Fine-Grained Motion Embedding for Landscape Animation

Hongwei Xue, Bei Liu, Huan Yang, Jianlong Fu, Houqiang Li, Jiebo Luo

Comments Accepted by ACM Multimedia 2021 as an oral paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.01361 2021-09-14 cs.CV

Cycle-Consistent Inverse GAN for Text-to-Image Synthesis

Hao Wang, Guosheng Lin, Steven C. H. Hoi, Chunyan Miao

Comments Accepted at ACM MM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.04153 2021-09-10 cs.CV cs.LG

Single Image 3D Object Estimation with Primitive Graph Networks

Qian He, Desen Zhou, Bo Wan, Xuming He

Comments Accepted by ACM MM'21

详情

展开后加载摘要…

URL PDF HTML 收藏