arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
2401.09802 2024-07-19 eess.AS cs.CV cs.SD

Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation

Minsu Kim, Jeong Hun Yeo, Se Jin Park, Hyeongseop Rha, Yong Man Ro

Comments ACMMM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12622 2024-07-18 cs.CV

Rethinking the Architecture Design for Efficient Generic Event Boundary Detection

Ziwei Zheng, Zechuan Zhang, Yulin Wang, Shiji Song, Gao Huang, Le Yang

Comments ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12339 2024-07-18 cs.CV

Exploring Deeper! Segment Anything Model with Depth Perception for Camouflaged Object Detection

Zhenni Yu, Xiaoqin Zhang, Li Zhao, Yi Bin, Guobao Xiao

Comments ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12255 2024-07-18 cs.CV

Dual-Hybrid Attention Network for Specular Highlight Removal

Xiaojiao Guo, Xuhang Chen, Shenghong Luo, Shuqiang Wang, Chi-Man Pun

Comments Accepted by ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.10471 2024-07-18 cs.CR cs.AI cs.SD eess.AS

GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis

Weizhi Liu, Yue Li, Dongdong Lin, Hui Tian, Haizhou Li

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09956 2024-07-18 cs.SD cs.AI cs.CL eess.AS

Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal, Wei-Ning Hsu, Rada Mihalcea, Soujanya Poria

Comments Accepted at ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01097 2024-07-18 cs.CV cs.LG

Spatio-Temporal Branching for Motion Prediction using Motion Increments

Jiexin Wang, Yujie Zhou, Wenwen Qiang, Ying Ba, Bing Su, Ji-Rong Wen

Comments The incremental information of our paper includes the displacement information from the last frame of the historical sequence, derived from the motion information of the first frame in the future sequence and the motion information of the last frame of the historical sequence. This implicitly contains future information, inadvertently giving an unfair advantage in the human motion prediction task

Journal ref ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11405 2024-07-17 cs.CR cs.CV

Cover-separable Fixed Neural Network Steganography via Deep Generative Models

Guobiao Li, Sheng Li, Zhenxing Qian, Xinpeng Zhang

Comments Accepetd at ACMMM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00934 2024-07-17 cs.CV

LanEvil: Benchmarking the Robustness of Lane Detection to Environmental Illusions

Tianyuan Zhang, Lu Wang, Hainan Li, Yisong Xiao, Siyuan Liang, Aishan Liu, Xianglong Liu, Dacheng Tao

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.16301 2024-06-25 cs.CV cs.AI cs.MM

UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Yuting Mei, Linli Yao, Qin Jin

Comments Accepted by ACM International Conference on Multimedia Retrieval (ICMR'24)

Journal ref Proceedings of the 2024 International Conference on Multimedia Retrieval, May 2024, Pages 1034-1042

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13409 2024-06-21 cs.CV cs.MM

PetalView: Fine-grained Location and Orientation Extraction of Street-view Images via Cross-view Local Search with Supplementary Materials

Wenmiao Hu, Yichen Zhang, Yuxuan Liang, Xianjing Han, Yifang Yin, Hannes Kruppa, See-Kiong Ng, Roger Zimmermann

Comments This paper has been accepted by ACM Multimedia 2023. This version contains additional supplementary materials

Journal ref Proceedings of the 31st ACM International Conference on Multimedia (2023) 56-66

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12316 2024-06-19 cs.CV cs.AI cs.MM

Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning

Ruiqi Wu, Bingliang Jiao, Wenxuan Wang, Meng Liu, Peng Wang

Comments Accepyed by ACM International Conference on Multimedia Retrieval (ICMR'24)

Journal ref ICMR'24: Proceedings of the 2024 International Conference on Multimedia Retrieval (2024) 579 - 588

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.08624 2024-06-18 cs.CV

Domain Camera Adaptation and Collaborative Multiple Feature Clustering for Unsupervised Person Re-ID

Yuanpeng Tu

Comments ACMMM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10497 2024-05-20 cs.MM cs.AI cs.CV cs.SI

SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge

Bo Wu, Peiye Liu, Wen-Huang Cheng, Bei Liu, Zhaoyang Zeng, Jia Wang, Qiushi Huang, Jiebo Luo

Comments ACM Multimedia. arXiv admin note: text overlap with arXiv:1910.01795

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18060 2024-04-30 cs.CV cs.LG

Prompt Customization for Continual Learning

Yong Dai, Xiaopeng Hong, Yabin Wang, Zhiheng Ma, Dongmei Jiang, Yaowei Wang

Comments ACM MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09530 2024-04-22 cs.CV cs.AI

RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization

Avinash Anand, Raj Jaiswal, Mohit Gupta, Siddhesh S Bangar, Pijush Bhuyan, Naman Lal, Rajeev Singh, Ritika Jha, Rajiv Ratn Shah, Shin'ichi Satoh

Comments 8 pages, 6 figures, MMAsia 2023 Proceedings of the 5th ACM International Conference on Multimedia in Asia

Journal ref In Proceedings of the 5th ACM International Conference on Multimedia in Asia 2023. Association for Computing Machinery, NY, USA, Article 74, pp. 1-6

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13938 2024-04-18 cs.CV cs.LG

Improving Semi-Supervised Semantic Segmentation with Dual-Level Siamese Structure Network

Zhibo Tain, Xiaolin Zhang, Peng Zhang, Kun Zhan

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07855 2024-04-12 cs.CV

Resolve Domain Conflicts for Generalizable Remote Physiological Measurement

Weiyu Sun, Xinyu Zhang, Hao Lu, Ying Chen, Yun Ge, Xiaolin Huang, Jie Yuan, Yingcong Chen

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06033 2024-04-11 cs.CV

Little Strokes Fell Great Oaks: Boosting the Hierarchical Features for Multi-exposure Image Fusion

Pan Mu, Zhiying Du, Jinyuan Liu, Cong Bai

Journal ref Proceedings of the 31st ACM International Conference on Multimedia, October 2023, Pages 2985-2993

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.12886 2024-04-09 cs.CL cs.DL

Unveiling Global Narratives: A Multilingual Twitter Dataset of News Media on the Russo-Ukrainian Conflict

Sherzod Hakimov, Gullal S. Cheema

Comments ICMR 2024

Journal ref ICMR 2024 - ACM International Conference on Multimedia Retrieval 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18254 2024-04-02 cs.CV cs.AI

Sketch Input Method Editor: A Comprehensive Dataset and Methodology for Systematic Input Recognition

Guangming Zhu, Siyuan Wang, Qing Cheng, Kelong Wu, Hao Li, Liang Zhang

Comments The paper has been accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14326 2024-03-28 cs.MM

Think before You Leap: Content-Aware Low-Cost Edge-Assisted Video Semantic Segmentation

Mingxuan Yan, Yi Wang, Xuedou Xiao, Zhiqing Luo, Jianhua He, Wei Wang

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14652 2024-03-25 cs.CY cs.AI cs.CL cs.MM

MemeCraft: Contextual and Stance-Driven Multimodal Meme Generation

Han Wang, Roy Ka-Wei Lee

Comments 8 pages, 7 figures, ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.04834 2024-03-21 cs.CV

View while Moving: Efficient Video Recognition in Long-untrimmed Videos

Ye Tian, Mengyu Yang, Lanshan Zhang, Zhizhen Zhang, Yang Liu, Xiaohui Xie, Xirong Que, Wendong Wang

Comments Published on ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12289 2024-03-20 cs.NI eess.SP

BostonTwin: the Boston Digital Twin for Ray-Tracing in 6G Networks

Paolo Testolina, Michele Polese, Pedram Johari, Tommaso Melodia

Comments 7 pages, 5 figures, 3 tables. This paper has been accepted for presentation at ACM Multimedia Systems Conference 2024 (MMSys '24). Copyright ACM 2024. Please cite it as: P.Testolina, M. Polese, P. Johari, and T. Melodia, "BostonTwin: the Boston Digital Twin for Ray-Tracing in 6G Networks," in Proceedings of the ACM Multimedia Systems Conference 2024, ser. MMSys '24. Bari, Italy, Apr. 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03850 2024-03-19 cs.MM

Predictive Sampling for Efficient Pairwise Subjective Image Quality Assessment

Shima Mohammadi, João Ascenso

Comments 9 pages, 5 figures, accepted by ACM MM 2023

Journal ref Shima Mohammadi and João Ascenso. 2023. Predictive Sampling for Efficient Pairwise Subjective Image Quality Assessment. In Proceedings of the 31st ACM International Conference on Multimedia (MM '23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05856 2024-03-12 cs.CV

POV: Prompt-Oriented View-Agnostic Learning for Egocentric Hand-Object Interaction in the Multi-View World

Boshen Xu, Sipeng Zheng, Qin Jin

Comments Accepted by ACM MM 2023. Project page: https://xuboshen.github.io/

Journal ref Proceedings of the 31st ACM International Conference on Multimedia (2023). Association for Computing Machinery, New York, NY, USA, 2807-2816

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01087 2024-03-05 cs.MM cs.CV cs.SD eess.AS

Towards Accurate Lip-to-Speech Synthesis in-the-Wild

Sindhu Hegde, Rudrabha Mukhopadhyay, C. V. Jawahar, Vinay Namboodiri

Comments 8 pages of content, 1 page of references and 4 figures

Journal ref In Proceedings of the 31st ACM International Conference on Multimedia, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18171 2024-02-29 cs.CV

Digging Into Normal Incorporated Stereo Matching

Zihua Liu, Songyan Zhang, Zhicheng Wang, Masatoshi Okutomi

Journal ref Proceedings of the 30th ACM International Conference on Multimedia (ACMMM2022), pp.6050-6060, October 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12252 2024-02-29 cs.CV cs.LG cs.MM

Long-Range Feature Propagating for Natural Image Matting

Qinglin Liu, Haozhe Xie, Shengping Zhang, Bineng Zhong, Rongrong Ji

Journal ref ACM International Conference on Multimedia (ACM MM) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏