arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
2310.08117 2023-10-13 cs.CV cs.AI

DUSA: Decoupled Unsupervised Sim2Real Adaptation for Vehicle-to-Everything Collaborative Perception

Xianghao Kong, Wentao Jiang, Jinrang Jia, Yifeng Shi, Runsheng Xu, Si Liu

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08032 2023-10-13 cs.AI

Incorporating Domain Knowledge Graph into Multimodal Movie Genre Classification with Self-Supervised Attention and Contrastive Learning

Jiaqi Li, Guilin Qi, Chuanyi Zhang, Yongrui Chen, Yiming Tan, Chenlong Xia, Ye Tian

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07287 2023-10-12 cs.MM

Interactive Interior Design Recommendation via Coarse-to-fine Multimodal Reinforcement Learning

He Zhang, Ying Sun, Weiyu Guo, Yafei Liu, Haonan Lu, Xiaodong Lin, Hui Xiong

Comments Accepted by ACM International Conference on Multimedia'23. 9 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07222 2023-10-12 cs.CV

Uni-paint: A Unified Framework for Multimodal Image Inpainting with Pretrained Diffusion Model

Shiyuan Yang, Xiaodong Chen, Jing Liao

Comments Accepted by ACMMM'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.06138 2023-10-11 cs.CV cs.AI cs.LG cs.MM cs.RO

Layout Sequence Prediction From Noisy Mobile Modality

Haichao Zhang, Yi Xu, Hongsheng Lu, Takayuki Shimizu, Yun Fu

Comments In Proceedings of the 31st ACM International Conference on Multimedia 2023 (MM 23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04999 2023-10-11 cs.CV

Symmetrical Linguistic Feature Distillation with CLIP for Scene Text Recognition

Zixiao Wang, Hongtao Xie, Yuxin Wang, Jianjun Xu, Boqiang Zhang, Yongdong Zhang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.03504 2023-10-11 cs.CV

Stroke-based Neural Painting and Stylization with Dynamically Predicted Painting Region

Teng Hu, Ran Yi, Haokun Zhu, Liang Liu, Jinlong Peng, Yabiao Wang, Chengjie Wang, Lizhuang Ma

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05847 2023-10-10 cs.LG cs.AI cs.CR cs.IR

Making Users Indistinguishable: Attribute-wise Unlearning in Recommender Systems

Yuyuan Li, Chaochao Chen, Xiaolin Zheng, Yizhao Zhang, Zhongxuan Han, Dan Meng, Jun Wang

Journal ref Proceedings of the 31st ACM International Conference on Multimedia (MM '23), October 29--November 3, 2023, Ottawa, ON, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05589 2023-10-10 cs.CL cs.MM

DRIN: Dynamic Relation Interactive Network for Multimodal Entity Linking

Shangyu Xing, Fei Zhao, Zhen Wu, Chunhui Li, Jianbing Zhang, Xinyu Dai

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04689 2023-10-10 cs.CV

SeeDS: Semantic Separable Diffusion Synthesizer for Zero-shot Food Detection

Pengfei Zhou, Weiqing Min, Yang Zhang, Jiajun Song, Ying Jin, Shuqiang Jiang

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04679 2023-10-10 eess.IV cs.CV

High Visual-Fidelity Learned Video Compression

Meng Li, Yibo Shi, Jing Wang, Yunqi Huang

Comments ACMMM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04456 2023-10-10 cs.CL cs.SD eess.AS

Multimodal Prompt Transformer with Hybrid Contrastive Learning for Emotion Recognition in Conversation

Shihao Zou, Xianying Huang, Xudong Shen

Comments Accepted to ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13551 2023-10-10 cs.HC cs.AI cs.GR

Dance with You: The Diversity Controllable Dancer Generation via Diffusion Models

Siyue Yao, Mingjie Sun, Bingliang Li, Fengyu Yang, Junle Wang, Ruimao Zhang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03272 2023-10-10 cs.CV

Feature-Suppressed Contrast for Self-Supervised Food Pre-training

Xinda Liu, Yaohui Zhu, Linhu Liu, Jiang Tian, Lili Wang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03670 2023-10-06 cs.CV

Regress Before Construct: Regress Autoencoder for Point Cloud Self-supervised Learning

Yang Liu, Chen Chen, Can Wang, Xulin King, Mengyuan Liu

Journal ref In Proceedings of the 31st ACM International Conference on Multimedia (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.02361 2023-10-05 cs.RO

Event-Enhanced Multi-Modal Spiking Neural Network for Dynamic Obstacle Avoidance

Yang Wang, Bo Dong, Yuji Zhang, Yunduo Zhou, Haiyang Mei, Ziqi Wei, Xin Yang

Comments In Proceedings of the 31st ACM International Conference on Multimedia (ACM MM 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.17104 2023-10-04 cs.CV

Prototype-guided Cross-modal Completion and Alignment for Incomplete Text-based Person Re-identification

Tiantian Gong, Guodong Du, Junsheng Wang, Yongkang Ding, Liyan Zhang

Comments Sorry, some collaborators do not agree to publish it on Arxiv, so please withdraw this paper

Journal ref ACM International Conference on Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.06034 2023-10-03 cs.LG

Normality Learning-based Graph Anomaly Detection via Multi-Scale Contrastive Learning

Jingcan Duan, Pei Zhang, Siwei Wang, Jingtao Hu, Hu Jin, Jiaxin Zhang, Haifang Zhou, Xinwang Liu

Comments 10 pages, 7 figures, accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.06466 2023-10-03 eess.IV

U2Net: A General Framework with Spatial-Spectral-Integrated Double U-Net for Image Fusion

Siran Peng, Chenhao Guo, Xiao Wu, Liang-Jian Deng

Comments Accepted by the 31st ACM International Conference on Multimedia (ACM MM '23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.09923 2023-10-02 cs.CV

Zero-shot point cloud segmentation by transferring geometric primitives

Runnan Chen, Xinge Zhu, Nenglun Chen, Wei Li, Yuexin Ma, Ruigang Yang, Wenping Wang

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16140 2023-09-29 cs.MM cs.CV

CLIP-Hand3D: Exploiting 3D Hand Pose Estimation via Context-Aware Prompting

Shaoxiang Guo, Qing Cai, Lin Qi, Junyu Dong

Comments Accepted In Proceedings of the 31st ACM International Conference on Multimedia (MM' 23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14359 2023-09-29 cs.AI cs.CL

Effect of Attention and Self-Supervised Speech Embeddings on Non-Semantic Speech Tasks

Payal Mohapatra, Akash Pandey, Yueyuan Sui, Qi Zhu

Comments Accepted to appear at ACM Multimedia 2023 Multimedia Grand Challenges Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11845 2023-09-27 cs.SD cs.LG cs.MM eess.AS

TMac: Temporal Multi-Modal Graph Learning for Acoustic Event Classification

Meng Liu, Ke Liang, Dayu Hu, Hao Yu, Yue Liu, Lingyuan Meng, Wenxuan Tu, Sihang Zhou, Xinwang Liu

Comments This work has been accepted by ACM MM 2023 for publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14673 2023-09-27 cs.LG cs.AI cs.IR cs.SI

ALEX: Towards Effective Graph Transfer Learning with Noisy Labels

Jingyang Yuan, Xiao Luo, Yifang Qin, Zhengyang Mao, Wei Ju, Ming Zhang

Comments Accepted by the ACM International Conference on Multimedia (MM) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14282 2023-09-26 cs.CV

Calibration-based Dual Prototypical Contrastive Learning Approach for Domain Generalization Semantic Segmentation

Muxin Liao, Shishun Tian, Yuhang Zhang, Guoguang Hua, Wenbin Zou, Xia Li

Comments Accepted by ACM MM'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16437 2023-09-26 cs.CV cs.AI cs.MM

KeyPosS: Plug-and-Play Facial Landmark Detection through GPS-Inspired True-Range Multilateration

Xu Bao, Zhi-Qi Cheng, Jun-Yan He, Chenyang Li, Wangmeng Xiang, Jingdong Sun, Hanbing Liu, Wei Liu, Bin Luo, Yifeng Geng, Xuansong Xie

Comments Accepted to ACM Multimedia 2023; 10 pages, 7 figures, 6 tables; the code is at https://github.com/zhiqic/KeyPosS

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.00468 2023-09-25 cs.CV

An Improved Encoder-Decoder Framework for Food Energy Estimation

Jack Ma, Jiangpeng He, Fengqing Zhu

Comments Accepted for Madima'23 in ACM Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.16247 2023-09-25 cs.CV

Ada3Diff: Defending against 3D Adversarial Point Clouds via Adaptive Diffusion

Kui Zhang, Hang Zhou, Jie Zhang, Qidong Huang, Weiming Zhang, Nenghai Yu

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.12134 2023-09-22 cs.SD cs.IR cs.LG eess.AS

Self-Supervised Contrastive Learning for Robust Audio-Sheet Music Retrieval Systems

Luis Carvalho, Tobias Washüttl, Gerhard Widmer

Journal ref Proceedings of the 14th ACM Multimedia Systems Conference (MMSys '23), June 7-10, 2023, Vancouver, BC, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.05070 2023-09-22 cs.CV

mmBody Benchmark: 3D Body Reconstruction Dataset and Analysis for Millimeter Wave Radar

Anjun Chen, Xiangyu Wang, Shaohao Zhu, Yanxu Li, Jiming Chen, Qi Ye

Comments Accepted to ACM Multimedia 2022, Project Page: https://chen3110.github.io/mmbody/index.html

详情

展开后加载摘要…

URL PDF HTML 收藏