arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
2307.10636 2023-08-03 cs.CV

Learning and Evaluating Human Preferences for Conversational Head Generation

Mohan Zhou, Yalong Bai, Wei Zhang, Ting Yao, Tiejun Zhao, Tao Mei

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.00655 2023-08-02 cs.CV

Toward Zero-shot Character Recognition: A Gold Standard Dataset with Radical-level Annotations

Xiaolei Diao, Daqian Shi, Jian Li, Lida Shi, Mingzhe Yue, Ruihua Qi, Chuntao Li, Hao Xu

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.00508 2023-08-02 cs.CV

Relational Contrastive Learning for Scene Text Recognition

Jinglei Zhang, Tiancheng Lin, Yi Xu, Kai Chen, Rui Zhang

Comments Accepted by ACMMM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06321 2023-08-02 cs.CV cs.MM eess.IV

SepMark: Deep Separable Watermarking for Unified Source Tracing and Deepfake Detection

Xiaoshuai Wu, Xin Liao, Bo Ou

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.13137 2023-08-02 cs.LG cs.DC

FedGH: Heterogeneous Federated Learning with Generalized Global Header

Liping Yi, Gang Wang, Xiaoguang Liu, Zhuan Shi, Han Yu

Comments 11 pages, 5 figures,accepted by Proceedings of the 31st ACM International Conference on Multimedia (MM 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.00313 2023-08-02 cs.CV

Zero-Shot Learning by Harnessing Adversarial Samples

Zhi Chen, Pengfei Zhang, Jingjing Li, Sen Wang, Zi Huang

Comments Accepted to ACM International Conference on Multimedia (MM) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.00255 2023-08-02 cs.CV cs.AI

LGViT: Dynamic Early Exiting for Accelerating Vision Transformer

Guanyu Xu, Jiawei Hao, Li Shen, Han Hu, Yong Luo, Hui Lin, Jialie Shen

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16620 2023-08-02 cs.SD cs.CV eess.AS

Audio-Visual Segmentation by Exploring Cross-Modal Mutual Semantics

Chen Liu, Peike Li, Xingqun Qi, Hu Zhang, Lincheng Li, Dadong Wang, Xin Yu

Comments This paper has been received by ACM MM 23

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.04224 2023-08-02 cs.CV

Visual Causal Scene Refinement for Video Question Answering

Yushen Wei, Yang Liu, Hong Yan, Guanbin Li, Liang Lin

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.06607 2023-08-02 cs.CL cs.MM

Few-shot Multimodal Sentiment Analysis based on Multimodal Probabilistic Fusion Prompts

Xiaocui Yang, Shi Feng, Daling Wang, Pengfei Hong, Soujanya Poria

Comments 9 pages, 2 figures, 7 tables. It has been accepted ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.04660 2023-08-02 cs.LG cs.AI

Dynamic Collective Intelligence Learning: Finding Efficient Sparse Model via Refined Gradients for Pruned Weights

Jangho Kim, Jayeon Yoo, Yeji Song, KiYoon Yoo, Nojun Kwak

Comments Accepted to ACM MM 2023, code is in https://github.com/Jangho-Kim/DCIL-pytorch

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16813 2023-08-01 cs.CV

Capturing Co-existing Distortions in User-Generated Content for No-reference Video Quality Assessment

Kun Yuan, Zishang Kong, Chuanchuan Zheng, Ming Sun, Xing Wen

Comments 10 pages, 7 figures, to appear in ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16765 2023-08-01 cs.CV

Lightweight Super-Resolution Head for Human Pose Estimation

Haonan Wang, Jie Liu, Jie Tang, Gangshan Wu

Comments ACM MM 2023 accepted

Journal ref ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16226 2023-08-01 cs.CV cs.MM

ScribbleVC: Scribble-supervised Medical Image Segmentation with Vision-Class Embedding

Zihan Li, Yuan Zheng, Xiangde Luo, Dandan Shan, Qingqi Hong

Comments Accepted by ACM MM 2023, project page: https://github.com/HUANGLIZI/ScribbleVC

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.14454 2023-08-01 cs.AI cs.CL

MEAformer: Multi-modal Entity Alignment Transformer for Meta Modality Hybrid

Zhuo Chen, Jiaoyan Chen, Wen Zhang, Lingbing Guo, Yin Fang, Yufeng Huang, Yichi Zhang, Yuxia Geng, Jeff Z. Pan, Wenting Song, Huajun Chen

Comments ACM Multimedia 2023 Accpeted, Repo: https://github.com/zjukg/MEAformer

Journal ref ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15574 2023-07-31 cs.DC cs.MM

FleXR: A System Enabling Flexibly Distributed Extended Reality

Jin Heo, Ketan Bhardwaj, Ada Gavrilovska

Comments 11 pages, 11 figures, conference paper

Journal ref In Proceedings of the 14th Conference on ACM Multimedia Systems (pp. 1-13) June, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15097 2023-07-31 cs.CL cs.LG cs.MM eess.AS

Cascaded Cross-Modal Transformer for Request and Complaint Detection

Nicolae-Catalin Ristea, Radu Tudor Ionescu

Comments Accepted at ACMMM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14613 2023-07-28 cs.LG cs.AI

Self-Contrastive Graph Diffusion Network

Yixian Ma, Kun Zhan

Comments ACM Multimedia 2013 Accpeted

Journal ref ACM MM 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.09160 2023-07-28 cs.CV cs.AI

SUG: Single-dataset Unified Generalization for 3D Point Cloud Classification

Siyuan Huang, Bo Zhang, Botian Shi, Peng Gao, Yikang Li, Hongsheng Li

Comments Accepted by ACM MM-2023, and our code is available at https://github.com/SiyuanHuang95/SUG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13948 2023-07-27 cs.CV cs.SD eess.AS

Rethinking Voice-Face Correlation: A Geometry View

Xiang Li, Yandong Wen, Muqiao Yang, Jinglu Wang, Rita Singh, Bhiksha Raj

Comments ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13944 2023-07-27 cs.LG cs.AI

Entropy Neural Estimation for Graph Contrastive Learning

Yixuan Ma, Xiaolin Zhang, Peng Zhang, Kun Zhan

Comments ACM MM 2023 accepted

Journal ref ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13908 2023-07-27 cs.CV

Points-to-3D: Bridging the Gap between Sparse Points and Shape-Controllable Text-to-3D Generation

Chaohui Yu, Qiang Zhou, Jingliang Li, Zhe Zhang, Zhibin Wang, Fan Wang

Comments Accepted by ACMMM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.05152 2023-07-27 cs.SD cs.MM eess.AS

Who is Speaking Actually? Robust and Versatile Speaker Traceability for Voice Conversion

Yanzhen Ren, Hongcheng Zhu, Liming Zhai, Zongkun Sun, Rubing Shen, Lina Wang

Comments has been accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.11100 2023-07-24 cs.CV cs.MM

CSSL-RHA: Contrastive Self-Supervised Learning for Robust Handwriting Authentication

Jingyao Wang, Luntian Mou, Changwen Zheng, Wen Gao

Comments 10 pages, 4 figures, 3 tables, submitted to ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10184 2023-07-21 cs.CR cs.AI cs.LG

A Dual Stealthy Backdoor: From Both Spatial and Frequency Perspectives

Yudong Gao, Honglong Chen, Peng Sun, Junjian Li, Anqing Zhang, Zhibo Wang

Comments 10 pages, 7 figures. Submit to ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.09821 2023-07-20 cs.CV cs.MM

Hierarchical Semantic Perceptual Listener Head Video Generation: A High-performance Pipeline

Zhigang Chang, Weitai Hu, Qing Yang, Shibao Zheng

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.09416 2023-07-20 cs.CV cs.CL

Let's ViCE! Mimicking Human Cognitive Behavior in Image Generation Evaluation

Federico Betti, Jacopo Staiano, Lorenzo Baraldi, Lorenzo Baraldi, Rita Cucchiara, Nicu Sebe

Comments Accepted as oral at ACM MultiMedia 2023 (Brave New Ideas track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.03398 2023-07-14 cs.CV

Beyond Geo-localization: Fine-grained Orientation of Street-view Images by Cross-view Matching with Satellite Imagery with Supplementary Materials

Wenmiao Hu, Yichen Zhang, Yuxuan Liang, Yifang Yin, Andrei Georgescu, An Tran, Hannes Kruppa, See-Kiong Ng, Roger Zimmermann

Comments This paper has been accepted by ACM Multimedia 2022. This version contains additional supplementary materials

Journal ref Proceedings of the 30th ACM International Conference on Multimedia (2022) 6155-6164

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11515 2023-07-13 cs.CV

Marior: Margin Removal and Iterative Content Rectification for Document Dewarping in the Wild

Jiaxin Zhang, Canjie Luo, Lianwen Jin, Fengjun Guo, Kai Ding

Comments This paper has been accepted by ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.06284 2023-07-11 cs.SD cs.LG cs.MM eess.AS

Everybody Compose: Deep Beats To Music

Conghao Shen, Violet Z. Yao, Yixin Liu

Comments Accepted MMSys '23

Journal ref Proceedings of the 14th Conference on ACM Multimedia Systems (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏