arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
2208.01954 2022-08-04 cs.CV

Dilated Context Integrated Network with Cross-Modal Consensus for Temporal Emotion Localization in Videos

Juncheng Li, Junlin Xie, Linchao Zhu, Long Qian, Siliang Tang, Wenqiao Zhang, Haochen Shi, Shengyu Zhang, Longhui Wei, Qi Tian, Yueting Zhuang

Comments Accepted by ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.01827 2022-08-04 cs.CV eess.IV

Fast Hierarchical Deep Unfolding Network for Image Compressed Sensing

Wenxue Cui, Shaohui Liu, Debin Zhao

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.01313 2022-08-03 cs.CV cs.MM

Unified Normalization for Accelerating and Stabilizing Transformers

Qiming Yang, Kai Zhang, Chaoxiang Lan, Zhi Yang, Zheyang Li, Wenming Tan, Jun Xiao, Shiliang Pu

Comments ACM MM'22

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.01195 2022-08-03 cs.CV

Making the Best of Both Worlds: A Domain-Oriented Transformer for Unsupervised Domain Adaptation

Wenxuan Ma, Jinming Zhang, Shuang Li, Chi Harold Liu, Yulin Wang, Wei Li

Comments Accepted at ACMMM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.00945 2022-08-02 cs.CV

DoF-NeRF: Depth-of-Field Meets Neural Radiance Fields

Zijin Wu, Xingyi Li, Juewen Peng, Hao Lu, Zhiguo Cao, Weicai Zhong

Comments Accepted by ACMMM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.00428 2022-08-02 cs.CV eess.IV

Robust Real-World Image Super-Resolution against Adversarial Attacks

Jiutao Yue, Haofeng Li, Pengxu Wei, Guanbin Li, Liang Lin

Comments ACM-MM 2021, Code: https://github.com/lhaof/Robust-SR-against-Adversarial-Attacks

Journal ref Proceedings of the 29th ACM International Conference on Multimedia (2021) 5148-5157

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.14571 2022-08-02 cs.CV

Prompting for Multi-Modal Tracking

Jinyu Yang, Zhe Li, Feng Zheng, Aleš Leonardis, Jingkuan Song

Comments Accepted at ACMMM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.05333 2022-08-02 cs.CV cs.LG

IDEA: Increasing Text Diversity via Online Multi-Label Recognition for Vision-Language Pre-training

Xinyu Huang, Youcai Zhang, Ying Cheng, Weiwei Tian, Ruiwei Zhao, Rui Feng, Yuejie Zhang, Yaqian Li, Yandong Guo, Xiaobo Zhang

Comments Accepted by the 30th ACM International Conference on Multimedia (ACM MM 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.14428 2022-08-01 cs.CV

Paired Cross-Modal Data Augmentation for Fine-Grained Image-to-Text Retrieval

Hao Wang, Guosheng Lin, Steven C. H. Hoi, Chunyan Miao

Comments Accepted at ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.13457 2022-07-28 cs.CV

Reducing the Vision and Language Bias for Temporal Sentence Grounding

Daizong Liu, Xiaoye Qu, Wei Hu

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.13450 2022-07-28 cs.CV

Skimming, Locating, then Perusing: A Human-Like Framework for Natural Language Video Localization

Daizong Liu, Wei Hu

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.12622 2022-07-27 cs.CV

Multi-Attention Network for Compressed Video Referring Object Segmentation

Weidong Chen, Dexiang Hong, Yuankai Qi, Zhenjun Han, Shuhui Wang, Laiyun Qing, Qingming Huang, Guorong Li

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11964 2022-07-26 eess.AS cs.LG cs.MM cs.SD

ConceptBeam: Concept Driven Target Speech Extraction

Yasunori Ohishi, Marc Delcroix, Tsubasa Ochiai, Shoko Araki, Daiki Takeuchi, Daisuke Niizumi, Akisato Kimura, Noboru Harada, Kunio Kashino

Comments Accepted to ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11844 2022-07-26 cs.CV eess.IV

Enhancing Image Rescaling using Dual Latent Variables in Invertible Neural Network

Min Zhang, Zhihong Pan, Xin Zhou, C. -C. Jay Kuo

Comments Accepted by ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.07820 2022-07-26 cs.CV

FCL-GAN: A Lightweight and Real-Time Baseline for Unsupervised Blind Image Deblurring

Suiyi Zhao, Zhao Zhang, Richang Hong, Mingliang Xu, Yi Yang, Meng Wang

Comments Please cite this work as: Suiyi Zhao, Zhao Zhang, Richang Hong, Mingliang Xu, Yi Yang and Meng Wang, "FCL-GAN: A Lightweight and Real-Time Baseline for Unsupervised Blind Image Deblurring," In: Proceedings of the 30th ACM International Conference on Multimedia (ACM MM), Lisbon, Portugal, June 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.08029 2022-07-26 cs.CV

Domain Generalization via Frequency-domain-based Feature Disentanglement and Interaction

Jingye Wang, Ruoyi Du, Dongliang Chang, Kongming Liang, Zhanyu Ma

Comments The paper is accepted by ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.00182 2022-07-26 cs.CV

You Only Hypothesize Once: Point Cloud Registration with Rotation-equivariant Descriptors

Haiping Wang, Yuan Liu, Zhen Dong, Wenping Wang

Comments Accepted by ACM Multimedia(MM) 2022, Project page: https://hpwang-whu.github.io/YOHO/

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.05526 2022-07-26 cs.CV cs.MM

Long-term Leap Attention, Short-term Periodic Shift for Video Classification

Hao Zhang, Lechao Cheng, Yanbin Hao, Chong-Wah Ngo

Comments Accepted by ACM Multimedia 2022, 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11118 2022-07-25 cs.CV

Rethinking the Reference-based Distinctive Image Captioning

Yangjun Mao, Long Chen, Zhihong Jiang, Dong Zhang, Zhimeng Zhang, Jian Shao, Jun Xiao

Comments ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.10833 2022-07-25 cs.CV

Few-shot Image Generation Using Discrete Content Representation

Yan Hong, Li Niu, Jianfu Zhang, Liqing Zhang

Comments This paper is accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.10817 2022-07-25 cs.SD cs.LG eess.AS

End-to-End and Self-Supervised Learning for ComParE 2022 Stuttering Sub-Challenge

Shakeel Ahmad Sheikh, Md Sahidullah, Fabrice Hirsch, Slim Ouni

Comments Accepted in ACM MM 2022 Conference : Grand Challenges, "\c{opyright} {Owner/Author | ACM} {2022}. This is the author's version of the work. It is posted here for your personal use. Not for redistribution

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.08601 2022-07-22 cs.CV

Geometry-Aware Reference Synthesis for Multi-View Image Super-Resolution

Ri Cheng, Yuqi Sun, Bo Yan, Weimin Tan, Chenxi Ma

Comments 16 pages, 10 figures, ACM MULTIMEDIA 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.12096 2022-07-21 cs.CV

Towards Unbiased Visual Emotion Recognition via Causal Intervention

Yuedong Chen, Xu Yang, Tat-Jen Cham, Jianfei Cai

Comments Accepted to ACM Multimedia 2022, code is available at https://github.com/donydchen/causal_emotion

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.09059 2022-07-20 cs.CV

Few-shot Open-set Recognition Using Background as Unknowns

Nan Song, Chi Zhang, Guosheng Lin

Comments Accpeted to ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.09046 2022-07-20 cs.CV

Dynamic Prototype Mask for Occluded Person Re-Identification

Lei Tan, Pingyang Dai, Rongrong Ji, Yongjian Wu

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07798 2022-07-20 cs.CV

CharFormer: A Glyph Fusion based Attentive Framework for High-precision Character Image Denoising

Daqian Shi, Xiaolei Diao, Lida Shi, Hao Tang, Yang Chi, Chuntao Li, Hao Xu

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07795 2022-07-20 cs.CV

RCRN: Real-world Character Image Restoration Network via Skeleton Extraction

Daqian Shi, Xiaolei Diao, Hao Tang, Xiaomin Li, Hao Xing, Hao Xu

Comments Accepted to ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.13317 2022-07-20 cs.CV cs.AI

MMRotate: A Rotated Object Detection Benchmark using PyTorch

Yue Zhou, Xue Yang, Gefan Zhang, Jiabao Wang, Yanyi Liu, Liping Hou, Xue Jiang, Xingzhao Liu, Junchi Yan, Chengqi Lyu, Wenwei Zhang, Kai Chen

Comments 5 pages, 2 tables, MMRotate is accepted by ACM MM 2022 (OS Track). Yue Zhou and Xue Yang provided equal contribution. The code is publicly released at https://github.com/open-mmlab/mmrotate

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.08387 2022-07-20 cs.CL cs.CV

LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking

Yupan Huang, Tengchao Lv, Lei Cui, Yutong Lu, Furu Wei

Comments ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.02378 2022-07-20 cs.CV

DiT: Self-supervised Pre-training for Document Image Transformer

Junlong Li, Yiheng Xu, Tengchao Lv, Lei Cui, Cha Zhang, Furu Wei

Comments ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏