arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
2207.05418 2022-07-13 cs.CV cs.LG

A Baseline for Detecting Out-of-Distribution Examples in Image Captioning

Gabi Shalev, Gal-Lev Shalev, Joseph Keshet

Comments Accepted to ACM Multimedia (MM) 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.05340 2022-07-13 cs.CV

Dual Contrastive Learning for Spatio-temporal Representation

Shuangrui Ding, Rui Qian, Hongkai Xiong

Comments ACM MM 2022 camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.05334 2022-07-13 cs.CV

Cycle Self-Training for Semi-Supervised Object Detection with Distribution Consistency Reweighting

Hao Liu, Bin Chen, Bo Wang, Chunpeng Wu, Feng Dai, Peng Wu

Comments ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02812 2022-07-13 cs.CV

Towards Counterfactual Image Manipulation via CLIP

Yingchen Yu, Fangneng Zhan, Rongliang Wu, Jiahui Zhang, Shijian Lu, Miaomiao Cui, Xuansong Xie, Xian-Sheng Hua, Chunyan Miao

Comments This paper has been accepted to ACM MM 2022, code may be found here: https://github.com/yingchen001/CF-CLIP

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.01030 2022-07-13 cs.CV

Boosting Single-Frame 3D Object Detection by Simulating Multi-Frame Point Clouds

Wu Zheng, Li Jiang, Fanbin Lu, Yangyang Ye, Chi-Wing Fu

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.09248 2022-07-13 cs.SD cs.CV cs.GR cs.LG cs.MM eess.AS

MESH2IR: Neural Acoustic Impulse Response Generator for Complex 3D Scenes

Anton Ratnarajah, Zhenyu Tang, Rohith Chandrashekar Aralikatti, Dinesh Manocha

Comments Accepted to ACM Multimedia 2022. More results and source code is available at https://anton-jeran.github.io/M2IR/

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.05946 2022-07-13 cs.SI

Understanding Political Polarization via Jointly Modeling Users, Connections and Multimodal Contents on Heterogeneous Graphs

Hanjia Lyu, Jiebo Luo

Comments Accepted for publication in Proceedings of the 30th ACM International Conference on Multimedia, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.12374 2022-07-13 cs.CV

MM-Pyramid: Multimodal Pyramid Attentional Network for Audio-Visual Event Localization and Video Parsing

Jiashuo Yu, Ying Cheng, Rui-Wei Zhao, Rui Feng, Yuejie Zhang

Comments ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.09337 2022-07-13 cs.CV

PC$^2$-PU: Patch Correlation and Point Correlation for Effective Point Cloud Upsampling

Chen Long, Wenxiao Zhang, Ruihui Li, Hao Wang, Zhen Dong, Bisheng Yang

Comments Accepted to ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.06441 2022-07-13 cs.SD cs.LG eess.AS

Structure-Enhanced Pop Music Generation via Harmony-Aware Learning

Xueyao Zhang, Jinchao Zhang, Yao Qiu, Li Wang, Jie Zhou

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.06187 2022-07-12 cs.CV

Calibrating Class Weights with Multi-Modal Information for Partial Video Domain Adaptation

Xiyu Wang, Yuecong Xu, Kezhi Mao, Jianfei Yang

Comments Accepted by ACM Multimedia (ACMMM) 2022, update to camera-ready version. 8 pages of text, 5 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.04214 2022-07-12 cs.IR

Adaptive Structural Similarity Preserving for Unsupervised Cross Modal Hashing

Liang Li, Baihua Zheng, Weiwei Sun

Comments Accepted to ACM Multimedia 2022 as Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.03723 2022-07-11 cs.CV cs.MM eess.IV

Exploring the Effectiveness of Video Perceptual Representation in Blind Video Quality Assessment

Liang Liao, Kangmin Xu, Haoning Wu, Chaofeng Chen, Wenxiu Sun, Qiong Yan, Weisi Lin

Comments Will appear on ACM MM 2022

Journal ref 2022 ACM International Conference on Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.01798 2022-07-11 cs.CV cs.LG

GSMFlow: Generation Shifts Mitigating Flow for Generalized Zero-Shot Learning

Zhi Chen, Yadan Luo, Sen Wang, Jingjing Li, Zi Huang

Comments IEEE Transactions on Multimedia 2022. Journal Extension from "Mitigating Generation Shifts for Generalized Zero-Shot Learning", ACM MM 2021. arXiv admin note: substantial text overlap with arXiv:2107.03163

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.03056 2022-07-08 cs.MM

Privacy-preserving Reflection Rendering for Augmented Reality

Yiqin Zhao, Sheng Wei, Tian Guo

Comments Accepted to ACM Multimedia 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.01241 2022-07-05 cs.CV

OS-MSL: One Stage Multimodal Sequential Link Framework for Scene Segmentation and Classification

Ye Liu, Lingfeng Qiao, Di Yin, Zhuoxuan Jiang, Xinghua Jiang, Deqiang Jiang, Bo Ren

Comments Accepted by ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.00965 2022-07-05 cs.CV eess.IV

Cycle-Interactive Generative Adversarial Network for Robust Unsupervised Low-Light Enhancement

Zhangkai Ni, Wenhan Yang, Hanli Wang, Shiqi Wang, Lin Ma, Sam Kwong

Comments 9 pages, 7 figures, accepted to ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.00887 2022-07-05 cs.CV

Towards Robust Video Object Segmentation with Adaptive Object Calibration

Xiaohao Xu, Jinglu Wang, Xiang Ming, Yan Lu

Comments 19 pages, 17 figures, ACM Multimedia 2022 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.00784 2022-07-05 cs.CV

Learning Cross-Image Object Semantic Relation in Transformer for Few-Shot Fine-Grained Image Classification

Bo Zhang, Jiakang Yuan, Baopu Li, Tao Chen, Jiayuan Fan, Botian Shi

Comments Accepted by ACM MM-2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.11582 2022-07-01 cs.CV

Graph-DETR3D: Rethinking Overlapping Regions for Multi-View 3D Object Detection

Zehui Chen, Zhenyu Li, Shiquan Zhang, Liangji Fang, Qinhong Jiang, Feng Zhao

Comments Accepted to ACM MM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.08249 2022-06-10 cs.CV

Learnable Optimal Sequential Grouping for Video Scene Detection

Daniel Rotman, Yevgeny Yaroker, Elad Amrani, Udi Barzelay, Rami Ben-Ari

Journal ref In Proceedings of the 28th ACM International Conference on Multimedia, pp. 1958-1966. 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.06799 2022-05-16 cs.SD cs.LG eess.AS

The ACM Multimedia 2022 Computational Paralinguistics Challenge: Vocalisations, Stuttering, Activity, & Mosquitoes

Björn W. Schuller, Anton Batliner, Shahin Amiriparian, Christian Bergler, Maurice Gerczuk, Natalie Holz, Pauline Larrouy-Maestri, Sebastian P. Bayerl, Korbinian Riedhammer, Adria Mallol-Ragolta, Maria Pateraki, Harry Coppock, Ivan Kiskin, Marianne Sinka, Stephen Roberts

Comments 5 pages, part of the ACM Multimedia 2022 Grand Challenge "The ACM Multimedia 2022 Computational Paralinguistics Challenge (ComParE 2022)"

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.09523 2022-04-21 cs.GR

SILVR: A Synthetic Immersive Large-Volume Plenoptic Dataset

Martijn Courteaux, Julie Artois, Stijn De Pauw, Peter Lambert, Glenn Van Wallendael

Comments In 13th ACM Multimedia Systems Conference (MMSys '22), June 14-17, 2022, Athlone, Ireland. ACM, New York, NY, USA, 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.09151 2022-04-11 cs.CV cs.CL cs.LG

Group-based Distinctive Image Captioning with Memory Attention

Jiuniu Wang, Wenjia Xu, Qingzhong Wang, Antoni B. Chan

Comments Accepted at ACM MM 2021 (oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.00792 2022-04-05 cs.CV

IR-GAN: Image Manipulation with Linguistic Instruction by Increment Reasoning

Zhenhuan Liu, Jincan Deng, Liang Li, Shaofei Cai, Qianqian Xu, Shuhui Wang, Qingming Huang

Journal ref Proceedings of the 28th ACM International Conference on Multimedia,2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.06254 2022-03-29 cs.CV

ConsNet: Learning Consistency Graph for Zero-Shot Human-Object Interaction Detection

Ye Liu, Junsong Yuan, Chang Wen Chen

Comments Accepted to Proceedings of the 28th ACM International Conference on Multimedia (MM 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.00361 2022-03-25 cs.CV

Answer-Driven Visual State Estimator for Goal-Oriented Visual Dialogue

Zipeng Xu, Fangxiang Feng, Xiaojie Wang, Yushu Yang, Huixing Jiang, Zhongyuan Wang

Comments Accepted at ACM International Conference on Multimedia (ACM MM 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.04222 2022-03-15 cs.CV cs.MM

Classification-Then-Grounding: Reformulating Video Scene Graphs as Temporal Bipartite Graphs

Kaifeng Gao, Long Chen, Yulei Niu, Jian Shao, Jun Xiao

Comments Accepted by CVPR 2022. Code is available at https://github.com/Dawn-LX/VidSGG-BIG. We also won the 1st place of Video Relation Understanding (VRU) Grand Challenge in ACM Multimedia 2021, with a simplified version of our model.(The code for object tracklets generation is available at https://github.com/Dawn-LX/VidVRD-tracklets)

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.10301 2022-02-22 cs.CV cs.AI

VLAD-VSA: Cross-Domain Face Presentation Attack Detection with Vocabulary Separation and Adaptation

Jiong Wang, Zhou Zhao, Weike Jin, Xinyu Duan, Zhen Lei, Baoxing Huai, Yiling Wu, Xiaofei He

Comments ACM MM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00866 2022-02-03 cs.CV

Decoupled IoU Regression for Object Detection

Yan Gao, Qimeng Wang, Xu Tang, Haochen Wang, Fei Ding, Jing Li, Yao Hu

Comments ACMMM 2021 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏