arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1966
2104.13553 2021-04-29 eess.AS cs.LG cs.SD

AMSS-Net: Audio Manipulation on User-Specified Sources with Textual Queries

Woosung Choi, Minseok Kim, Marco A. Martínez Ramírez, Jaehwa Chung, Soonyoung Jung

Comments 10 pages, 8 figures, 3 tables, under reviewing of ACMMM 21

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.12471 2021-04-27 cs.CV cs.AI cs.CL cs.IR cs.MM

Contextualized Keyword Representations for Multi-modal Retinal Image Captioning

Jia-Hong Huang, Ting-Wei Wu, Marcel Worring

Comments This paper is accepted by ACM International Conference on Multimedia Retrieval (ICMR), 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.12465 2021-04-27 cs.CV cs.AI cs.CL cs.MM

GPT2MVS: Generative Pre-trained Transformer-2 for Multi-modal Video Summarization

Jia-Hong Huang, Luka Murn, Marta Mrak, Marcel Worring

Comments This paper is accepted by ACM International Conference on Multimedia Retrieval (ICMR), 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.05977 2021-04-13 cs.CV

Hybrid Dynamic-static Context-aware Attention Network for Action Assessment in Long Videos

Ling-An Zeng, Fa-Ting Hong, Wei-Shi Zheng, Qi-Zhi Yu, Wei Zeng, Yao-Wei Wang, Jian-Huang Lai

Comments ACM International Conference on Multimedia 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.04268 2021-04-12 cs.CR cs.CV

Reversible Watermarking in Deep Convolutional Neural Networks for Integrity Authentication

Xiquan Guan, Huamin Feng, Weiming Zhang, Hang Zhou, Jie Zhang, Nenghai Yu

Comments Accepted to ACM MM 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.02299 2021-04-07 cs.CV

Change Detection from SAR Images Based on Deformable Residual Convolutional Neural Networks

Junjie Wang, Feng Gao, Junyu Dong

Comments Accepted by ACM Multimedia Asia 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.15576 2021-04-02 cs.CV

Dense Scene Multiple Object Tracking with Box-Plane Matching

Jinlong Peng, Yueyang Gu, Yabiao Wang, Chengjie Wang, Jilin Li, Feiyue Huang

Comments ACM Multimedia 2020 GC paper. ACM Multimedia Grand Challenge HiEve 2020 Track-1 Winner

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.03212 2021-03-23 cs.CV cs.CL

Text-Guided Neural Image Inpainting

Lisai Zhang, Qingcai Chen, Baotian Hu, Shuoran Jiang

Comments ACM MM'2020 (Oral). 9 pages, 4 tables, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.02516 2021-03-16 eess.AS cs.CL cs.CV cs.LG cs.SD

FastLR: Non-Autoregressive Lipreading Model with Integrate-and-Fire

Jinglin Liu, Yi Ren, Zhou Zhao, Chen Zhang, Baoxing Huai, Nicholas Jing Yuan

Comments Accepted by ACM MM 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.08525 2021-03-12 cs.MM cs.CV

Building Movie Map -- A Tool for Exploring Areas in a City -- and its Evaluation

Naoki Sugimoto, Yoshihito Ebine, Kiyoharu Aizawa

Journal ref ACM Multimedia 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.14084 2021-03-12 cs.MM eess.IV eess.SP

Kalman Filter-based Head Motion Prediction for Cloud-based Mixed Reality

Serhan Gül, Sebastian Bosse, Dimitri Podborski, Thomas Schierl, Cornelius Hellge

Comments Accepted at the ACM Multimedia Conference (ACMMM) 2020. 9 pages, 9 figures

Journal ref Proceedings of the 28th ACM International Conference on Multimedia (2020) 3632-3641

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.05676 2021-03-04 cs.CV

Forest R-CNN: Large-Vocabulary Long-Tailed Object Detection and Instance Segmentation

Jialian Wu, Liangchen Song, Tiancai Wang, Qian Zhang, Junsong Yuan

Comments Accepted to ACM MM 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13273 2021-02-26 cs.CV

Group-Skeleton-Based Human Action Recognition in Complex Events

Tingtian Li, Zixun Sun, Xiao Chen

Comments accpeted by ACM MM 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.05907 2021-01-26 eess.IV cs.AI cs.CV cs.LG cs.MM

Attention Cube Network for Image Restoration

Yucheng Hang, Qingmin Liao, Wenming Yang, Yupeng Chen, Jie Zhou

Comments Accepted by the 28th ACM International Conference on Multimedia (ACM MM 2020); Code is available at https://github.com/YCHang686/A-CubeNet

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.01985 2021-01-25 cs.MM cs.CV

Intrinsic Image Popularity Assessment

Keyan Ding, Kede Ma, Shiqi Wang

Comments Accepted by ACM Multimedia 2019

Journal ref Proceedings of the 27th ACM International Conference on Multimedia, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.09402 2020-12-18 cs.CV

LIGHTEN: Learning Interactions with Graph and Hierarchical TEmporal Networks for HOI in videos

Sai Praneeth Reddy Sunkesula, Rishabh Dabral, Ganesh Ramakrishnan

Comments 9 pages, 6 figures, ACM Multimedia Conference 2020

Journal ref MM20 Proceedings of the 28th ACM International Conference on Multimedia, October 2020, Pages 691 to 699

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.05117 2020-12-16 cs.MM cs.DC cs.LG

Hysia: Serving DNN-Based Video-to-Retail Applications in Cloud

Huaizheng Zhang, Yuanming Li, Qiming Ai, Yong Luo, Yonggang Wen, Yichao Jin, Nguyen Binh Duong Ta

Comments 4 pages, 4 figures

Journal ref In Proceedings of the 28th ACM International Conference on Multimedia (2020) 4457-4460

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.05096 2020-12-16 cs.DC cs.LG

MLModelCI: An Automatic Cloud Platform for Efficient MLaaS

Huaizheng Zhang, Yuanming Li, Yizheng Huang, Yonggang Wen, Jianxiong Yin, Kyle Guan

Comments 4 pages, 4 figures

Journal ref In Proceedings of the 28th ACM International Conference on Multimedia (2020) 4453-4456

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.04260 2020-11-10 cs.CV

Robust Visual Tracking via Statistical Positive Sample Generation and Gradient Aware Learning

Lijian Lin, Haosheng Chen, Yanjie Liang, Yan Yan, Hanzi Wang

Comments 6 pages

Journal ref ACM MM Asia2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.03363 2020-11-09 cs.CV

Domain Adaptive Person Re-Identification via Coupling Optimization

Xiaobin Liu, Shiliang Zhang

Comments ACM MM 2020 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.05586 2020-11-03 cs.CV

ATRW: A Benchmark for Amur Tiger Re-identification in the Wild

Shuyuan Li, Jianguo Li, Hanlin Tang, Rui Qian, Weiyao Lin

Comments ACM Multimedia (MM) 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.06112 2020-10-28 cs.CV cs.LG eess.IV

Unified Generative Adversarial Networks for Controllable Image-to-Image Translation

Hao Tang, Hong Liu, Nicu Sebe

Comments Accepted to TIP, an extended version of a paper published in ACM MM 2018. arXiv admin note: substantial text overlap with arXiv:1808.04859

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.10008 2020-10-22 cs.CV

Towards Accurate Human Pose Estimation in Videos of Crowded Scenes

Li Yuan, Shuning Chang, Xuecheng Nie, Ziyuan Huang, Yichen Zhou, Yunpeng Chen, Jiashi Feng, Shuicheng Yan

Comments 2nd Place in ACM Multimedia Grand Challenge: Human in Events, Track2: Crowd Pose Estimation in Complex Events. ACM Multimedia 2020. arXiv admin note: substantial text overlap with arXiv:2010.08365, arXiv:2010.10007

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.10007 2020-10-22 cs.CV

A Simple Baseline for Pose Tracking in Videos of Crowded Scenes

Li Yuan, Shuning Chang, Ziyuan Huang, Yichen Zhou, Yunpeng Chen, Xuecheng Nie, Francis E. H. Tay, Jiashi Feng, Shuicheng Yan

Comments 2nd Place in ACM Multimedia Grand Challenge: Human in Events, Track3: Crowd Pose Tracking in Complex Events. ACM Multimedia 2020. arXiv admin note: substantial text overlap with arXiv:2010.08365, arXiv:2010.10008

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09982 2020-10-21 cs.CV

Depth Guided Adaptive Meta-Fusion Network for Few-shot Video Recognition

Yuqian Fu, Li Zhang, Junke Wang, Yanwei Fu, Yu-Gang Jiang

Comments accepted by ACM Multimedia 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09978 2020-10-21 cs.CV

Stronger, Faster and More Explainable: A Graph Convolutional Baseline for Skeleton-based Action Recognition

Yi-Fan Song, Zhang Zhang, Caifeng Shan, Liang Wang

Comments Accepted by ACM MultiMedia 2020, 9 pages, 4 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09768 2020-10-21 cs.SI

Not Judging a User by Their Cover: Understanding Harm in Multi-Modal Processing within Social Media Research

Jiachen Jiang, Soroush Vosoughi

Comments In proceedings of the 2nd International Workshop on Fairness, Accountability, Transparency and Ethics in Multimedia (FATE/MM'20). Held in conjunction with ACM Multimedia 2020 (MM 20)

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.03864 2020-10-20 cs.CV cs.LG eess.IV

Nighttime Dehazing with a Synthetic Benchmark

Jing Zhang, Yang Cao, Zheng-Jun Zha, Dacheng Tao

Comments ACM MM 2020. Both the dataset and source code will be available at \url{https://github.com/chaimi2013/3R}

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.03545 2020-10-20 cs.CL cs.LG

MISA: Modality-Invariant and -Specific Representations for Multimodal Sentiment Analysis

Devamanyu Hazarika, Roger Zimmermann, Soujanya Poria

Comments Accepted at ACM MM 2020 (Oral Paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.08365 2020-10-19 cs.CV

Toward Accurate Person-level Action Recognition in Videos of Crowded Scenes

Li Yuan, Yichen Zhou, Shuning Chang, Ziyuan Huang, Yunpeng Chen, Xuecheng Nie, Tao Wang, Jiashi Feng, Shuicheng Yan

Comments 1'st Place in ACM Multimedia Grand Challenge: Human in Events, Track4: Person-level Action Recognition in Complex Events

Journal ref ACM Multimedia 2020

详情

展开后加载摘要…

URL PDF HTML 收藏