arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1944
2311.17955 2024-07-24 cs.CV

PEAN: A Diffusion-Based Prior-Enhanced Attention Network for Scene Text Image Super-Resolution

Zuoyan Zhao, Hui Xue, Pengfei Fang, Shipeng Zhu

Comments Accepted by ACMMM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15155 2024-07-23 cs.CV cs.AI cs.MM

Distilling Vision-Language Foundation Models: A Data-Free Approach via Prompt Diversification

Yunyi Xuan, Weijie Chen, Shicai Yang, Di Xie, Luojun Lin, Yueting Zhuang

Comments Accepted by ACMMM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15050 2024-07-23 cs.LG cs.AI cs.CR cs.MM

Arondight: Red Teaming Large Vision Language Models with Auto-generated Multi-modal Jailbreak Prompts

Yi Liu, Chengjun Cai, Xiaoli Zhang, Xingliang Yuan, Cong Wang

Comments To be published in ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14796 2024-07-23 cs.CV cs.AI

PASSION: Towards Effective Incomplete Multi-Modal Medical Image Segmentation with Imbalanced Missing Rates

Junjie Shi, Caozhi Shang, Zhaobin Sun, Li Yu, Xin Yang, Zengqiang Yan

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13768 2024-07-23 cs.CV cs.AI

Addressing Imbalance for Class Incremental Learning in Medical Image Classification

Xuze Hao, Wenqian Ni, Xuhao Jiang, Weimin Tan, Bo Yan

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12876 2024-07-23 cs.SI cs.AI cs.CL

Exploring the Use of Abusive Generative AI Models on Civitai

Yiluo Wei, Yiming Zhu, Pan Hui, Gareth Tyson

Comments Accepted to ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09342 2024-07-23 cs.CV cs.SD eess.AS

Face-voice Association in Multilingual Environments (FAME) Challenge 2024 Evaluation Plan

Muhammad Saad Saeed, Shah Nawaz, Muhammad Salman Tahir, Rohan Kumar Das, Muhammad Zaigham Zaheer, Marta Moscati, Markus Schedl, Muhammad Haris Khan, Karthik Nandakumar, Muhammad Haroon Yousaf

Comments ACM Multimedia Conference - Grand Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13976 2024-07-22 cs.CV

PlacidDreamer: Advancing Harmony in Text-to-3D Generation

Shuo Huang, Shikun Sun, Zixuan Wang, Xiaoyu Qin, Yanmin Xiong, Yuan Zhang, Pengfei Wan, Di Zhang, Jia Jia

Comments Accepted by ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.09802 2024-07-19 eess.AS cs.CV cs.SD

Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation

Minsu Kim, Jeong Hun Yeo, Se Jin Park, Hyeongseop Rha, Yong Man Ro

Comments ACMMM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12622 2024-07-18 cs.CV

Rethinking the Architecture Design for Efficient Generic Event Boundary Detection

Ziwei Zheng, Zechuan Zhang, Yulin Wang, Shiji Song, Gao Huang, Le Yang

Comments ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12339 2024-07-18 cs.CV

Exploring Deeper! Segment Anything Model with Depth Perception for Camouflaged Object Detection

Zhenni Yu, Xiaoqin Zhang, Li Zhao, Yi Bin, Guobao Xiao

Comments ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12255 2024-07-18 cs.CV

Dual-Hybrid Attention Network for Specular Highlight Removal

Xiaojiao Guo, Xuhang Chen, Shenghong Luo, Shuqiang Wang, Chi-Man Pun

Comments Accepted by ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.10471 2024-07-18 cs.CR cs.AI cs.SD eess.AS

GROOT: Generating Robust Watermark for Diffusion-Model-Based Audio Synthesis

Weizhi Liu, Yue Li, Dongdong Lin, Hui Tian, Haizhou Li

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09956 2024-07-18 cs.SD cs.AI cs.CL eess.AS

Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization

Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal, Wei-Ning Hsu, Rada Mihalcea, Soujanya Poria

Comments Accepted at ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01097 2024-07-18 cs.CV cs.LG

Spatio-Temporal Branching for Motion Prediction using Motion Increments

Jiexin Wang, Yujie Zhou, Wenwen Qiang, Ying Ba, Bing Su, Ji-Rong Wen

Comments The incremental information of our paper includes the displacement information from the last frame of the historical sequence, derived from the motion information of the first frame in the future sequence and the motion information of the last frame of the historical sequence. This implicitly contains future information, inadvertently giving an unfair advantage in the human motion prediction task

Journal ref ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11405 2024-07-17 cs.CR cs.CV

Cover-separable Fixed Neural Network Steganography via Deep Generative Models

Guobiao Li, Sheng Li, Zhenxing Qian, Xinpeng Zhang

Comments Accepetd at ACMMM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00934 2024-07-17 cs.CV

LanEvil: Benchmarking the Robustness of Lane Detection to Environmental Illusions

Tianyuan Zhang, Lu Wang, Hainan Li, Yisong Xiao, Siyuan Liang, Aishan Liu, Xianglong Liu, Dacheng Tao

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.16301 2024-06-25 cs.CV cs.AI cs.MM

UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Yuting Mei, Linli Yao, Qin Jin

Comments Accepted by ACM International Conference on Multimedia Retrieval (ICMR'24)

Journal ref Proceedings of the 2024 International Conference on Multimedia Retrieval, May 2024, Pages 1034-1042

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13409 2024-06-21 cs.CV cs.MM

PetalView: Fine-grained Location and Orientation Extraction of Street-view Images via Cross-view Local Search with Supplementary Materials

Wenmiao Hu, Yichen Zhang, Yuxuan Liang, Xianjing Han, Yifang Yin, Hannes Kruppa, See-Kiong Ng, Roger Zimmermann

Comments This paper has been accepted by ACM Multimedia 2023. This version contains additional supplementary materials

Journal ref Proceedings of the 31st ACM International Conference on Multimedia (2023) 56-66

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12316 2024-06-19 cs.CV cs.AI cs.MM

Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning

Ruiqi Wu, Bingliang Jiao, Wenxuan Wang, Meng Liu, Peng Wang

Comments Accepyed by ACM International Conference on Multimedia Retrieval (ICMR'24)

Journal ref ICMR'24: Proceedings of the 2024 International Conference on Multimedia Retrieval (2024) 579 - 588

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.08624 2024-06-18 cs.CV

Domain Camera Adaptation and Collaborative Multiple Feature Clustering for Unsupervised Person Re-ID

Yuanpeng Tu

Comments ACMMM 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10497 2024-05-20 cs.MM cs.AI cs.CV cs.SI

SMP Challenge: An Overview and Analysis of Social Media Prediction Challenge

Bo Wu, Peiye Liu, Wen-Huang Cheng, Bei Liu, Zhaoyang Zeng, Jia Wang, Qiushi Huang, Jiebo Luo

Comments ACM Multimedia. arXiv admin note: text overlap with arXiv:1910.01795

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18060 2024-04-30 cs.CV cs.LG

Prompt Customization for Continual Learning

Yong Dai, Xiaopeng Hong, Yabin Wang, Zhiheng Ma, Dongmei Jiang, Yaowei Wang

Comments ACM MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09530 2024-04-22 cs.CV cs.AI

RanLayNet: A Dataset for Document Layout Detection used for Domain Adaptation and Generalization

Avinash Anand, Raj Jaiswal, Mohit Gupta, Siddhesh S Bangar, Pijush Bhuyan, Naman Lal, Rajeev Singh, Ritika Jha, Rajiv Ratn Shah, Shin'ichi Satoh

Comments 8 pages, 6 figures, MMAsia 2023 Proceedings of the 5th ACM International Conference on Multimedia in Asia

Journal ref In Proceedings of the 5th ACM International Conference on Multimedia in Asia 2023. Association for Computing Machinery, NY, USA, Article 74, pp. 1-6

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13938 2024-04-18 cs.CV cs.LG

Improving Semi-Supervised Semantic Segmentation with Dual-Level Siamese Structure Network

Zhibo Tain, Xiaolin Zhang, Peng Zhang, Kun Zhan

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07855 2024-04-12 cs.CV

Resolve Domain Conflicts for Generalizable Remote Physiological Measurement

Weiyu Sun, Xinyu Zhang, Hao Lu, Ying Chen, Yun Ge, Xiaolin Huang, Jie Yuan, Yingcong Chen

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06033 2024-04-11 cs.CV

Little Strokes Fell Great Oaks: Boosting the Hierarchical Features for Multi-exposure Image Fusion

Pan Mu, Zhiying Du, Jinyuan Liu, Cong Bai

Journal ref Proceedings of the 31st ACM International Conference on Multimedia, October 2023, Pages 2985-2993

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.12886 2024-04-09 cs.CL cs.DL

Unveiling Global Narratives: A Multilingual Twitter Dataset of News Media on the Russo-Ukrainian Conflict

Sherzod Hakimov, Gullal S. Cheema

Comments ICMR 2024

Journal ref ICMR 2024 - ACM International Conference on Multimedia Retrieval 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18254 2024-04-02 cs.CV cs.AI

Sketch Input Method Editor: A Comprehensive Dataset and Methodology for Systematic Input Recognition

Guangming Zhu, Siyuan Wang, Qing Cheng, Kelong Wu, Hao Li, Liang Zhang

Comments The paper has been accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14326 2024-03-28 cs.MM

Think before You Leap: Content-Aware Low-Cost Edge-Assisted Video Semantic Segmentation

Mingxuan Yan, Yi Wang, Xuedou Xiao, Zhiqing Luo, Jianhua He, Wei Wang

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏