arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1944
2403.14652 2024-03-25 cs.CY cs.AI cs.CL cs.MM

MemeCraft: Contextual and Stance-Driven Multimodal Meme Generation

Han Wang, Roy Ka-Wei Lee

Comments 8 pages, 7 figures, ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.04834 2024-03-21 cs.CV

View while Moving: Efficient Video Recognition in Long-untrimmed Videos

Ye Tian, Mengyu Yang, Lanshan Zhang, Zhizhen Zhang, Yang Liu, Xiaohui Xie, Xirong Que, Wendong Wang

Comments Published on ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12289 2024-03-20 cs.NI eess.SP

BostonTwin: the Boston Digital Twin for Ray-Tracing in 6G Networks

Paolo Testolina, Michele Polese, Pedram Johari, Tommaso Melodia

Comments 7 pages, 5 figures, 3 tables. This paper has been accepted for presentation at ACM Multimedia Systems Conference 2024 (MMSys '24). Copyright ACM 2024. Please cite it as: P.Testolina, M. Polese, P. Johari, and T. Melodia, "BostonTwin: the Boston Digital Twin for Ray-Tracing in 6G Networks," in Proceedings of the ACM Multimedia Systems Conference 2024, ser. MMSys '24. Bari, Italy, Apr. 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03850 2024-03-19 cs.MM

Predictive Sampling for Efficient Pairwise Subjective Image Quality Assessment

Shima Mohammadi, João Ascenso

Comments 9 pages, 5 figures, accepted by ACM MM 2023

Journal ref Shima Mohammadi and João Ascenso. 2023. Predictive Sampling for Efficient Pairwise Subjective Image Quality Assessment. In Proceedings of the 31st ACM International Conference on Multimedia (MM '23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05856 2024-03-12 cs.CV

POV: Prompt-Oriented View-Agnostic Learning for Egocentric Hand-Object Interaction in the Multi-View World

Boshen Xu, Sipeng Zheng, Qin Jin

Comments Accepted by ACM MM 2023. Project page: https://xuboshen.github.io/

Journal ref Proceedings of the 31st ACM International Conference on Multimedia (2023). Association for Computing Machinery, New York, NY, USA, 2807-2816

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01087 2024-03-05 cs.MM cs.CV cs.SD eess.AS

Towards Accurate Lip-to-Speech Synthesis in-the-Wild

Sindhu Hegde, Rudrabha Mukhopadhyay, C. V. Jawahar, Vinay Namboodiri

Comments 8 pages of content, 1 page of references and 4 figures

Journal ref In Proceedings of the 31st ACM International Conference on Multimedia, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18171 2024-02-29 cs.CV

Digging Into Normal Incorporated Stereo Matching

Zihua Liu, Songyan Zhang, Zhicheng Wang, Masatoshi Okutomi

Journal ref Proceedings of the 30th ACM International Conference on Multimedia (ACMMM2022), pp.6050-6060, October 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12252 2024-02-29 cs.CV cs.LG cs.MM

Long-Range Feature Propagating for Natural Image Matting

Qinglin Liu, Haozhe Xie, Shengping Zhang, Bineng Zhong, Rongrong Ji

Journal ref ACM International Conference on Multimedia (ACM MM) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14473 2024-02-23 cs.IR cs.AI

Personalized Behavior-Aware Transformer for Multi-Behavior Sequential Recommendation

Jiajie Su, Chaochao Chen, Zibin Lin, Xi Li, Weiming Liu, Xiaolin Zheng

Journal ref Proceedings of the 31st ACM International Conference on Multimedia. 2023: 6321-6331

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.09841 2024-02-22 cs.CV

Exploring Inconsistent Knowledge Distillation for Object Detection with Data Augmentation

Jiawei Liang, Siyuan Liang, Aishan Liu, Ke Ma, Jingzhi Li, Xiaochun Cao

Comments ACMMM 2023 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.11812 2024-02-20 cs.CV cs.MM

Interpretable Embedding for Ad-hoc Video Search

Jiaxin Wu, Chong-Wah Ngo

Comments accepted in ACMMM 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08213 2024-02-09 cs.MM cs.CV

Accelerated Event-Based Feature Detection and Compression for Surveillance Video Systems

Andrew C. Freeman, Ketan Mayer-Patel, Montek Singh

Comments Accepted for publication in the proceedings of ACM Multimedia Systems '24

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.02982 2024-01-26 cs.CV

Beyond First Impressions: Integrating Joint Multi-modal Cues for Comprehensive 3D Representation

Haowei Wang, Jiji Tang, Jiayi Ji, Xiaoshuai Sun, Rongsheng Zhang, Yiwei Ma, Minda Zhao, Lincheng Li, zeng zhao, Tangjie Lv, Rongrong Ji

Comments ACM MM 2023, 3D Understanding, JM3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.01339 2024-01-23 cs.CV

DSE-GAN: Dynamic Semantic Evolution Generative Adversarial Network for Text-to-Image Generation

Mengqi Huang, Zhendong Mao, Penghui Wang, Quan Wang, Yongdong Zhang

Journal ref ACM Multimedia 2022 Best Student Paper Award

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.02119 2024-01-22 cs.CV

Hierarchical Masked 3D Diffusion Model for Video Outpainting

Fanda Fan, Chaoxu Guo, Litong Gong, Biao Wang, Tiezheng Ge, Yuning Jiang, Chunjie Luo, Jianfeng Zhan

Comments Accepted to ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.09883 2024-01-19 cs.CV

Question-Answer Cross Language Image Matching for Weakly Supervised Semantic Segmentation

Songhe Deng, Wei Zhuo, Jinheng Xie, Linlin Shen

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.13700 2023-12-25 cs.CV

UniNeXt: Exploring A Unified Architecture for Vision Recognition

Fangjian Lin, Jianlong Yuan, Sitong Wu, Fan Wang, Zhibin Wang

Comments Accep by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.06154 2023-12-20 cs.LG cs.AI

HypLL: The Hyperbolic Learning Library

Max van Spengler, Philipp Wirth, Pascal Mettes

Comments ACM Multimedia Open-Source Software Competition 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.14628 2023-12-20 cs.CV cs.RO

Multi-Frame Self-Supervised Depth Estimation with Multi-Scale Feature Fusion in Dynamic Scenes

Jiquan Zhong, Xiaolin Huang, Xiao Yu

Comments 11 pages, 8 figures, ACM MM'23 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00347 2023-12-19 cs.CV cs.CL cs.MM

RTQ: Rethinking Video-language Understanding Based on Image-text Model

Xiao Wang, Yaoyu Li, Tian Gan, Zheng Zhang, Jingjing Lv, Liqiang Nie

Comments Accepted by ACM MM 2023 as Oral representation

Journal ref In International Conference on Multimedia. ACM, 557--566 (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.00872 2023-12-19 cs.CV

Fearless Luminance Adaptation: A Macro-Micro-Hierarchical Transformer for Exposure Correction

Gehui Li, Jinyuan Liu, Long Ma, Zhiying Jiang, Xin Fan, Risheng Liu

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09753 2023-12-18 cs.MM

MORE: A Multimodal Object-Entity Relation Extraction Dataset with a Benchmark Evaluation

Liang He, Hongke Wang, Yongchang Cao, Zhen Wu, Jianbing Zhang, Xinyu Dai

Journal ref In Proceedings of the 31st ACM International Conference on Multimedia, pp. 4564-4573. 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06094 2023-12-12 cs.CL cs.CV cs.MM

MATK: The Meme Analytical Tool Kit

Ming Shan Hee, Aditi Kumaresan, Nguyen Khoi Hoang, Nirmalendu Prakash, Rui Cao, Roy Ka-Wei Lee

Comments Accepted at ACM Multimedia'23 Open-Source Software Competition Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06093 2023-12-12 cs.CL cs.CV cs.MM

PromptMTopic: Unsupervised Multimodal Topic Modeling of Memes using Large Language Models

Nirmalendu Prakash, Han Wang, Nguyen Khoi Hoang, Ming Shan Hee, Roy Ka-Wei Lee

Comments Accepted at ACM Multimedia'23 Research Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18286 2023-12-01 cs.CV

SimulFlow: Simultaneously Extracting Feature and Identifying Target for Unsupervised Video Object Segmentation

Lingyi Hong, Wei Zhang, Shuyong Gao, Hong Lu, WenQiang Zhang

Comments Accepted to ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.05189 2023-11-30 cs.CL cs.CV

SUR-adapter: Enhancing Text-to-Image Pre-trained Diffusion Models with Large Language Models

Shanshan Zhong, Zhongzhan Huang, Wushao Wen, Jinghui Qin, Liang Lin

Comments accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.14911 2023-11-28 cs.CV

CUCL: Codebook for Unsupervised Continual Learning

Chen Cheng, Jingkuan Song, Xiaosu Zhu, Junchen Zhu, Lianli Gao, Hengtao Shen

Comments MM '23: Proceedings of the 31st ACM International Conference on Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.14905 2023-11-28 cs.CV

Class Gradient Projection For Continual Learning

Cheng Chen, Ji Zhang, Jingkuan Song, Lianli Gao

Comments MM '22: Proceedings of the 30th ACM International Conference on Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09600 2023-11-27 cs.CV

Hawkeye: A PyTorch-based Library for Fine-Grained Image Recognition with Deep Learning

Jiabei He, Yang Shen, Xiu-Shen Wei, Ye Wu

Comments ACM Multimedia 2023 Open Source Software Competition Winner Entry. X.-S. Wei is the corresponding author

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13929 2023-11-27 cs.CV

MetaFBP: Learning to Learn High-Order Predictor for Personalized Facial Beauty Prediction

Luojun Lin, Zhifeng Shen, Jia-Li Yin, Qipeng Liu, Yuanlong Yu, Weijie Chen

Comments Accepted by ACM MM 2023. Source code: https://github.com/MetaVisionLab/MetaFBP

详情

展开后加载摘要…

URL PDF HTML 收藏