arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1944
2303.00952 2024-08-07 cs.CV cs.RO eess.IV

Towards Activated Muscle Group Estimation in the Wild

Kunyu Peng, David Schneider, Alina Roitberg, Kailun Yang, Jiaming Zhang, Chen Deng, Kaiyu Zhang, M. Saquib Sarfraz, Rainer Stiefelhagen

Comments Accepted to ACM MM 2024. The database and code can be found at https://github.com/KPeng9510/MuscleMap

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.02100 2024-08-06 cs.CV

View-consistent Object Removal in Radiance Fields

Yiren Lu, Jing Ma, Yu Yin

Comments Accepted to ACM Multimedia (MM) 2024. Project website is accessible at https://vulab-ai.github.io/View-consistent_Object_Removal_in_Radiance_Fields

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01952 2024-08-06 cs.CV

CACE-Net: Co-guidance Attention and Contrastive Enhancement for Effective Audio-Visual Event Localization

Xiang He, Xiangxi Liu, Yang Li, Dongcheng Zhao, Guobin Shen, Qingqun Kong, Xin Yang, Yi Zeng

Comments Accepted by ACM MM 2024. Code is available at this https://github.com/Brain-Cog-Lab/CACE-Net

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01678 2024-08-06 cs.CV

iControl3D: An Interactive System for Controllable 3D Scene Generation

Xingyi Li, Yizheng Wu, Jun Cen, Juewen Peng, Kewei Wang, Ke Xian, Zhe Wang, Zhiguo Cao, Guosheng Lin

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01664 2024-08-06 cs.CV cs.AI

SAT3D: Image-driven Semantic Attribute Transfer in 3D

Zhijun Zhai, Zengmao Wang, Xiaoxiao Long, Kaixuan Zhou, Bo Du

Journal ref In Proceedings of the 32nd ACM International Conference on Multimedia, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01428 2024-08-06 cs.CV cs.AI

Transferable Adversarial Facial Images for Privacy Protection

Minghui Li, Jiangxiong Wang, Hao Zhang, Ziqi Zhou, Shengshan Hu, Xiaobing Pei

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01355 2024-08-06 cs.CV cs.MM

Hallu-PI: Evaluating Hallucination in Multi-modal Large Language Models within Perturbed Inputs

Peng Ding, Jingyu Wu, Jun Kuang, Dan Ma, Xuezhi Cao, Xunliang Cai, Shi Chen, Jiajun Chen, Shujian Huang

Comments Acccepted by ACM MM 2024, 14 pages, 11 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09026 2024-08-06 cs.CV cs.LG cs.MM eess.IV

HPC: Hierarchical Progressive Coding Framework for Volumetric Video

Zihan Zheng, Houqiang Zhong, Qiang Hu, Xiaoyun Zhang, Li Song, Ya Zhang, Yanfeng Wang

Comments 11 pages, 7 figures, ACM Multimedia 24

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.17762 2024-08-06 cs.CV

Large Multi-modality Model Assisted AI-Generated Image Quality Assessment

Puyi Wang, Wei Sun, Zicheng Zhang, Jun Jia, Yanwei Jiang, Zhichao Zhang, Xiongkuo Min, Guangtao Zhai

Comments ACM MM'24

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01349 2024-08-05 cs.MM cs.AI cs.CV cs.IR cs.LG

PC$^2$: Pseudo-Classification Based Pseudo-Captioning for Noisy Correspondence Learning in Cross-Modal Retrieval

Yue Duan, Zhangxuan Gu, Zhenzhe Ying, Lei Qi, Changhua Meng, Yinghuan Shi

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01269 2024-08-05 cs.CV

A General Framework to Boost 3D GS Initialization for Text-to-3D Generation by Lexical Richness

Lutao Jiang, Hangyu Li, Lin Wang

Journal ref ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21043 2024-08-05 cs.CL cs.AI cs.LG

CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning

Yu Feng, Zhen Tian, Yifan Zhu, Zongfu Han, Haoran Luo, Guangwei Zhang, Meina Song

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01218 2024-08-05 cs.CV

S2TD-Face: Reconstruct a Detailed 3D Face with Controllable Texture from a Single Sketch

Zidu Wang, Xiangyu Zhu, Jiang Yu, Tianshuo Zhang, Zhen Lei

Comments ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00644 2024-08-02 cs.CV

Towards End-to-End Explainable Facial Action Unit Recognition via Vision-Language Joint Learning

Xuri Ge, Junchen Fu, Fuhai Chen, Shan An, Nicu Sebe, Joemon M. Jose

Comments 10 pages, 5 figures, 4 tables

Journal ref ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00491 2024-08-02 cs.CL cs.CV cs.MM

GalleryGPT: Analyzing Paintings with Large Multimodal Models

Yi Bin, Wenhao Shi, Yujuan Ding, Zhiqiang Hu, Zheng Wang, Yang Yang, See-Kiong Ng, Heng Tao Shen

Comments Accepted as Oral Presentation at ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00441 2024-08-02 cs.CV cs.AI

Focus, Distinguish, and Prompt: Unleashing CLIP for Efficient and Flexible Scene Text Retrieval

Gangyan Zeng, Yuan Zhang, Jin Wei, Dongbao Yang, Peng Zhang, Yiwen Gao, Xugong Qin, Yu Zhou

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00305 2024-08-02 cs.MM cs.IR

Leveraging Weak Cross-Modal Guidance for Coherence Modelling via Iterative Learning

Yi Bin, Junrong Liao, Yujuan Ding, Haoxuan Li, Yang Yang, See-Kiong Ng, Heng Tao Shen

Comments Accepted by ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00257 2024-08-02 cs.AI

RoCo:Robust Collaborative Perception By Iterative Object Matching and Pose Adjustment

Zhe Huang, Shuo Wang, Yongcai Wang, Wanting Li, Deying Li, Lei Wang

Comments ACM MM2024

Journal ref Proceedings of the 32nd ACM International Conference on Multimedia (MM '24), October 28-November 1, 2024, Melbourne, VIC, Australia

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21491 2024-08-02 cs.CL cs.SD eess.AS

Generative Expressive Conversational Speech Synthesis

Rui Liu, Yifan Hu, Yi Ren, Xiang Yin, Haizhou Li

Comments 14 pages, 6 figures, 8 tables. Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14360 2024-08-02 cs.CV

Deblurring Neural Radiance Fields with Event-driven Bundle Adjustment

Yunshan Qi, Lin Zhu, Yifan Zhao, Nan Bao, Jia Li

Comments Accepted by 32nd ACM International Conference on Multimedia (MM 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03781 2024-08-02 cs.CV cs.AI

Lite-Mind: Towards Efficient and Robust Brain Representation Network

Zixuan Gong, Qi Zhang, Guangyin Bao, Lei Zhu, Ke Liu, Liang Hu, Duoqian Miao, Yu Zhang

Comments 17 pages, ACM MM 2024 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21742 2024-08-01 cs.LG cs.AI

HGOE: Hybrid External and Internal Graph Outlier Exposure for Graph Out-of-Distribution Detection

Junwei He, Qianqian Xu, Yangbangyan Jiang, Zitai Wang, Yuchen Sun, Qingming Huang

Comments Proceedings of the 32nd ACM International Conference on Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21721 2024-08-01 cs.MM cs.AI

Open-Vocabulary Audio-Visual Semantic Segmentation

Ruohao Guo, Liao Qu, Dantong Niu, Yanyu Qi, Wenzhen Yue, Ji Shi, Bowei Xing, Xianghua Ying

Comments Accepted by ACM MM 2024 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.14433 2024-08-01 cs.IR cs.MM

Attribute-driven Disentangled Representation Learning for Multimodal Recommendation

Zhenyang Li, Fan Liu, Yinwei Wei, Zhiyong Cheng, Liqiang Nie, Mohan Kankanhalli

Comments ACM Multimedia 2024 Accepted

Journal ref In Proceedings of the 32st ACM International Conference on Multimedia (MM '24), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20664 2024-08-01 cs.CV

3D-GRES: Generalized 3D Referring Expression Segmentation

Changli Wu, Yihang Liu, Jiayi Ji, Yiwei Ma, Haowei Wang, Gen Luo, Henghui Ding, Xiaoshuai Sun, Rongrong Ji

Comments Accepted by ACM MM 2024 (Oral), Code: https://github.com/sosppxo/MDIN

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.12682 2024-08-01 cs.CV

Transferring to Real-World Layouts: A Depth-aware Framework for Scene Adaptation

Mu Chen, Zhedong Zheng, Yi Yang

Comments ACM MM 2024 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20766 2024-07-31 eess.IV cs.CV

Highly Efficient No-reference 4K Video Quality Assessment with Full-Pixel Covering Sampling and Training Strategy

Xiaoheng Tan, Jiabin Zhang, Yuhui Quan, Jing Li, Yajing Wu, Zilin Bian

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20730 2024-07-31 cs.CV

Autogenic Language Embedding for Coherent Point Tracking

Zikai Song, Ying Tang, Run Luo, Lintao Ma, Junqing Yu, Yi-Ping Phoebe Chen, Wei Yang

Comments accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20693 2024-07-31 cs.CV cs.AI cs.MM

Boosting Audio Visual Question Answering via Key Semantic-Aware Cues

Guangyao Li, Henghui Du, Di Hu

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20124 2024-07-31 cs.MM cs.AI

AxiomVision: Accuracy-Guaranteed Adaptive Visual Model Selection for Perspective-Aware Video Analytics

Xiangxiang Dai, Zeyu Zhang, Peng Yang, Yuedong Xu, Xutong Liu, John C. S. Lui

Comments Accepted by ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏