arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

ACM International Conference on Multimedia · 会议 · Multimedia

共收录 1944
2310.12724 2023-10-20 cs.CV

Query-aware Long Video Localization and Relation Discrimination for Deep Video Understanding

Yuanxing Xu, Yuting Wei, Bin Wu

Comments ACM MM 2023 Grand Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.11210 2023-10-18 cs.CV cs.MM

Learning Comprehensive Representations with Richer Self for Text-to-Image Person Re-Identification

Shuanglin Yan, Neng Dong, Jun Liu, Liyan Zhang, Jinhui Tang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09678 2023-10-17 cs.CV cs.AI cs.MM cs.RO

PoSynDA: Multi-Hypothesis Pose Synthesis Domain Adaptation for Robust 3D Human Pose Estimation

Hanbing Liu, Jun-Yan He, Zhi-Qi Cheng, Wangmeng Xiang, Qize Yang, Wenhao Chai, Gaoang Wang, Xu Bao, Bin Luo, Yifeng Geng, Xuansong Xie

Comments Accepted to ACM Multimedia 2023; 10 pages, 4 figures, 8 tables; the code is at https://github.com/hbing-l/PoSynDA

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.08696 2023-10-17 cs.CV cs.AI cs.MM

Improving Anomaly Segmentation with Multi-Granularity Cross-Domain Alignment

Ji Zhang, Xiao Wu, Zhi-Qi Cheng, Qi He, Wei Li

Comments Accepted to ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.10090 2023-10-17 cs.LG cs.AI

Orthogonal Uncertainty Representation of Data Manifold for Robust Long-Tailed Learning

Yanbiao Ma, Licheng Jiao, Fang Liu, Shuyuan Yang, Xu Liu, Lingling Li

Comments 10pages,Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09533 2023-10-17 cs.CV

Towards End-to-End Unsupervised Saliency Detection with Self-Supervised Top-Down Context

Yicheng Song, Shuyong Gao, Haozhe Xing, Yiting Cheng, Yan Wang, Wenqiang Zhang

Comments accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05648 2023-10-17 cs.CV

Counterfactual Cross-modality Reasoning for Weakly Supervised Video Moment Localization

Zezhong Lv, Bing Su, Ji-Rong Wen

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08892 2023-10-16 cs.CV

Image Cropping under Design Constraints

Takumi Nishiyasu, Wataru Shimoda, Yoichi Sato

Comments ACMMM Asia accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08117 2023-10-13 cs.CV cs.AI

DUSA: Decoupled Unsupervised Sim2Real Adaptation for Vehicle-to-Everything Collaborative Perception

Xianghao Kong, Wentao Jiang, Jinrang Jia, Yifeng Shi, Runsheng Xu, Si Liu

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08032 2023-10-13 cs.AI

Incorporating Domain Knowledge Graph into Multimodal Movie Genre Classification with Self-Supervised Attention and Contrastive Learning

Jiaqi Li, Guilin Qi, Chuanyi Zhang, Yongrui Chen, Yiming Tan, Chenlong Xia, Ye Tian

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07287 2023-10-12 cs.MM

Interactive Interior Design Recommendation via Coarse-to-fine Multimodal Reinforcement Learning

He Zhang, Ying Sun, Weiyu Guo, Yafei Liu, Haonan Lu, Xiaodong Lin, Hui Xiong

Comments Accepted by ACM International Conference on Multimedia'23. 9 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07222 2023-10-12 cs.CV

Uni-paint: A Unified Framework for Multimodal Image Inpainting with Pretrained Diffusion Model

Shiyuan Yang, Xiaodong Chen, Jing Liao

Comments Accepted by ACMMM'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.06138 2023-10-11 cs.CV cs.AI cs.LG cs.MM cs.RO

Layout Sequence Prediction From Noisy Mobile Modality

Haichao Zhang, Yi Xu, Hongsheng Lu, Takayuki Shimizu, Yun Fu

Comments In Proceedings of the 31st ACM International Conference on Multimedia 2023 (MM 23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04999 2023-10-11 cs.CV

Symmetrical Linguistic Feature Distillation with CLIP for Scene Text Recognition

Zixiao Wang, Hongtao Xie, Yuxin Wang, Jianjun Xu, Boqiang Zhang, Yongdong Zhang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.03504 2023-10-11 cs.CV

Stroke-based Neural Painting and Stylization with Dynamically Predicted Painting Region

Teng Hu, Ran Yi, Haokun Zhu, Liang Liu, Jinlong Peng, Yabiao Wang, Chengjie Wang, Lizhuang Ma

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05847 2023-10-10 cs.LG cs.AI cs.CR cs.IR

Making Users Indistinguishable: Attribute-wise Unlearning in Recommender Systems

Yuyuan Li, Chaochao Chen, Xiaolin Zheng, Yizhao Zhang, Zhongxuan Han, Dan Meng, Jun Wang

Journal ref Proceedings of the 31st ACM International Conference on Multimedia (MM '23), October 29--November 3, 2023, Ottawa, ON, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05589 2023-10-10 cs.CL cs.MM

DRIN: Dynamic Relation Interactive Network for Multimodal Entity Linking

Shangyu Xing, Fei Zhao, Zhen Wu, Chunhui Li, Jianbing Zhang, Xinyu Dai

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04689 2023-10-10 cs.CV

SeeDS: Semantic Separable Diffusion Synthesizer for Zero-shot Food Detection

Pengfei Zhou, Weiqing Min, Yang Zhang, Jiajun Song, Ying Jin, Shuqiang Jiang

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04679 2023-10-10 eess.IV cs.CV

High Visual-Fidelity Learned Video Compression

Meng Li, Yibo Shi, Jing Wang, Yunqi Huang

Comments ACMMM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04456 2023-10-10 cs.CL cs.SD eess.AS

Multimodal Prompt Transformer with Hybrid Contrastive Learning for Emotion Recognition in Conversation

Shihao Zou, Xianying Huang, Xudong Shen

Comments Accepted to ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13551 2023-10-10 cs.HC cs.AI cs.GR

Dance with You: The Diversity Controllable Dancer Generation via Diffusion Models

Siyue Yao, Mingjie Sun, Bingliang Li, Fengyu Yang, Junle Wang, Ruimao Zhang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03272 2023-10-10 cs.CV

Feature-Suppressed Contrast for Self-Supervised Food Pre-training

Xinda Liu, Yaohui Zhu, Linhu Liu, Jiang Tian, Lili Wang

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03670 2023-10-06 cs.CV

Regress Before Construct: Regress Autoencoder for Point Cloud Self-supervised Learning

Yang Liu, Chen Chen, Can Wang, Xulin King, Mengyuan Liu

Journal ref In Proceedings of the 31st ACM International Conference on Multimedia (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.02361 2023-10-05 cs.RO

Event-Enhanced Multi-Modal Spiking Neural Network for Dynamic Obstacle Avoidance

Yang Wang, Bo Dong, Yuji Zhang, Yunduo Zhou, Haiyang Mei, Ziqi Wei, Xin Yang

Comments In Proceedings of the 31st ACM International Conference on Multimedia (ACM MM 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.17104 2023-10-04 cs.CV

Prototype-guided Cross-modal Completion and Alignment for Incomplete Text-based Person Re-identification

Tiantian Gong, Guodong Du, Junsheng Wang, Yongkang Ding, Liyan Zhang

Comments Sorry, some collaborators do not agree to publish it on Arxiv, so please withdraw this paper

Journal ref ACM International Conference on Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.06034 2023-10-03 cs.LG

Normality Learning-based Graph Anomaly Detection via Multi-Scale Contrastive Learning

Jingcan Duan, Pei Zhang, Siwei Wang, Jingtao Hu, Hu Jin, Jiaxin Zhang, Haifang Zhou, Xinwang Liu

Comments 10 pages, 7 figures, accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.06466 2023-10-03 eess.IV

U2Net: A General Framework with Spatial-Spectral-Integrated Double U-Net for Image Fusion

Siran Peng, Chenhao Guo, Xiao Wu, Liang-Jian Deng

Comments Accepted by the 31st ACM International Conference on Multimedia (ACM MM '23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.09923 2023-10-02 cs.CV

Zero-shot point cloud segmentation by transferring geometric primitives

Runnan Chen, Xinge Zhu, Nenglun Chen, Wei Li, Yuexin Ma, Ruigang Yang, Wenping Wang

Comments ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16140 2023-09-29 cs.MM cs.CV

CLIP-Hand3D: Exploiting 3D Hand Pose Estimation via Context-Aware Prompting

Shaoxiang Guo, Qing Cai, Lin Qi, Junyu Dong

Comments Accepted In Proceedings of the 31st ACM International Conference on Multimedia (MM' 23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14359 2023-09-29 cs.AI cs.CL

Effect of Attention and Self-Supervised Speech Embeddings on Non-Semantic Speech Tasks

Payal Mohapatra, Akash Pandey, Yueyuan Sui, Qi Zhu

Comments Accepted to appear at ACM Multimedia 2023 Multimedia Grand Challenges Track

详情

展开后加载摘要…

URL PDF HTML 收藏