arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6929 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6929 篇

2207.08485 2022-07-20 cs.CV 57%

Hierarchical Feature Alignment Network for Unsupervised Video Object Segmentation

Gensheng Pei, Fumin Shen, Yazhou Yao, Guo-Sen Xie, Zhenmin Tang, Jinhui Tang

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments Accepted by ECCV-2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.08811 2022-07-20 cs.LG cs.AI 57%

Fusion of Physiological and Behavioural Signals on SPD Manifolds with Application to Stress and Pain Detection

Yujin WU, Mohamed Daoudi, Ali Amad, Laurent Sparrow, Fabien D'Hondt

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI

Comments International Conference on Systems, Man, and Cybernetics, IEEE SMC 2022, October 9-12, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.11094 2022-07-19 cs.CV 57%

GroupViT: Semantic Segmentation Emerges from Text Supervision

Jiarui Xu, Shalini De Mello, Sifei Liu, Wonmin Byeon, Thomas Breuel, Jan Kautz, Xiaolong Wang

专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV

Comments CVPR 2022. Project page and code: https://jerryxu.net/GroupViT

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.04185 2022-07-19 cs.CV cs.LG 57%

Transformaly -- Two (Feature Spaces) Are Better Than One

Matan Jacob Cohen, Shai Avidan

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments CVPR Workshop, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.06041 2022-07-14 cs.LG cs.AI 57%

Multiple Kernel Clustering with Dual Noise Minimization

Junpu Zhang, Liang Li, Siwei Wang, Jiyuan Liu, Yue Liu, Xinwang Liu, En Zhu

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.01607 2022-07-13 cs.LG cs.AI q-bio.NC 57%

Modern Views of Machine Learning for Precision Psychiatry

Zhe Sage Chen, Prathamesh, Kulkarni, Isaac R. Galatzer-Levy, Benedetta Bigio, Carla Nasca, Yu Zhang

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.04287 2022-07-06 cs.CV cs.LG 57%

Amplitude Spectrum Transformation for Open Compound Domain Adaptive Semantic Segmentation

Jogendra Nath Kundu, Akshay Kulkarni, Suvaansh Bhambri, Varun Jampani, R. Venkatesh Babu

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments AAAI 2022. Project page: http://sites.google.com/view/ast-ocdaseg

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.01172 2022-07-05 cs.CV 57%

TANet: Transformer-based Asymmetric Network for RGB-D Salient Object Detection

Chang Liu, Gang Yang, Shuo Wang, Hangxu Wang, Yunhua Zhang, Yutao Wang

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.09581 2022-06-22 cs.CV 57%

Explicit and implicit models in infrared and visible image fusion

Zixuan Wang, Bin Sun

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments 8 pages, 5 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.05381 2022-06-15 cs.CV 57%

Self-supervised Vision Transformers for Joint SAR-optical Representation Learning

Yi Wang, Conrad M Albrecht, Xiao Xiang Zhu

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 4 pages, 1 figure; IGARSS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.13077 2022-05-19 cs.CV cs.LG cs.RO 57%

Learning Selective Sensor Fusion for States Estimation

Changhao Chen, Stefano Rosa, Chris Xiaoxuan Lu, Bing Wang, Niki Trigoni, Andrew Markham

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Neural Networks and Learning Systems (TNNLS). arXiv admin note: text overlap with arXiv:1903.01534

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.08094 2022-05-18 cs.CV 57%

MATrIX -- Modality-Aware Transformer for Information eXtraction

Thomas Delteil, Edouard Belval, Lei Chen, Luis Goncalves, Vijay Mahadevan

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.04684 2022-05-12 cs.CV cs.LG 57%

OTFPF: Optimal Transport-Based Feature Pyramid Fusion Network for Brain Age Estimation with 3D Overlapped ConvNeXt

Yu Fu, Yanyan Huang, Yalin Wang, Shunjie Dong, Le Xue, Xunzhao Yin, Qianqian Yang, Yiyu Shi, Cheng Zhuo

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.03100 2022-05-09 cs.SI cs.AI 57%

Fake News Detection with Heterogeneous Transformer

Tianle Li, Yushi Sun, Shang-ling Hsu, Yanjia Li, Raymond Chi-Wing Wong

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.02411 2022-05-06 cs.CL 57%

Relational Representation Learning in Visually-Rich Documents

Xin Li, Yan Zheng, Yiqing Hu, Haoyu Cao, Yunfei Wu, Deqiang Jiang, Yinsong Liu, Bo Ren

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.13408 2022-05-06 cs.CV 57%

Learning Rich Features for Gait Recognition by Integrating Skeletons and Silhouettes

Yunjie Peng, Kang Ma, Yang Zhang, Zhiqiang He

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments The paper is under consideration at Multimedia Tools and Applications

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.14044 2022-05-02 cs.CV 57%

C3-STISR: Scene Text Image Super-resolution with Triple Clues

Minyi Zhao, Miao Wang, Fan Bai, Bingjia Li, Jie Wang, Shuigeng Zhou

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments Accepted by IJCAI 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.10992 2022-04-26 cs.CV stat.ML 57%

Expansion-Squeeze-Excitation Fusion Network for Elderly Activity Recognition

Xiangbo Shu, Jiawen Yang, Rui Yan, Yan Song

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.10803 2022-04-25 cs.CV 57%

Pay "Attention" to Adverse Weather: Weather-aware Attention-based Object Detection

Saket S. Chaturvedi, Lan Zhang, Xiaoyong Yuan

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments This paper is accepted at IEEE International Conference on Pattern Recognition (ICPR), 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.09518 2022-04-21 eess.SP cs.AI 57%

Simulation of machine learning-based 6G systems in virtual worlds

Ailton Oliveira, Felipe Bastos, Isabela Trindade, Walter Frazao, Arthur Nascimento, Diego Gomes, Francisco Muller, Aldebaro Klautau

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI

Journal ref ITU Journal on Future and Evolving Technologies, Volume 2 (2021), Issue 4 - AI and machine learning solutions in 5G and future networks, Pages 113-123

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.03004 2022-04-04 cs.CV 57%

MmWave Radar and Vision Fusion for Object Detection in Autonomous Driving: A Review

Zhiqing Wei, Fengkai Zhang, Shuo Chang, Yangyang Liu, Huici Wu, Zhiyong Feng

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.17109 2022-04-01 cs.AI 57%

A Rich Recipe Representation as Plan to Support Expressive Multi Modal Queries on Recipe Content and Preparation Process

Vishal Pallagani, Priyadharsini Ramamurthy, Vedant Khandelwal, Revathy Venkataramanan, Kausik Lakkaraju, Sathyanarayanan N. Aakur, Biplav Srivastava

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15618 2022-03-30 cs.CV 57%

Exploring Body Texture from mmW Images for Person Recognition

E. Gonzalez-Sosa, J. Fierrez, R. Vera-Rodriguez, F. Alonso-Fernandez, V. M. Patel

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Published at IEEE Transactions on Biometrics, Behavior, and Identity Science

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.11481 2022-03-30 cs.CV cs.CR 57%

Mixed Differential Privacy in Computer Vision

Aditya Golatkar, Alessandro Achille, Yu-Xiang Wang, Aaron Roth, Michael Kearns, Stefano Soatto

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments Accepted at CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.14581 2022-03-29 cs.CV 57%

S2-Net: Self-supervision Guided Feature Representation Learning for Cross-Modality Images

Shasha Mei

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.06935 2022-03-22 cs.MM 57%

A Systematic Review on Affective Computing: Emotion Models, Databases, and Recent Advances

Yan Wang, Wei Song, Wei Tao, Antonio Liotta, Dawei Yang, Xinlei Li, Shuyong Gao, Yixuan Sun, Weifeng Ge, Wei Zhang, Wenqiang Zhang

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.MM

Comments Accepted for Information Fusion

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.09913 2022-03-21 cs.CV cs.LG 57%

Convolutional Simultaneous Sparse Approximation with Applications to RGB-NIR Image Fusion

Farshad G. Veshki, Sergiy A. Vorobyov

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.07300 2022-03-15 cs.CV 57%

Mobile Behavioral Biometrics for Passive Authentication

Giuseppe Stragapede, Ruben Vera-Rodriguez, Ruben Tolosana, Aythami Morales, Alejandro Acien, Gael Le Lan

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.13251 2022-03-01 cs.CV cs.LG 57%

Supervising Remote Sensing Change Detection Models with 3D Surface Semantics

Isaac Corley, Peyman Najafirad

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.06673 2022-02-15 cs.CV 57%

Convolutional Neural Network with Convolutional Block Attention Module for Finger Vein Recognition

Zhongxia Zhang, Mingwen Wang

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 11 pages, 6 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏