arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 9189 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态评测 9189 篇

2309.15332 2023-10-02 cs.RO cs.CV 79%

Multimodal Dataset for Localization, Mapping and Crop Monitoring in Citrus Tree Farms

Hanzhe Teng, Yipeng Wang, Xiaoao Song, Konstantinos Karydis

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments Accepted to the 18th International Symposium on Visual Computing (ISVC 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14090 2023-09-26 cs.LG cs.CV 79%

Convolutional autoencoder-based multimodal one-class classification

Firas Laakom, Fahad Sohrab, Jenni Raitoharju, Alexandros Iosifidis, Moncef Gabbouj

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments 5 pages, 1 figure, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.12568 2023-09-25 cs.RO cs.AI 79%

A Study on Learning Social Robot Navigation with Multimodal Perception

Bhabaranjan Panigrahi, Amir Hossain Raj, Mohammad Nazeri, Xuesu Xiao

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.09445 2023-09-25 cs.CV 79%

aiMotive Dataset: A Multimodal Dataset for Robust Autonomous Driving with Long-Range Perception

Tamás Matuszka, Iván Barton, Ádám Butykai, Péter Hajas, Dávid Kiss, Domonkos Kovács, Sándor Kunsági-Máté, Péter Lengyel, Gábor Németh, Levente Pető, Dezső Ribli, Dávid Szeghy, Szabolcs Vajna, Bálint Varga

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments The paper was accepted to ICLR 2023 Workshop Scene Representations for Autonomous Driving

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11839 2023-09-22 cs.CV cs.RO 79%

MoPA: Multi-Modal Prior Aided Domain Adaptation for 3D Semantic Segmentation

Haozhi Cao, Yuecong Xu, Jianfei Yang, Pengyu Yin, Shenghai Yuan, Lihua Xie

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09867 2023-09-19 cs.SE cs.AI 79%

EGFE: End-to-end Grouping of Fragmented Elements in UI Designs with Multimodal Learning

Liuqing Chen, Yunnong Chen, Shuhong Xiao, Yaxuan Song, Lingyun Sun, Yankun Zhen, Tingting Zhou, Yanfang Chang

专题命中 多模态评测 :multimodal(title);multi-modal(abstract);分类 cs.AI

Comments Accepted to 46th International Conference on Software Engineering (ICSE 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08922 2023-09-19 cs.CL 79%

Multimodal Multi-Hop Question Answering Through a Conversation Between Tools and Efficiently Finetuned Large Language Models

Hossein Rajabzadeh, Suyuchen Wang, Hyock Ju Kwon, Bang Liu

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08474 2023-09-18 cs.CR cs.AI 79%

VulnSense: Efficient Vulnerability Detection in Ethereum Smart Contracts by Multimodal Learning with Graph Neural Network and Language Model

Phan The Duy, Nghi Hoang Khoa, Nguyen Huu Quyen, Le Cong Trinh, Vu Trung Kien, Trinh Minh Hoang, Van-Hau Pham

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08160 2023-09-18 eess.IV cs.CV 79%

Cross-Modal Synthesis of Structural MRI and Functional Connectivity Networks via Conditional ViT-GANs

Yuda Bi, Anees Abrol, Jing Sui, Vince Calhoun

专题命中 多模态评测 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.10161 2023-09-13 cs.CV 79%

ThermRad: A Multi-modal Dataset for Robust 3D Object Detection under Challenging Conditions

Qiao Yan, Yihan Wang

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments At this time, we have not reached a definitive agreement regarding the ownership and copyright of this dataset. Due to the unresolved issue regarding the dataset, I am writing to formally request the withdrawal of our paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.04634 2023-08-29 cs.CV cs.CY eess.IV 79%

Multimodal Noisy Segmentation based fragmented burn scars identification in Amazon Rainforest

Satyam Mohla, Sidharth Mohla, Anupam Guha, Biplab Banerjee

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments 5 pages, 5 figures. Accepted at IEEE International Conference on Systems, Man and Cybernetics 2020. Earlier draft presented at Harvard CRCS AI for Social Good Workshop 2020

Journal ref 2020 IEEE International Conference on Systems, Man, and Cybernetics (SMC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13411 2023-08-28 cs.CV 79%

Harvard Glaucoma Detection and Progression: A Multimodal Multitask Dataset and Generalization-Reinforced Semi-Supervised Learning

Yan Luo, Min Shi, Yu Tian, Tobias Elze, Mengyu Wang

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12163 2023-08-24 cs.CV 79%

NPF-200: A Multi-Modal Eye Fixation Dataset and Method for Non-Photorealistic Videos

Ziyu Yang, Sucheng Ren, Zongwei Wu, Nanxuan Zhao, Junle Wang, Jing Qin, Shengfeng He

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted by ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.11983 2023-08-24 cs.RO cs.CV 79%

Multi-Modal Multi-Task (3MT) Road Segmentation

Erkan Milli, Özgür Erkent, Asım Egemen Yılmaz

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Journal ref in IEEE Robotics and Automation Letters, vol. 8, no. 9, pp. 5408-5415, Sept. 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.10621 2023-08-22 cs.CV 79%

Multi-Modal Dataset Acquisition for Photometrically Challenging Object

HyunJun Jung, Patrick Ruhkamp, Nassir Navab, Benjamin Busam

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted at ICCV 2023 TRICKY Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.10491 2023-08-22 cs.CV 79%

SynDrone -- Multi-modal UAV Dataset for Urban Scenarios

Giulia Rizzoli, Francesco Barbato, Matteo Caligiuri, Pietro Zanuttigh

专题命中 多模态评测 :multi-modal(title);multimodal(abstract);分类 cs.CV

Comments Accepted at ICCV Workshops, downloadable dataset with CC-BY license, 8 pages, 4 figures, 8 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.08517 2023-08-17 cs.CV cs.LG 79%

Building RadiologyNET: Unsupervised annotation of a large-scale multimodal medical database

Mateja Napravnik, Franko Hržić, Sebastian Tschauner, Ivan Štajduhar

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.14520 2023-08-11 cs.RO cs.CV 79%

Multimodal Dataset from Harsh Sub-Terranean Environment with Aerosol Particles for Frontier Exploration

Alexander Kyuroson, Niklas Dahlquist, Nikolaos Stathoulopoulos, Vignesh Kottayam Viswanathan, Anton Koval, George Nikolakopoulos

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments Accepted in the 31st Mediterranean Conference on Control and Automation [MED2023]

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.00131 2023-08-09 cs.CV 79%

Guided Hybrid Quantization for Object detection in Multimodal Remote Sensing Imagery via One-to-one Self-teaching

Jiaqing Zhang, Jie Lei, Weiying Xie, Yunsong Li, Xiuping Jia

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments This article has been delivered to TRGS and is under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03906 2023-08-09 cs.CV 79%

TIJO: Trigger Inversion with Joint Optimization for Defending Multimodal Backdoored Models

Indranil Sur, Karan Sikka, Matthew Walmer, Kaushik Koneripalli, Anirban Roy, Xiao Lin, Ajay Divakaran, Susmit Jha

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CV

Comments Published as conference paper at ICCV 2023. 13 pages, 6 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.16144 2023-08-09 cs.RO cs.AI 79%

Towards trustworthy multi-modal motion prediction: Holistic evaluation and interpretability of outputs

Sandra Carrasco Limeros, Sylwia Majchrowska, Joakim Johnander, Christoffer Petersson, Miguel Ángel Sotelo, David Fernández Llorca

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.AI

Comments 16 pages, 7 figures, 6 tables

Journal ref CAAI Transactions on Intelligence Technology 24686557 (ISSN) 24682322 (eISSN) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10511 2023-08-08 cs.CL 79%

General Debiasing for Multimodal Sentiment Analysis

Teng Sun, Juntong Ni, Wenjie Wang, Liqiang Jing, Yinwei Wei, Liqiang Nie

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.CL

Comments Accepted at ACM MM 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.02883 2023-08-08 cs.CV 79%

Cross-modal & Cross-domain Learning for Unsupervised LiDAR Semantic Segmentation

Yiyang Chen, Shanshan Zhao, Changxing Ding, Liyao Tang, Chaoyue Wang, Dacheng Tao

专题命中 多模态评测 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.00228 2023-08-02 cs.CV 79%

Using Scene and Semantic Features for Multi-modal Emotion Recognition

Zhifeng Wang, Ramesh Sankaranarayana

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13933 2023-08-02 cs.CV 79%

AIDE: A Vision-Driven Multi-View, Multi-Modal, Multi-Tasking Dataset for Assistive Driving Perception

Dingkang Yang, Shuai Huang, Zhi Xu, Zhenpeng Li, Shunli Wang, Mingcheng Li, Yuzheng Wang, Yang Liu, Kun Yang, Zhaoyu Chen, Yan Wang, Jing Liu, Peixuan Zhang, Peng Zhai, Lihua Zhang

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16366 2023-08-01 eess.IV cs.CV 79%

Multi-modal Graph Neural Network for Early Diagnosis of Alzheimer's Disease from sMRI and PET Scans

Yanteng Zhanga, Xiaohai He, Yi Hao Chan, Qizhi Teng, Jagath C. Rajapakse

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments 14 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15554 2023-07-31 cs.CL 79%

'What are you referring to?' Evaluating the Ability of Multi-Modal Dialogue Models to Process Clarificational Exchanges

Javier Chiyah-Garcia, Alessandro Suglia, Arash Eshghi, Helen Hastie

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CL

Comments Accepted at SIGDIAL'23 (upcoming). Repository with code and experiments available at https://github.com/JChiyah/what-are-you-referring-to

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13706 2023-07-27 cs.HC cs.AI 79%

Introducing CALMED: Multimodal Annotated Dataset for Emotion Detection in Children with Autism

Annanda Sousa, Karen Young, Mathieu D'aquin, Manel Zarrouk, Jennifer Holloway

专题命中 多模态评测 :multimodal(title,abstract);分类 cs.AI

Journal ref HCII 2023: Universal Access in Human-Computer Interaction, Margherita Antona; Constantine Stephanidis, Jul 2023, Copenhagen, Denmark. pp.657-677

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12067 2023-07-25 cs.CV 79%

Replay: Multi-modal Multi-view Acted Videos for Casual Holography

Roman Shapovalov, Yanir Kleiman, Ignacio Rocco, David Novotny, Andrea Vedaldi, Changan Chen, Filippos Kokkinos, Ben Graham, Natalia Neverova

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted for ICCV 2023. Roman, Yanir, and Ignacio contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.00255 2023-07-21 eess.IV cs.CV 79%

Cascaded Multi-Modal Mixing Transformers for Alzheimer's Disease Classification with Incomplete Data

Linfeng Liu, Siyu Liu, Lu Zhang, Xuan Vinh To, Fatima Nasrallah, Shekhar S. Chandra

专题命中 多模态评测 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏