DyFuLM: An Advanced Multimodal Framework for Sentiment Analysis
DyFuLM:一种用于情感分析的先进多模态框架
Ruohan Zhou, Jiachen Yuan, Churui Yang, Wenzheng Huang, Guoyan Zhang, Shiyao Wei, Jiazhen Hu, Ning Xin, Md Maruf Hasan
机构
*
Department of Applied Mathematics, Xi'an Jiaotong-Liverpool University(应用数学系,西安交通大学-利物浦大学)
;
School of AI and Advanced Computing, XJTLU Entrepreneur College (Taicang)(人工智能与先进计算学院,XJTLU创业学院(太仓))
;
Department of Intelligent Science, Xi'an Jiaotong-Liverpool University(智能科学系,西安交通大学-利物浦大学)
HMR3D: Hierarchical Multimodal Representation for 3D Scene Understanding with Large Vision-Language Model
HMR3D:用于大视觉-语言模型的层次多模态表示以实现3D场景理解
Chen Li, Eric Peh, Basura Fernando
机构
*
Institute of High-Performance Computing, Agency for Science, Technology and Research, Singapore(高性能计算研究所,科技研究局,新加坡)
;
Centre for Frontier AI Research, Agency for Science, Technology and Research, Singapore(前沿人工智能研究中心,科技研究局,新加坡)
;
College of Computing and Data Science, Nanyang Technological University, Singapore(计算与数据科学学院,南洋理工大学,新加坡)
机构
*
Department of Clinical Neurosciences, University of Cambridge, UK(剑桥大学临床神经科学系)
;
Department of Electrical & Electronic Engineering, The University of Manchester, UK(曼彻斯特大学电气与电子工程系)
;
Department of Pathology, State Key Laboratory of Oncology in South China, Guangdong Provincial Clinical Research Center for Cancer, Sun Yat-sen University Cancer Center, China(南方医科大学肿瘤学国家重点实验室、广东省癌症临床研究中心、中山大学肿瘤中心病理学部)
;
Department of Statistics and Actuarial Science, The University of Hong Kong, Hong Kong SAR, China(香港大学统计与精算科学系)
;
Department of Clinical Neurosciences and Department of Applied Mathematics and Theoretical Physics, University of Cambridge(剑桥大学临床神经科学系和应用数学与理论物理系;邓迪大学科学与工程学院和医学学院)
;
School of Science and Engineering and School of Medicine, University of Dundee, UK
TinyRS-R1: Compact Multimodal Language Model for Remote Sensing
TinyRS-R1:用于遥感的紧凑多模态语言模型
Aybora Koksal, A. Aydin Alatan
机构
*
Center for the Image Analysis (OGAM) and Department of Electrical and Electronics Engineering of Middle East Technical University (METU)(图像分析中心(OGAM)和中欧技术大学(METU)电子与电气工程系)
CommentsAccepted to IEEE Geoscience and Remote Sensing Letters (GRSL). Code, models, and the captions for datasets are available at https://github.com/aybora/TinyRS
机构
*
CVIU Lab, University of Arkansas, USA(大学实验室,亚利桑那大学,美国)
;
University of Florida, USA(佛罗里达大学,美国)
;
COSMOS Research Center, University of Arkansas, Little Rock, USA(研究机构,亚利桑那大学,小石城,美国)
;
ICSI, University of California, Berkeley, USA(研究机构,加州大学伯克利分校,美国)
A Theory-Inspired Framework for Few-Shot Cross-Modal Sketch Person Re-Identification
基于理论的少样本跨模态素描人重识别框架
Yunpeng Gong, Yongjie Hou, Jiangming Shi, Kim Long Diep, Min Jiang
机构
*
School of Informatics, Xiamen University(厦门大学信息学院)
;
School of Electronic Science and Engineering, Xiamen University(厦门大学电子科学与技术学院)
;
Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究院)
;
Department of Artificial Intelligence, Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, School of Informatics, Key Laboratory of Digital Protection and Intelligent Processing of Intangible CulturalHeritage of Fujian and Taiwan, Ministry of Culture and Tourism, Xiamen University(人工智能系,教育部多媒体可信感知与高效计算重点实验室,厦门大学信息学院,福建省和台湾非物质文化遗产数字化保护与智能处理重点实验室,文化和旅游部,厦门大学)
机构
*
School of Computer Science, Hangzhou Dianzi University(杭州电子科技大学计算机科学学院)
;
CAD & CG, Zhejiang University(浙江大学计算机辅助设计与图形学研究所)
;
School of Computer Science, University of Bristol(布里斯托大学计算机科学学院)
UGG-ReID: Uncertainty-Guided Graph Model for Multi-Modal Object Re-Identification
Xixi Wan, Aihua Zheng, Bo Jiang, Beibei Wang, Chenglong Li, Jin Tang
机构
*
Information Materials and Intelligent Sensing Laboratory of Anhui Province, School of Artificial Intelligence, Anhui University(安徽省信息材料与智能感知实验室,人工智能学院,安徽大学)
;
Anhui Provincial Key Laboratory of Multimodal Cognitive Computation, School of Computer Science and Technology, Anhui University(安徽省多模态认知计算重点实验室,计算机科学与技术学院,安徽大学)
CommentsTo be published in the Proceedings of the 40th Annual AAAI Conference on Artificial Intelligence (AAAI 2026 Special Track on AI for Social Impact )
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
Feiyang Jia, Caiyan Jia, Ailin Liu, Shaoqing Xu, Qiming Xia, Lin Liu, Lei Yang, Yan Gong, Ziying Song
机构
*
School of Computer Science and Technology, Beijing Key Laboratory of Traffic Data Mining and Embodied Intelligence, Beijing Jiaotong University(计算机科学与技术学院、交通数据挖掘与具身智能北京市重点实验室、北京交通大学)
;
State Key Laboratory of Internet of Things for Smart City and Department of Electrome chanical Engineering, University of Macau(智能城市物联网国家重点实验室、澳门大学机电工程系)
;
Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University(智能城市感知与计算福建省重点实验室、厦门大学)
;
School of Mechanical and Aerospace Engineering, Nanyang Technological University(机械与航空航天工程学院、南洋理工大学)
;
State Key Laboratory of Robotics and System, Harbin Institute of Technology(机器人系统国家重点实验室、哈尔滨工业大学)