机构
*
Indian Institute of Technology, Ropar, India(印度理工学院罗帕尔分校)
;
Machine Intelligence Group, Birla Institute of Technology and Science, Pilani, Hyderabad Campus, India(比拉理工科学院帕利尼 Hyderabad 分校机器智能小组)
;
Monash University, Melbourne, Australia(墨尔本大学)
DA-DPO: Cost-efficient Difficulty-aware Preference Optimization for Reducing MLLM Hallucinations
DA-DPO:面向减少多模态大语言模型幻觉的高效难度感知偏好优化
Longtian Qiu, Shan Ning, Chuyu Zhang, Jiaxuan Sun, Xuming He
机构
*
ShanghaiTech University(上海科技大学)
;
Lingang Laboratory(灵冈实验室)
;
Shanghai Engineering Research Center of Intelligent Vision and Imaging(上海智能视觉与成像工程技术研究中心)
Comments21 pages, 6 figures, 8 tables. Includes ancillary files with full benchmark results and ablation studies. Code available at https://github.com/athrael-soju/Snappy
Agentic TinyML for Intent-aware Handover in 6G Wireless Networks
代理型TinyML用于6G无线网络中的意图感知切换
Alaa Saleh, Roberto Morabito, Sasu Tarkoma, Anders Lindgren, Susanna Pirttikangas, Lauri Lovén
机构
*
Center for Ubiquitous Computing(无处不在计算中心)
;
University of Oulu(奥卢大学)
;
Department of Communication Systems(通信系统系)
;
EURECOM
;
Department of Computer Science(计算机科学系)
;
University of Helsinki(赫尔辛基大学)
;
RISE Research Institutes of Sweden(瑞典研究机构)
;
Department of Computer Science, Electrical and Space Engineering(计算机科学、电气与空间工程系)
;
Luleå University of Technology(卢勒奥技术大学)
机构
*
School of Computing, Mathematics, and Engineering, Charles Sturt University(计算、数学与工程学院,查尔斯·斯特劳特大学)
;
School of Dentistry and Medical Sciences, Charles Sturt University(牙科学院与医学科学学院,查尔斯·斯特劳特大学)
;
AI and Cyber Futures Centre (AICF)(人工智能与未来技术中心(AICF))
;
Hawkins Clinic General Medical Practice(霍金斯诊所普通医疗实践)
S1-MMAlign: A Large-Scale, Multi-Disciplinary Dataset for Scientific Figure-Text Understanding
S1-MMAlign: 一个大规模、跨学科的数据集用于科学图像-文本理解
He Wang, Longteng Guo, Pengkang Huo, Xuanxu Lin, Yichen Yuan, Jie Jiang, Jing Liu
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉学科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
机构
*
University of Hong Kong(香港大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Huawei Research(华为研究)
Benchmarking Preprocessing and Integration Methods in Single-Cell Genomics
单细胞组学中预处理和整合方法的基准测试
Ali Anaissi, Seid Miad Zandavi, Weidong Huang, Junaid Akram, Basem Suleiman, Ali Braytee, Jie Hua
机构
*
University of Technology Sydney, Australia(悉尼科技大学)
;
University of Sydney, Australia(悉尼大学)
;
Broad Institute, United States(Broad研究所)
;
University of New South Wales, Australia(新南威尔士大学)
;
Shaoyang University, China(韶阳大学)
VisualQuest: A Benchmark for Abstract Visual Reasoning in MLLMs
VisualQuest: 一个用于多模态大语言模型(MLLMs)抽象视觉推理的基准数据集
Kelaiti Xiao, Liang Yang, Dongyu Zhang, Paerhati Tulajiang, Hongfei Lin
机构
*
School of Computer Science and Technology, Dalian University of Technology, Dalian, China(大连理工大学计算机科学与技术学院)
;
School of Foreign Languages, Dalian University of Technology, Dalian, China(大连理工大学外语学院)
;
School of Computer Science and Technology, Xinjiang Normal University, Urumqi, China(新疆师范大学计算机科学与技术学院)
Bayesian Inverse Games with High-Dimensional Multi-Modal Observations
高维多模态观测下的贝叶斯逆游戏
Yash Jain, Xinjie Liu, Lasse Peters, David Fridovich-Keil, Ufuk Topcu
机构
*
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Delft University of Technology(代尔夫特理工大学)
;
Sunrise Setting Ltd
;
SAGE Publications Ltd(SAGE出版社有限公司)
Modality Dominance-Aware Optimization for Embodied RGB-Infrared Perception
具身RGB-红外感知的模态主导意识优化
Xianhui Liu, Siqi Jiang, Yi Xie, Yuqing Lin, Siao Liu
机构
*
College of Electronics and Information Engineering(电子信息工程学院)
;
School of Computer Science and Technology(计算机科学与技术学院)
;
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
School of Future Science and Engineering(未来科学与工程学院)
;
Key Laboratory of General Artificial Intelligence and Large Models in Provincial Universities(省属大学通用人工智能与大模型重点实验室)
OmniVaT: Single Domain Generalization for Multimodal Visual-Tactile Learning
OmniVaT:单域泛化用于多模态视觉-触觉学习
Liuxiang Qiu, Hui Da, Yuzhen Niu, Tiesong Zhao, Yang Cao, Zheng-Jun Zha
机构
*
Fujian Key Laboratory for Intelligent Processing and Wireless Transmission of Media Information(福建智能媒体信息处理与无线传输重点实验室)
;
College of Physics and Information Engineering(物理与信息工程学院)
;
Fuzhou University(福州市大学)
;
College of Computer and Data Science(计算机与数据科学学院)
;
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition(MoE脑启发智能感知与认知重点实验室)
;
University of Science and Technology of China(中国科学技术大学)