A Multimodal Symphony: Integrating Taste and Sound through Generative AI
Matteo Spanio, Massimiliano Zampini, Antonio Rodà, Franco Pierucci
机构
*
Centro di Sonologia Computazionale (CSC)(计算声学中心)
;
Department of Information Engineering University of Padova(信息工程系帕多瓦大学)
;
Center for Mind/Brain Sciences (CIMeC)(心智/大脑科学中心)
;
University of Trento(特伦托大学)
;
SoundFood s.r.l.(SoundFood公司)
CommentsThis paper is accepted to Interspeech 2025. This publication is part of the project Responsible AI for Voice Diagnostics (RAIVD) with file number NGF.1607.22.013 of the research programme NGF AiNed Fellowship Grants which is financed by the Dutch Research Council (NWO)
机构
*
Zhejiang Normal University(浙江师范大学)
;
Hong Kong University of Science and Technology(香港理工大学)
;
Zhejiang University(浙江大学)
;
Tencent(腾讯)
;
Soochow University(苏州大学)
机构
*
College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院)
;
School of Artificial Intelligence, Shenzhen University(深圳大学人工智能学院)
;
Guangdong Provincial Key Laboratory of Intelligent Information Processing(广东省智能信息处理重点实验室)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Renmin University of China(中国人民大学)
;
National Yang Ming Chiao Tung University
;
Taipei Veterans General Hospital(台北荣民总医院)
;
School of Biomedical Engineering, Shenzhen University(深圳大学生物医学工程学院)
机构
*
Faculty of Computing, Harbin Institute of Technology(计算机学院,哈尔滨工业大学)
;
Department of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术系,哈尔滨工业大学)
;
Harbin Institute of Technology Suzhou Research Institute(哈尔滨工业大学苏州研究院长)
;
Peng Cheng Laboratory, Shenzhen, China(鹏城实验室,深圳,中国)
机构
*
Lanzhou University \& Westlake University Lanzhou China
;
Xi'an Jiaotong University Xi'an China
;
Westlake University \& Chinese University of Hong Kong University Hang Zhou, China
;
Southern University of Science
;
Lanzhou University Lanzhou China
;
University of Amsterdam Amsterdam Netherland
;
Westlake University Hangzhou China
;
Lanzhou University \& Westlake University
;
Xi'an Jiaotong University
;
Westlake University \& Chinese University of Hong Kong University
;
Lanzhou University
;
University of Amsterdam
;
Westlake University
SYNBUILD-3D: A large, multi-modal, and semantically rich synthetic dataset of 3D building models at Level of Detail 4
Kevin Mayer, Alex Vesel, Xinyi Zhao, Martin Fischer
机构
*
Department of Civil and Environmental Engineering, Stanford University(土木与环境工程系,斯坦福大学)
;
Department of Computer Science, Stanford University(计算机科学系,斯坦福大学)
Integrating Pathology and CT Imaging for Personalized Recurrence Risk Prediction in Renal Cancer
Daniël Boeke, Cedrik Blommestijn, Rebecca N. Wray, Kalina Chupetlovska, Shangqi Gao, Zeyu Gao, Regina G. H. Beets-Tan, Mireia Crispin-Ortuzar, James O. Jones, Wilson Silva, Ines P. Machado
机构
*
Department of Radiology, Antoni van Leeuwenhoek-Netherlands Cancer Institute, Amsterdam, The Netherlands(放射科,Antoni van Leeuwenhoek荷兰癌症研究所,阿姆斯特丹,荷兰)
;
University of Amsterdam(阿姆斯特丹大学)
;
AI Technology for Life, Department of Information and Computing Sciences, Department of Biology, Utrecht University(AI技术生命,信息与计算科学系,生物学系,乌得勒支大学)
;
GROW Oncology, Maastricht University(GROW肿瘤学,马斯特里赫特大学)
;
Department of Oncology, University of Cambridge(肿瘤科,剑桥大学)
;
Cancer Research UK Cambridge Centre, University of Cambridge(英国癌症研究会剑桥中心,剑桥大学)
;
Early Cancer Institute, University of Cambridge(早期癌症研究所,剑桥大学)
;
Cambridge University Hospitals NHS Foundation Trust(剑桥大学医院 NHS 基础信托)
Comments12 pages, 2 figures, 1 table. Accepted at the Multimodal Learning and Fusion Across Scales for Clinical Decision Support (ML-CDS) Workshop, MICCAI 2025. This is the submitted version with authors, affiliations, and acknowledgements included; it has not undergone peer review or revisions. The final version will appear in the Springer Lecture Notes in Computer Science (LNCS) proceedings