The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue
图像重建游戏:通过迭代多模态对话建立共同基础
Sherzod Hakimov, Mattia D'Agostini, Ivan Samodelkin, David Schlangen
机构
*
Computational Linguistics, Department of Linguistics University of Potsdam(波恩大学语言学系计算语言学部)
;
German Research Center for Artificial Intelligence (DFKI), Berlin(德国人工智能研究中心(DFKI)柏林)
A multimodal dataset of photoplethysmography and continuous behavioral responses to ASMR and nature videos
光电容积描记术和ASMR及自然视频连续行为反应的多模态数据集
Tushar Das, Daigo Hozaki, Koushlendra Kumar Singh, Hirohito M. Kondo
机构
*
Machine Vision & Intelligence Lab, National Institute of Technology Jamshedpur(机器视觉与智能实验室,jamshedpur国家理工学院)
;
School of Psychology, Chukyo University(心理学系,chukyo大学)
FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes
FigSIM:用于自杀迷因的细粒度自杀严重程度和比喻语言数据集
Liuliu Chen, Elise R. Carrotte, Brian E. Chapman, Jo Robinson, Mike Conway
机构
*
School of Computing and Information Systems, University of Melbourne, Australia(墨尔本大学计算与信息学院)
;
Orygen, The National Centre of Excellence in Youth Mental Health, Australia(奥里根青少年心理健康国家研究中心)
;
Centre for Youth Mental Health, University of Melbourne, Australia(墨尔本大学青少年心理健康中心)
;
O’Donnell School of Public Health, UT Southwestern Medical Center, United States(奥唐奈公共卫生学院,西南医学中心)
机构
*
National University of Singapore(国立新加坡大学)
;
Yunnan University(云南大学)
;
The Ohio State University(俄亥俄州立大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
What to Format and How: A Benchmark and Workflow Approach for Document Formatting
格式化什么以及如何格式化:文档格式化的基准与工作流方法
Shihao Rao, Liang Li, Jiapeng Liu, Tong Lin, Bing Li, Xiyan Gao, Peng Fu, Jing Huang, Can Ma
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(信息工程研究所,中国科学院)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
FAIR^2 Drones: An AI-Ready Standard for Cross-Domain Wildlife Drone Datasets
FAIR^2 Drones:跨领域野生动物无人机数据集的AI就绪标准
Jenna Kline, Kilian Meier, Vandita Shukla, Edouard G. A. Rolland, Elena Iannino, Lucie Laporte-Devylder, Constanza Andrea Molina Catricheo, Blair Costelloe, Elizabeth Campolongo, Henrik S. Midtiby, Devis Tuia, Benjamin Risse, Ulrik P. S. Lundquist, Anders Lyhne Christensen, Fabio Remondino, Thomas Richardson, Tanya Berger-Wolf
机构
*
The Ohio State University, Department of Computer Science and Engineering(俄亥俄州立大学计算机科学与工程系)
;
School of Civil, Aerospace and Design Engineering, University of Bristol(布里斯托尔大学土木、航空航天与设计工程学院)
;
D Optical Metrology (3DOM), Fondazione Bruno Kessler (FBK)(3DOM光学计量(3DOM),布鲁诺·克塞勒基金会(FBK))
;
Computer Vision and Machine Learning Systems Group, Institute for Geoinformatics, University of Muenster(计算机视觉与机器学习系统组,地理信息研究所,穆恩斯特大学)
;
Unmanned Aerial Systems Center, University of Southern Denmark(无人飞行系统中心,南部丹麦大学)
;
Department of Collective Behavior, Max Planck Institute of Animal Behavior(集体行为部门,动物行为马克斯·普朗克研究所)
;
Department of Biology, University of Konstanz(生物学系,康斯坦茨大学)
;
Department of Biology, University of Southern Denmark(生物学系,南部丹麦大学)
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction
WorldMemArena: 通过动作-世界交互评估多模态智能体记忆
Chengzhi Liu, Yuzhe Yang, Sophia Xiao Pu, Yepeng Liu, Lin Long, Yichen Guo, Nuo Chen, Zhaotian Weng, Elena Kochkina, Simerjot Kaur, Charese Smiley, Xiaomo Liu, James Zou, Sheng Liu, Yuheng Bu, Songyou Peng, Xin Eric Wang
机构
*
University of California, Santa Barbara(加州大学圣芭芭拉分校)
;
J.P. Morgan Chase(摩根大通)
;
ETH Zurich(苏黎世联邦理工学院)
;
Stanford University(斯坦福大学)
;
Johns Hopkins University(约翰霍普金斯大学)
;
Carnegie Mellon University(卡内基梅隆大学)