Listening with the Eyes: Benchmarking Egocentric Co-Speech Grounding across Space and Time
用眼睛倾听:跨时空的自体视觉共指基准测试
Weijie Zhou, Xuantang Xiong, Zhenlin Hu, Xiaomeng Zhu, Chaoyang Zhao, Honghui Dong, Zhengyou Zhang, Ming Tang, Jinqiao Wang
机构
*
Beijing Jiaotong University(北京交通大学)
;
Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences (CASIA)(基础模型研究中心、自动化研究所、中国科学院(CASIA))
;
Tencent Robotics X(腾讯机器人X)
;
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology (HKUST)(计算机科学与工程系、香港科学与技术大学(HKUST))
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学深圳学院)
PhysLLM: Harnessing Large Language Models for Cross-Modal Remote Physiological Sensing
PhysLLM:利用大语言模型进行跨模态远程生理传感
Yiping Xie, Bo Zhao, Mingtong Dai, Jian-Ping Zhou, Yue Sun, Tao Tan, Weicheng Xie, Linlin Shen, Zitong Yu
机构
*
Shenzhen University(深圳大学)
;
Great Bay University(大鹏大学)
;
National Engineering Laboratory for Big Data System Computing Technology(大数据系统计算技术国家工程实验室)
;
Dongguan Key Laboratory for Intelligence and Information Technology(东莞智能与信息技术重点实验室)
;
Guangdong Medical University(广东医科大学)
;
Southern Medical University (Dongguan People’s Hospital)(南方医科大学(东莞人民医院))
;
Macao Polytechnic University(澳门理工学院)
Bloom: Designing for LLM-Augmented Behavior Change Interactions
Bloom:为LLM增强的行为改变交互设计
Matthew Jörke, Defne Genç, Valentin Teutschbein, Shardul Sapkota, Sarah Chung, Paul Schmiedmayer, Maria Ines Campero, Abby C. King, Emma Brunskill, James A. Landay
(hu)Man vs. Machine: In the Future of Motorsport, can Autonomous Vehicles Compete?
人与机器:在赛车未来中,自动驾驶车辆能竞争吗?
Armand Amaritei, Amber-Lily Blackadder, Sebastian Donnelly, Lora Hernandez, James Vine, Alexander Rast, Matthias Rolf, Andrew Bradley
机构
*
Autonomous Driving and Intelligent Transport Group, Oxford Brookes University, UK(自主驾驶与智能交通组,奥克斯伯里大学,英国)
;
School of Engineering, Computing & Mathematics, Oxford Brookes University, UK(工程、计算与数学学院,奥克斯伯里大学,英国)
;
Artificial Intelligence, Data Analysis and Systems Institute, Oxford Brookes University, UK(人工智能、数据分析与系统研究所,奥克斯伯里大学,英国)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Institute of Digital Twin, Eastern Institute of Technology(数字孪生研究院,东部技术研究所)
;
Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative(宁波空间智能与数字衍生关键实验室)
Can Unified Generation and Understanding Models Maintain Semantic Equivalence Across Different Output Modalities?
统一的生成与理解模型能否在不同输出模态间保持语义等价性?
Hongbo Jiang, Jie Li, Yunhang Shen, Pingyang Dai, Xing Sun, Haoyu Cao, Liujuan Cao
机构
*
Tencent Youtu Lab(腾讯云图实验室)
;
Xiamen University(厦门大学)
;
Computing Lab, Department of Artificial Intelligence, School of Informatics(计算实验室,人工智能系,信息学院)
MediX-R1: Open Ended Medical Reinforcement Learning
MediX-R1:开放端医疗强化学习
Sahal Shaji Mullappilly, Mohammed Irfan Kurpath, Omair Mohamed, Mohamed Zidan, Fahad Khan, Salman Khan, Rao Anwer, Hisham Cholakkal
机构
*
Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI)(迈赫迈德·本·扎耶德人工智能大学)
;
Jubilee Mission Medical College(jubilee mission 医学院)
;
Research Institute(研究院)
;
JJM Medical College(JJM 医学院)
机构
*
Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou)(香港理工大学(广州)人工智能研究所)
;
Hubei Key Laboratory of Inland Shipping Technology (Wuhan University of Technology)(湖北内河航运技术重点实验室(武汉理工大学))
;
School of Navigation, Wuhan University of Technology(武汉理工大学航海学院)
;
School of Advanced Technology, Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学先进科技学院)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
School of Information Engineering, Yancheng Institute of Technology(盐城职业技术学院信息工程学院)
;
School of Engineering, Stanford University(斯坦福大学工程学院)
;
Centre for AI and Data Science Innovation and the School of Science and Engineering, James Cook University(詹姆斯库克大学人工智能与数据科学创新中心及科学与工程学院)
AR&D: A Framework for Retrieving and Describing Concepts for Interpreting AudioLLMs
AR&D: 一种用于音频大语言模型解释的检索与描述框架
Townim Faisal Chowdhury, Ta Duc Huy, Siqi Pan, Jeremy Stoddard, Zhibin Liao
机构
*
Australian Institute for Machine Learning, University of Adelaide(澳大利亚机器学习研究所,阿德莱德大学)
;
Dolby Laboratories(杜比实验室)
;
School of Computer and Mathematical Sciences, University of Adelaide, Australia(计算机与数学科学学院,阿德莱德大学,澳大利亚)
机构
*
School of Big Data & Software Engineering, Chongqing University(大数据与软件工程学院,重庆大学)
;
School of Computer Science & Technology, Chongqing University(计算机科学与技术学院,重庆大学)
Do Large Language Models Understand Data Visualization Rules?
大语言模型理解数据可视化规则吗?
Martin Sinnona, Valentin Bonas, Emmanuel Iarussi, Viviana Siless
机构
*
Universidad Torcuato Di Tella(托科托迪特拉大学)
;
Consejo Nacional de Investigaciones Científicas y Técnicas(国家科学与技术研究院)
;
Universidad de Buenos Aires(布宜诺斯艾利斯大学)