VideoWeaver: Multimodal Multi-View Video-to-Video Transfer for Embodied Agents
VideoWeaver:多模态多视角视频到视频转换用于具身智能体
George Eskandar, Fengyi Shen, Mohammad Altillawi, Dong Chen, Yang Bai, Liudi Yang, Ziyuan Liu
机构
*
Huawei Heisenberg Research Center(华为海森堡研究中心)
;
Ludwig Maximilian University of Munich(慕尼黑大学)
;
University of Freiburg(弗莱堡大学)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
DyMRL: Dynamic Multispace Representation Learning for Multimodal Event Forecasting in Knowledge Graph
DyMRL: 动态多空间表征学习用于知识图谱中的多模态事件预测
Feng Zhao, Kangzheng Liu, Teng Peng, Yu Yang, Guandong Xu
机构
*
Huazhong University of Science and Technology(华中科技大学)
;
Centre for Learning, Teaching and Technology(学习、教学与技术中心)
;
The Education University of Hong Kong(香港教育大学)
专题命中
视频多模态
:multimodal(title,abstract);分类 cs.AI
AI总结
DyMRL通过动态多空间表征学习,解决多模态知识动态获取与融合问题,提升事件预测性能。
CommentsAccepted to The ACM Web Conference 2026 (WWW '26). This version is published under a CC BY license
机构
*
Department of Computer Science and Technology(计算机科学与技术系)
;
University of Southern California(南加州大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
State Key Laboratory of Internet Architecture(互联网架构国家重点实验室)
CommentsarXiv admin comment: This version has been removed by arXiv administrators as the submitter did not have the rights to agree to the license at the time of submission
MSEG-VCUQ: Multimodal SEGmentation with Enhanced Vision Foundation Models, Convolutional Neural Networks, and Uncertainty Quantification for High-Speed Video Phase Detection Data
Spatiotemporal Heterogeneity of AI-Driven Traffic Flow Patterns and Land Use Interaction: A GeoAI-Based Analysis of Multimodal Urban Mobility
人工智能驱动交通流模式与土地利用相互作用的时空异质性:基于GeoAI的多模式城市交通分析
Olaf Yunus Laitinen Imanov
机构
*
Department of Applied Mathematics(应用数学系)
;
Computer Science (DTU Compute), Technical University of Denmark, Kongens Lyngby, Denmark(计算机科学(DTU Compute),丹麦技术大学,Kongens Lyngby)
Recognition of Daily Activities through Multi-Modal Deep Learning: A Video, Pose, and Object-Aware Approach for Ambient Assisted Living
通过多模态深度学习识别日常活动:一种面向环境辅助养老的视频、姿态和物体感知方法
Kooshan Hashemifard, Pau Climent-Pérez, Francisco Florez-Revuelta
机构
*
Alicante Institute for Health and Biomedical Research (ISABIAL)(阿利坎特健康与生物医学研究所(ISABIAL))
;
ValgrAI - Valencian Graduate School and Research Network of Artificial Intelligence(瓦尔格AI——瓦伦西亚人工智能研究生院与研究网络)
Zhibin Lan, Liqiang Niu, Fandong Meng, Jie Zhou, Jinsong Su
机构
*
School of Informatics, Xiamen University, China(厦门大学信息学院)
;
WeChat AI, Tencent Inc, China(腾讯公司微信AI部门)
;
Key Laboratory of Digital Protection and Intelligent Processing of Intangible Cultural Heritage of Fujian and Taiwan (Xiamen University), Ministry of Culture and Tourism, China(福建省和台湾非物质文化遗产数字化保护与智能处理重点实验室)
;
Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室)
Enhancing Vision-Language Navigation with Multimodal Event Knowledge from Real-World Indoor Tour Videos
通过真实世界室内游览视频的多模态事件知识增强视觉语言导航
Haoxuan Xu, Tianfu Li, Wenbo Chen, Yi Liu, Xingxing Zuo, Yaoxian Song, Haoang Li
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Tsinghua University(清华大学)
;
Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI)(马尔代夫 bin Zayed 大学人工智能学院)
;
Hangzhou City University(杭州城市学院)