SAVOR: Skill Affordance Learning from Visuo-Haptic Perception for Robot-Assisted Bite Acquisition
机构 * Cornell University(康奈尔大学) ; UC San Diego(南加州大学)
专题命中 视频多模态 :multi-modal(abstract)
Comments Conference on Robot Learning, Oral
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Cornell University(康奈尔大学) ; UC San Diego(南加州大学)
专题命中 视频多模态 :multi-modal(abstract)
Comments Conference on Robot Learning, Oral
机构 * DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) ; Imperial College London(帝国理工学院) ; Tsinghua University(清华大学) ; Hupan Lab(华潘实验室)
专题命中 跨模态检索 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments Technical Report
机构 * Department of Computer Science, Toronto Metropolitan University(计算机科学系,多伦多 Metropolitan 大学)
专题命中 跨模态检索 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CL
Journal ref Machine Learning and Knowledge Extraction. 2025; 7(3):89
机构 * Vocational School of Social Sciences, Marmara University(马尔马拉大学社会科学职业学校) ; Geospatial Data Science Group, School of Geographical & Earth Sciences, University of Glasgow(格拉斯哥大学地理与地球科学学院空间数据科学小组)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * Department of Computer Science Clarkson University(计算机科学系克拉克逊大学)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments 8 pages, 5 figures, 2 tables
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; The Hong Kong University of Science and Technology(香港科学与技术大学)
专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV、cs.AI
Comments 9 pages
机构 * City University of Hong Kong(香港城市大学) ; Tencent Inc.(腾讯公司) ; Tsinghua University(清华大学)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.AI
Comments CIKM 2025 Full Research Paper
机构 * University of Science, VNU-HCM(越南国家大学科学学院)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV
Comments ACM Multimedia 2025
机构 * School of Astronautics, Beihang University(北京航空航天大学航天学院) ; Department of Computer Science, Dartmouth College(达特茅斯学院计算机科学系)
专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV
机构 * Shanghai Jiao Tong University(上海交通大学) ; Zhejiang University(浙江大学) ; Westlake University(西湖大学)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CL、cs.AI
Comments Our platform is publicly accessible at https://www.tbox.cn/about/model-ranking
机构 * Department of Applied Mathematics, Faculty of Information Technology, Czech Technical University in Prague(应用数学系,信息科技学院,布拉格捷克技术大学)
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.AI
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.CV
Comments 9 pages, 5 figures. Accepted by the 18th International Conference on Knowledge Science, Engineering and Management (KSEM 2025)
机构 * Alibaba Group(阿里巴巴集团) ; Beijing Electronic Science and Technology Institute(北京电子科技研究所) ; Nanjing University(南京大学) ; Renmin University of China(中国人民大学) ; Northeastern University(东北大学) ; BraneMatrix AI ; Nanyang Technological University(南洋理工大学)
专题命中 跨模态检索 :cross-modal(abstract);分类 cs.AI
Comments Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization
机构 * Shanghai AI Laboratory(上海人工智能实验室) ; Peking University(北京大学) ; The University of HongKong(香港大学) ; Shanghai Jiaotong University(上海交通大学) ; Beihang University(北京航空航天大学)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV
Comments Accepted by ICCV 2025
专题命中 跨模态检索 :multimodal(abstract)
机构 * University of Washington(华盛顿大学)
专题命中 跨模态检索 :multimodal(abstract)
Comments Published at ICML '25 (Oral)
机构 * The Hong Kong Polytechnic University(香港理工大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments ACM MM 2025 Camera Ready
机构 * ETH Zürich(苏黎世联邦理工学院) ; DisneyResearch | Studios(迪士尼研究室)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
Comments Added more evaluation since the first version. Accepted to SMI 2025. Computers & Graphics
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
Comments Project Page: https://personavlog-paper.github.io/
专题命中 多模态生成 :multi-modal(title);multimodal(abstract)
专题命中 多模态生成 :multimodal(title)
Comments MICCAI'25
机构 * School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院) ; South China University of Technology(华南理工大学) ; Guangdong Engineering Center for Large Model and GenAI Technology(广东省大模型与生成式人工智能技术工程中心) ; Guangdong Provincial Key Lab of Computational Intelligence and Cyberspace Information(广东省计算智能与网络信息重点实验室) ; Centre for Smart Health, Hong Kong Polytechnic University(香港理工大学智能健康研究中心) ; CAS-Hong Kong Joint Laboratory for Multimodal Medical Molecular Imaging(中国科学院-香港联合多模态医学分子成像联合实验室) ; School of Electronic and Information Engineering, South China University of Technology(华南理工大学电子与信息学院) ; Pazhou Lab, Guangzhou, Guangdong, China(广州琶洲实验室) ; School of Computing and Information Systems, Singapore Management University(新加坡国立大学计算机与信息系统学院)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI、cs.MM
机构 * ONE Lab, HUST(华中科技大学 ONE 实验室) ; ONE Lab, HUST University of Maryland(华中科技大学 与 马里兰大学 ONE 实验室) ; University of Washington(华盛顿大学) ; Zhejiang University(浙江大学)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV、cs.CL
Comments Technical Report
机构 * Technical University of Munich(慕尼黑技术大学) ; Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) ; ETH Zurich(苏黎世联邦理工学院) ; MBZUAI(穆桑比克大学人工智能研究所)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments 13 pages
机构 * Beijing Jiaotong University(北京交通大学) ; Ant Group(蚂蚁集团) ; Qinghai University(青海大学) ; Tsinghua University(清华大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments 14 pages, 11 figures
机构 * Consultant, high-throughput microscopy and hardware-accelerated algorithms(咨询顾问,高通量显微镜和硬件加速算法)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CL
Comments Minor changes: resolve HTML rendering issues of sideways tables; Code listing in dark mode. Cite three more journal articles
机构 * Stanford University(斯坦福大学) ; Adobe Research(Adobe研究)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments ICML 2025. Code: https://github.com/Lakonik/GMFlow
机构 * Department of Computer Science, North Carolina State University(计算机科学系,北卡罗来纳州立大学) ; Department of Computer Science and Engineering, University of Louisville(计算机科学与工程系,路易斯维尔大学) ; Department of Electrical and Computer Engineering and Frost Institute for Data Science and Computing, University of Miami(电气与计算机工程系及弗罗斯特数据科学与计算研究所,迈阿密大学)
专题命中 多模态生成 :multi-modal(abstract)
Comments Accepted by IEEE MobiWac 2025
专题命中 多模态生成 :multi-modal(abstract)
机构 * University of Washington(华盛顿大学) ; UC San Diego(圣地亚哥大学) ; Nvidia(英伟达)
专题命中 多模态生成 :multi-modal(abstract)