SDGOCC: Semantic and Depth-Guided Bird's-Eye View Transformation for 3D Multimodal Occupancy Prediction
机构 * Huazhong University of Science and Technology(华中科技大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments accepted by CVPR2025
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Huazhong University of Science and Technology(华中科技大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments accepted by CVPR2025
机构 * Hong Kong Baptist University(香港 Baptist 大学) ; Shanxi University(山西大学) ; Shanghai Institute for Advanced Study of Zhejiang University(浙江大学上海先进研究院) ; Hong Kong University of Science and Technology(香港科技大学) ; The University of Manchester(曼彻斯特大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CL、cs.AI
Comments 13 pages. 7 figures
Journal ref This paper is accpeted by ACL2025(Main)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CL、cs.AI
Comments Preprint
机构 * ByteDance(字节跳动)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.MM
Comments 19 pages, 5 figures
机构 * AppliedML, Cerebras(应用机器学习,Cerebras)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * University of Massachusetts, Amherst(马萨诸塞大学阿默斯特分校) ; Massachusetts Institute of Technology(麻省理工学院)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments Project page: https://vlm-mirage.github.io/
机构 * College of Computer and Mathematics, Central South University of Forestry and Technology(计算机与数学学院,中央南大学林业科技学院) ; Department of Computer Science, Anhui Normal University(计算机科学系,安徽师范大学) ; Future Technology Institute, South China University of Technology(未来技术研究院,华南理工大学) ; Department of Computer Science, State University of New York(计算机科学系,纽约州立大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CL、eess.AS
机构 * The Pennsylvania State University(宾夕法尼亚州立大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments 18 pages, 10 figures; Accepted to ACL 2025 Findings
机构 * Machine Learning Center, Georgia Institute of Technology, Atlanta, GA ; School of Mathematics, Georgia Institute of Technology, Atlanta, GA ; Department of Mathematics \& Statistics, SUNY Albany, NY
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments Accepted to ICML 2025. Code available at https://github.com/KevinRojas1499/Diffuse-Everything
机构 * Department of Computer Science, University of Maryland, Baltimore County(计算机科学系,马里兰大学巴尔的摩县分校) ; Department of Dermatology, Johns Hopkins University School of Medicine(皮肤科系,约翰霍普金斯大学医学院)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments Accepted at IEEE/CVF Computer Society Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)
机构 * Intel Labs(英特尔实验室) ; Amazon(亚马逊)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL
Comments ICML 2025 Spotlight Oral
机构 * Shanghai Jiao Tong University(上海交通大学) ; Tencent Hunyuan(腾讯文英) ; Zhejiang University(浙江大学)
专题命中 多模态生成 :cross-modal(title);MLLM(abstract);分类 cs.CV、cs.AI
机构 * School of Automation(自动化学院) ; Beijing Institute of Technology(北京理工大学)
专题命中 多模态生成 :multimodal(title);MLLM(abstract);分类 cs.CV、cs.AI
Comments 16 pages, 11 figures
机构 * Nankai University(南开大学) ; Shanghai Innovation Institute(上海创新研究院) ; University of Science and Technology of China(中国科学技术大学) ; Wuhan University(武汉大学) ; Shanghai AI Laboratory(上海人工智能实验室)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * College of Science and Engineering(科学与工程学院) ; James Cook University(詹姆斯库克大学) ; School of Computing(计算学院)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * UCL(伦敦大学学院) ; Technical University of Munich(慕尼黑技术大学) ; University of Oxford(牛津大学) ; University of Glasgow(格拉斯哥大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments Preprint, manuscript under-review
机构 * Duke University(杜克大学) ; University of California, San Diego(加州大学圣迭戈分校) ; MBZUAI(马克斯·普朗克人工智能研究所) ; MIT(麻省理工学院) ; University of Michigan(密歇根大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * Shanghai Key Lab of Intell. Info. Processing, Fudan University(上海智能信息处理关键实验室,复旦大学) ; Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心) ; School of Computer Science and Technology, East China Normal University(东华大学计算机科学与技术学院)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * School of Computer Science and Engineering(计算机科学与工程学院) ; South China University of Technology(华南理工大学) ; School of Future Technology(未来技术学院) ; University of Oxford(牛津大学) ; Pengcheng Laboratory(鹏城实验室) ; Department of Computer Science(计算机科学系)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * Department of Computer Science and Engineering, Kyung Hee University(计算机科学与工程系,庆熙大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL
Comments 10 pages, 2 figures
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CL、cs.AI
机构 * Shanghai Jiao Tong University(上海交通大学) ; Huawei(华为) ; Tongji University(同济大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) ; Pengcheng Laboratory(鹏城实验室) ; MAIS, Institute of Automation, Chinese Academy of Sciences(自动化所MAIS部) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) ; University of Chinese Academy of Sciences(中国科学院大学) ; ByteDance Inc.(字节跳动公司) ; National Cheng Kung University(国立成功大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
机构 * Gulf Coast Research and Education Center (GCREC), University of Florida(佛罗里达大学 Gulf Coast Research and Education Center) ; Citrus Research and Education Center (CREC), University of Florida(佛罗里达大学 Citrus Research and Education Center)
专题命中 多模态生成 :multi-modal(title,abstract);分类 cs.CV、cs.AI
机构 * Dept. of Electrical and Computer Engineering, Clemson University(电子与计算机工程系,克莱姆森大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments 7 pages, 4 figures, 4 tables
机构 * Qatar Computing Research Institute(卡塔尔计算研究所) ; Hamad Bin Khalifa University(哈马德·本·卡尔法大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments 20 pages
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments Short Paper The Web Conference
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments 12 pages, 11 figures, 13IHMMSec2025