Towards SFW sampling for diffusion models via external conditioning
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepcted at IJCNN 2025
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepcted at IJCNN 2025
机构 * Samsung R&D Institute China–Beijing(三星中国北京研发中心) ; School of Mathematics, Southwestern University of Finance and Economics(西南财经大学数学学院) ; State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统国家重点实验室) ; Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所复杂系统认知与决策智能实验室) ; Beijing Academy of Artificial Intelligence(北京人工智能研究院)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments 16 pages, 4 figures
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Journal ref ICML 2025
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Tsinghua University(清华大学) ; Duke University(杜克大学) ; Peking University(北京大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Department of Mathematics and Computing Science, Saint Mary’s University(数学与计算科学系,圣玛丽大学) ; School of Health Policy and Management, York University(健康政策与管理学院,约克大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Submitted to PLOS Digital Health, Revision 1
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments 30 Pages, 3 figures, 1 table
机构 * School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院) ; State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University(清华大学智能绿色车辆与移动系统国家重点实验室) ; School of Instrumentation and Optoelectronic Engineering, BeiHang University(北航仪器与光电工程学院) ; Department of Automation, Tsinghua University(清华大学自动化系)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
机构 * Aerospace Information Research Institute, Chinese Academy of Sciences, Beijing 100094, China(中国科学院 aerospace information research institute, 北京) ; School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences, Beijing 100049, China(中国科学院大学电子电气与通信工程学院, 北京) ; School of Mathematics, Southeast University, Nanjing 210096, China(东南大学数学学院, 南京) ; Graduate School of Frontier Sciences, the University of Tokyo, Chiba 277-8561, Japan(东京大学前沿科学研究生院, 日本) ; RIKEN Center for Advanced Intelligence Project, Tokyo 103-0027, Japan(RIKEN 高度智能项目中心, 日本)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Xidian University(西安电子科技大学) ; Griffith University(格里菲斯大学) ; Yunnan University(云南大学) ; Carnegie Mellon University(卡内基梅隆大学) ; Zhejiang University(浙江大学) ; Augusta University(奥古斯塔大学) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Accepted by the 34th International Joint Conference on Artificial Intelligence (IJCAI 2025)
机构 * Carnegie Mellon University(卡内基梅隆大学) ; Brown University(布朗大学)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.AI
Comments The first three authors contributed equally to this work. Published at ICLR 2025
机构 * Xiaobo Jin School of Information Engineering Taizhou Vocational College of Science & Technology(金晓波 学校信息工程学院 太zhou 职业科技学院) ; Traditional Chinese Medicine department Taizhou First People’s Hospital(传统中医部门 太zhou 第一人民医院)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments 14 pages,2 figures,4 tables
机构 * ReLER, CCAI, Zhejiang University(ReLER、CCAI、浙江大学) ; DBMI, HMS, Harvard University(DBMI、HMS、哈佛大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * University of Science and Technology of China(中国科学技术大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Zhejiang University(浙江大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepted at ICLR 2025
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Harbin Institute of Technology(哈尔滨工业大学) ; Li Auto Inc.(Li汽车公司) ; Tsinghua University(清华大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Northeastern University(东北大学) ; Meta GenAI(Meta 生成人工智能) ; Meta FAIR ; National University of Singapore(国立新加坡大学) ; The Chinese University of Hong Kong(香港中文大学) ; University of Washington(华盛顿大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Project Page: https://ma-xu.github.io/token-shuffle/ Add related works
机构 * USTC(中国科学技术大学) ; SJTU(上海交通大学) ; PKUSZ(北京大学软件学院) ; Tencent(腾讯) ; BFA(北京航空航天大学)
专题命中 多模态生成 :MLLM(abstract);分类 cs.CV
Comments project:https://github.com/AILab-CVC/VideoGen-Eval
专题命中 多模态生成 :audio-visual(abstract);分类 cs.MM
机构 * Apple(苹果公司)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
机构 * University of Cambridge(剑桥大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Published in Artificial Intelligence Review
Journal ref Artif Intell Rev 58, 214 (2025)
机构 * CUHK MMLab(香港中文大学多模态实验室) ; KAUST(科威特科学与技术研究中心) ; Hugging Face(Hugging Face公司) ; Shanghai AI Lab(上海人工智能实验室)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments All code, checkpoints, and datasets are available at \url{https://diffusion-cot.github.io/reflection2perfection}
机构 * Zhejiang University(浙江大学) ; Harvard University(哈佛大学) ; Nanyang Technological University(南洋理工大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Tsinghua University(清华大学) ; Kuaishou Technology(快手科技) ; CASIA(中国科学院自动化研究所)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments 8 pages, 7 figures
机构 * School of Systems and Computing University of New South Wales(系统与计算学院 新南威尔士大学) ; School of Engineering Australian National University(工程学院 澳大利亚国立大学) ; School of Business University of New South Wales(商学院 新南威尔士大学) ; Australian Institute for Machine Learning University of Adelaide(机器学习研究所 阿德莱德大学) ; School of Computer Science and Engineering University of New South Wales(计算机科学与工程学院 新南威尔士大学) ; School of Computer Science University of Technology Sydney(计算机科学学院 技术大学悉尼)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments accepted by IJCV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Project page: https://github.com/KaiyueSun98/T2I-Personalization-with-AR
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Project page: https://wuyan01.github.io/uniphys-project/
专题命中 多模态生成 :cross-modal(abstract);分类 cs.CV