LatRef-Diff: Latent and Reference-Guided Diffusion for Facial Attribute Editing and Style Manipulation
LatRef-Diff: 基于潜在和参考引导的扩散模型用于面部属性编辑和风格操控
Wenmin Huang, Weiqi Luo, Xiaochun Cao, Jiwu Huang
机构
*
GuangDong Province Key Lab of Information Security Technology and School of Computer Science and Engineering(广东信息安全技术重点实验室和计算机科学与工程学院)
;
School of Cyber Science and Technology, Shenzhen Campus(深圳校区网络科学与技术学院)
;
Guangdong Laboratory of Machine Perception and Intelligent Computing, Faculty of Engineering(广东机器感知与智能计算实验室,工程学院)
VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models
VFM-VAE:视觉基础模型可以作为潜在扩散模型的良好分词器
Tianci Bi, Xiaoyi Zhang, Yan Lu, Nanning Zheng
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学)
;
Microsoft Research Asia(微软亚洲研究院)
机构
*
Department of Electrical and Computer Engineering, University of Massachusetts Lowell(马萨诸塞大学洛厄尔分校电子与计算机工程系)
;
School of Biomedical Engineering, Sun Yat-Sen University(中山大学生物医学工程学院)
DynamicRad: Content-Adaptive Sparse Attention for Long Video Diffusion
动态Rad:面向长视频扩散的内容自适应稀疏注意力
Yongji Long, Shijun Liang, Jintao Li, Yun Li
机构
*
University of Electronic Science and Technology of China(电子科学与技术大学)
;
Shenzhen Institute for Advanced Study, UESTC(深圳先进研究院)
;
Computational Mathematics, Science, & Engineering at Michigan State University (MSU)(密歇根州立大学计算数学、科学与工程)
GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers
GeoRelight: 基于灵活多模态扩散变换器的联合几何重照明与重建
Yuxuan Xue, Ruofan Liang, Egor Zakharov, Timur Bagautdinov, Chen Cao, Giljoo Nam, Shunsuke Saito, Gerard Pons-Moll, Javier Romero
机构
*
Codec Avatars Lab, Meta(Meta编码器动画实验室)
;
University of Tübingen(图宾根大学)
;
Max Planck Institute for Informatics, Saarland Informatics Campus(马克斯·普朗克信息学院,萨尔兰信息校园)
Single-Step Reconstruction-Free Anomaly Detection and Segmentation via Diffusion Models
基于扩散模型的实时无重建异常检测与分割
Mehrdad Moradi, Marco Grasso, Bianca Maria Colosimo, Kamran Paynabar
机构
*
H. Milton Stewart School of Industrial and Systems Engineering(H. Milton Stewart工业与系统工程学院)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Department of Mechanical Engineering(机械工程系)
;
Polytechnic University of Milan(米兰理工学院)
机构
*
Key Laboratory of Intelligent Information Processing, Institute of Computing Technology (ICT), Chinese Academy of Sciences (CAS)(智能信息处理重点实验室,计算技术研究所(ICT),中国科学院(CAS))
;
Jilin University (JLU)(吉林大学(JLU))
;
Beijing University of Posts and Telecommunications (BUPT)(北京邮电大学(BUPT))
;
University of the Chinese Academy of Sciences(中国科学院大学)
StomaD2: An All-in-One System for Intelligent Stomatal Phenotype Analysis via Diffusion-Based Restoration Detection Network
StomaD2:一种基于扩散恢复检测网络的智能气孔表型分析一体化系统
Quanling Zhao, Meng'en Qin, Yanfeng Sun, Yuan Miao, Xiaohui Yang
机构
*
Henan Engineering Research Center for Artificial Intelligence Theory and Algorithms(河南人工智能理论与算法工程研究中心)
;
School of Mathematics and Statistics(数学与统计学学院)
;
International Joint Research Laboratory for Global Change Ecology(全球变化生态学联合研究实验室)
;
School of Life Sciences(生命科学学院)
;
State Key Laboratory of Cotton Biology(棉花生物学国家重点实验室)
A Generalist Model for Diverse Text-Guided Medical Image Synthesis
一种通用模型用于多样化文本引导的医学图像合成
Joseph Cho, Mrudang Mathur, Cyril Zakka, Dhamanpreet Kaur, Matthew Leipzig, Alex Dalal, Aravind Krishnan, Eubee Koo, Karen Wai, Cindy S. Zhao, Akshay Chaudhari, Matthew Duda, Ashley Choi, Ehsan Rahimy, Lyna Azzouz, Robyn Fong, Rohan Shad, William Hiesinger
机构
*
Department of Cardiothoracic Surgery(心脏外科部门)
;
Stanford Medicine(斯坦福医学)
;
Department of Ophthalmology(眼科部门)
;
Division of Cardiovascular Surgery(心血管外科分会)
;
Penn Medicine(宾夕法尼亚医学)
Denoise and Align: Diffusion-Driven Foreground Knowledge Prompting for Open-Vocabulary Temporal Action Detection
去噪与对齐:基于扩散的前景知识提示用于开放词汇时序动作检测
Sa Zhu, Wanqian Zhang, Lin Wang, Jinchao Zhang, Cong Wang, Bo Li
机构
*
Institute of Information Engineering, Chinese Academy of Sciences School of Cyber Security, University of Chinese Academy of Sciences State Key Laboratory of Cyberspace Security Defense Beijing China
;
Institute of Information Engineering, Chinese Academy of Sciences Beijing China
;
Hangzhou Dianzi University Hangzhou China
;
Institute of Information Engineering, Chinese Academy of Sciences\ Key Laboratory of Cyberspace Security Defense Beijing China
;
Engineering, Zhejiang University Hangzhou China
;
Institute of Information Engineering, Chinese Academy of Sciences State Key Laboratory of Cyberspace Security Defense Beijing China
;
Institute of Information Engineering, Chinese Academy of Sciences School of Cyber Security, University of Chinese Academy of Sciences State Key Laboratory of Cyberspace Security Defense
;
Institute of Information Engineering, Chinese Academy of Sciences
;
Hangzhou Dianzi University
;
Institute of Information Engineering, Chinese Academy of Sciences\ Key Laboratory of Cyberspace Security Defense
;
Engineering, Zhejiang University
;
Institute of Information Engineering, Chinese Academy of Sciences State Key Laboratory of Cyberspace Security Defense
Diffusion-Based Feature Denoising and Using NNMF for Robust Brain Tumor Classification
基于扩散的特征去噪和使用NNMF进行鲁棒性脑肿瘤分类
Hiba Adil Al-kharsan, Róbert Rajkó
机构
*
Doctoral School of Computer Science, University of Szeged(计算机科学博士学院,塞格德大学)
;
Academic Staff, Doctoral School of Computer Science, University of Szeged(学术人员,计算机科学博士学院,塞格德大学)
New Fourth-Order Grayscale Indicator-Based Telegraph Diffusion Model for Image Despeckling
新的第四阶灰度指标基于电报扩散模型用于图像去斑
Rajendra K. Ray, Manish Kumar
机构
*
School of Mathematical and Statistical Sciences, Indian Institute of Technology Mandi(印度理工学院曼迪数学与统计科学学院)
;
Department of Mathematics, Indian Institute of Technology Delhi(印度理工学院德里数学系)
机构
*
University of California, Irvine(加州大学伊维特分校)
;
Stony Brook University(石溪大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of Florida(佛罗里达大学)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
R3D2:通过扩散实现自动驾驶模拟中的真实3D资产插入
William Ljungbergh, Bernardo Taveira, Wenzhao Zheng, Adam Tonderski, Chensheng Peng, Fredrik Kahl, Christoffer Petersson, Michael Felsberg, Kurt Keutzer, Masayoshi Tomizuka, Wei Zhan
机构
*
Institute of High Performance Computing, Agency for Science, Technology and Research, Singapore(高性能计算研究所,科技研究局,新加坡)
;
Institute for Infocomm Research, Agency for Science, Technology and Research, Singapore(信息通信研究所,科技研究局,新加坡)
;
Johns Hopkins University(约翰·霍普金斯大学)
机构
*
University of Alabama(阿拉巴马大学)
;
Emory University(埃默里大学)
;
University of Michigan(密歇根大学)
;
University of South Florida(佛罗里达州立大学)
;
New York University(纽约大学)