Hoi3DGen: Generating High-Quality Human-Object-Interactions in 3D
Hoi3DGen:生成高质量的人-物交互的3D模型
Agniv Sharma, Xianghui Xie, Tom Fischer, Eddy Ilg, Gerard Pons-Moll
机构
*
University of Tübingen(图宾根大学)
;
Tübingen AI Center(图宾根人工智能中心)
;
Max Planck Institute for Informatics(马克斯·普朗克信息学院)
;
Technische Universität Nürnberg(纽伦堡技术大学)
机构
*
Computer Science and Engineering, Northeastern University, Shenyang, China(东北大学计算机科学与工程系,中国沈阳)
;
Key Laboratory of Intelligent Computing in Medical Image of Ministry of Education, Northeastern University, Shenyang, China(教育部医学图像智能计算重点实验室,东北大学,中国沈阳)
;
National Frontiers Science Center for Industrial Intelligence and Systems Optimization, Shenyang, China(工业智能与系统优化国家级前沿科学中心,中国沈阳)
;
AiShiWeiLai AI Research, China(艾世维来人工智能研究,中国)
;
Amii, University of Alberta, Edmonton, Alberta, Canada(阿尔伯塔大学艾米人工智能研究所,加拿大埃德蒙顿,阿尔伯塔)
专题命中
多模态生成
:cross-modal(abstract);分类 cs.CV
AI总结
本文提出视觉引导的文本解耦框架,通过细粒度语义解耦提升医学图像生成的可控性和生成质量。
Comments10 pages, 7 figures. Currently under review
StruVis: Enhancing Reasoning-based Text-to-Image Generation via Thinking with Structured Vision
StruVis: 通过结构化视觉进行推理的文本到图像生成增强
Yuanhuiyi Lyu, Kaiyu Lei, Ziqiao Weng, Xu Zheng, Lutao Jiang, Teng Li, Yangfu Li, Ziyuan Huang, Linfeng Zhang, Xuming Hu
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Ant Group(蚂蚁集团)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
East China Normal University(华东师范大学)
A Simple Baseline for Unifying Understanding, Generation, and Editing via Vanilla Next-token Prediction
一种通过 vanilla next-token 预测统一理解、生成和编辑的简单基线
Jie Zhu, Hanghang Ma, Jia Wang, Yayong Guan, Yanbing Zeng, Lishuai Gao, Junqiang Wu, Jie Hu, Leye Wang
机构
*
Key Lab of High Confidence Software Technologies (Peking University), Ministry of Education, China(高可信软件技术重点实验室(北京大学),教育部,中国)
;
School of Computer Science, Peking University, Beijing, China(北京大学计算机学院,北京,中国)
;
School of Computer Science and Technology, University of Chinese Academy of Sciences, Beijing, China(中国科学院大学计算机科学与技术学院,北京,中国)
Factuality Matters: When Image Generation and Editing Meet Structured Visuals
事实性至关重要:当图像生成与编辑遇见结构化视觉
Le Zhuo, Songhao Han, Yuandong Pu, Boxiang Qiu, Sayak Paul, Yue Liao, Yihao Liu, Jie Shao, Xi Chen, Si Liu, Hongsheng Li
机构
*
CUHK MMLab(香港大学多模态实验室)
;
Beihang University(北京航空航天大学)
;
Krea AI(Krea人工智能)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai AI Lab(上海人工智能实验室)
;
Hugging Face
;
National University of Singapore(新加坡国立大学)
;
ByteDance(字节跳动)
;
The University of Hong Kong(香港大学)