arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-09-03 至 2025-09-03 共收录 4 信号源:cs.CV, cs.GR, cs.MM

1. 图像生成评测 4 篇

2509.02474 2025-09-03 cs.GR cs.CV cs.LG 62%

Unifi3D: A Study on 3D Representations for Generation and Reconstruction in a Common Framework

Nina Wiedemann, Sainan Liu, Quentin Leboutet, Katelyn Gao, Benjamin Ummenhofer, Michael Paulitsch, Kai Yuan

机构 * Intel Corporation(英特尔公司)

专题命中 图像生成评测 :image generation(abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00777 2025-09-03 cs.GR cs.CV 62%

IntrinsicReal: Adapting IntrinsicAnything from Synthetic to Real Objects

Xiaokang Wei, Zizheng Yan, Zhangyang Xiong, Yiming Hao, Yipeng Qin, Xiaoguang Han

机构 * The Hong Kong Polytechnic University(香港理工大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Cardiff University(卡迪夫大学) NanJing XiaoZhuang University(南京小庄大学)

专题命中 图像生成评测 :diffusion(abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01411 2025-09-03 cs.CV 57%

MILO: A Lightweight Perceptual Quality Metric for Image and Latent-Space Optimization

Uğur Çoğalan, Mojtaba Bemana, Karol Myszkowski, Hans-Peter Seidel, Colin Groth

机构 * Max Planck Institute for Informatics(马克斯·普朗克研究所(信息研究所))

专题命中 图像生成评测 :diffusion(abstract);分类 cs.CV

Comments 11 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00381 2025-09-03 cs.CV cs.HC 57%

Visually Grounded Narratives: Reducing Cognitive Burden in Researcher-Participant Interaction

Runtong Wu, Jiayao Song, Fei Teng, Xianhao Ren, Yuyan Gao, Kailun Yang

机构 * School of Mathematics, Hunan University(湖南大学数学学院) School of Artificial Intelligence and Robotics, Hunan University(湖南大学人工智能与机器人学院) International Business School, Henan University of Economics and Law(河南财经政法大学国际商学院) Science & Technology College, University of Lorraine(洛林大学科学技术学院)

专题命中 图像生成评测 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏