Visual Autoregressive Modelling for Monocular Depth Estimation
单目深度估计的视觉自回归建模
Amir El-Ghoussani, André Kaup, Nassir Navab, Gustavo Carneiro, Vasileios Belagiannis
机构
*
Friedrich-Alexander University Erlangen-Nuremberg, Germany(埃朗根-纽伦堡弗里德里希-亚历山大大学)
;
Technical University of Munich, Germany(慕尼黑技术大学)
;
University of Surrey, United Kingdom(萨里大学)
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
ByteDance(字节跳动)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
The Hong Kong Polytechnic University(香港理工大学)
机构
*
Queen Mary University of London(伦敦女王学院)
;
Centre for AI, AstraZeneca(AstraZeneca人工智能中心)
;
University of Bedfordshire(贝德福德大学)
;
Samsung AI Center(三星人工智能中心)
CommentsThe manuscript is 23 pages, with five main figures and one table. The supplemental material includes 23 pages with fourteen figures and four tables
AniMer+: Unified Pose and Shape Estimation Across Mammalia and Aves via Family-Aware Transformer
Liang An, Jin Lyu, Li Lin, Pujin Cheng, Yebin Liu, Xiaoying Tang
机构
*
Department of Electronic and Electrical Engineering, Southern University of Science and Technology, Shenzhen, China(南方科技大学电子与电气工程系)
;
Department of Automation, Tsinghua University, Beijing, China(清华大学自动化系)
;
Jiaxing Research Institute, Southern University of Science and Technology, Jiaxing, China(南方科技大学嘉兴研究所)
;
Department of Electrical and Electronic Engineering, the University of Hong Kong, Hong Kong, China(香港大学电子与电气工程系)
Target-Guided Bayesian Flow Networks for Quantitatively Constrained CAD Generation
Wenhao Zheng, Chenwei Sun, Wenbo Zhang, Jiancheng Lv, Xianggen Liu
机构
*
College of Computer Science, Sichuan University(四川大学计算机科学学院)
;
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Engineering Research Center of Machine Learning and Industry Intelligence, Ministry of Education, Chengdu, China(教育部机器学习与工业智能工程研究中心)
InfiniHuman: Infinite 3D Human Creation with Precise Control
Yuxuan Xue, Xianghui Xie, Margaret Kostyrko, Gerard Pons-Moll
机构
*
University of Tübingen, Tübingen AI Center(图宾根大学,图宾根人工智能中心)
;
University of Tübingen, Tübingen AI Center, MPI for Informatics, SIC(图宾根大学,图宾根人工智能中心,马克斯·普朗克信息研究所,SIC)
;
University of Tübingen(图宾根大学)
SemanticControl: A Training-Free Approach for Handling Loosely Aligned Visual Conditions in ControlNet
Woosung Joung, Daewon Chae, Jinkyu Kim
机构
*
Department of Computer Science and Engineering(计算机科学与工程系)
;
Korea University(韩国大学)
;
Electrical and Computer Engineering(电气与计算机工程系)
;
University of Michigan(密歇根大学)
DanceText: A Training-Free Layered Framework for Controllable Multilingual Text Transformation in Images
Zhenyu Yu, Mohd Yamani Idna Idris, Hua Wang, Pei Wang, Rizwan Qureshi, Shaina Raza, Aman Chadha, Yong Xiang, Zhixiang Chen
机构
*
Universiti Malaya(马来大学)
;
Zhejiang University of Finance and Economics Dongfang College(浙江财经大学东阳学院)
;
Kunming University of Science and Technology(昆明理工大学)
;
University of Central Florida(佛罗里达中央大学)
;
Toronto metropolitan university(多伦多 Metropolitan 大学)
;
Amazon Web Services(亚马逊网络服务)
;
Deakin University(德肯大学)
;
University of Sheffield(谢菲尔德大学)