AnimateAnyMesh++: A Flexible Feed-Forward Framework for High-Fidelity Text-Driven Mesh Animation
AnimateAnyMesh++: 一种灵活的4D基础模型用于高质量文本驱动的网格动画
Zijie Wu, Chaohui Yu, Fan Wang, Xiang Bai
机构
*
School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)
;
DAMO Academy, Alibaba Group(阿里巴巴达摩院)
;
Hupan Lab, Hangzhou, China(湖畔实验室)
;
School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)
Roberto Di Via, Irina Voiculescu, Francesca Odone, Vito Paolo Pastore
机构
*
MaLGa
;
DIBRIS, University of Genoa(DIBRIS,热那亚大学)
;
University of Genoa(热那亚大学)
;
Department of Computer Science, University of Oxford(奥大利大学计算机科学系)
MuellerPT: Decomposition Driven Pretraining for Dense Learning in Mueller Polarimetry
MuellerPT: 穆勒偏振测量中密集学习的分解驱动预训练
Adam Tlemsani, Yingdian Li, Maxime Giot, Naim Slim, Christopher J. Peters, Abhijeet Ghosh, Daniel S. Elson
机构
*
Department of Computing, Imperial College London(帝国理工学院计算机系)
;
Hamlyn Centre for Robotic Surgery, Imperial College London(帝国理工学院机器人外科中心)
;
Department of Surgery and Cancer, Imperial College London(帝国理工学院外科与癌症系)
;
Xi’an Institute of Optics and Precision Mechanics, Chinese Academy of Sciences(中国科学院西安光学精密机械研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
UniE2F: A Unified Diffusion Framework for Event-to-Frame Reconstruction with Video Foundation Models
UniE2F: 一种基于视频基础模型的统一扩散框架用于事件到帧重建
Gang Xu, Zhiyu Zhu, Junhui Hou
机构
*
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室(深圳))
;
Department of Computer Science, City University of Hong Kong (Dongguan)(香港城市大学(东莞)计算机科学系)
机构
*
Beijing Institute of Technology(北京理工大学)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(国家一般人工智能重点实验室,BIGAI)
;
Deep Glint
;
Zhejiang University(浙江大学)
;
The University of Hong Kong(香港大学)
;
Shandong Agricultural University(山东农业大学)
PanoSplatt3R: Leveraging Perspective Pretraining for Generalized Unposed Wide-Baseline Panorama Reconstruction
Jiahui Ren, Mochu Xiang, Jiajun Zhu, Yuchao Dai
机构
*
School of Electronics and Information, Northwestern Polytechnical University and Shaanxi Key Laboratory of Information Acquisition and Processing(电子工程学院、西北工业大学和陕西省信息获取与处理重点实验室)
SceneSplat: Gaussian Splatting-based Scene Understanding with Vision-Language Pretraining
Yue Li, Qi Ma, Runyi Yang, Huapeng Li, Mengjiao Ma, Bin Ren, Nikola Popovic, Nicu Sebe, Ender Konukoglu, Theo Gevers, Luc Van Gool, Martin R. Oswald, Danda Pani Paudel
机构
*
University of Amsterdam(阿姆斯特丹大学)
;
Computer Vision Lab, ETH Zurich(苏黎世联邦理工学院计算机视觉实验室)
;
INSAIT, Sofia University ”St. Kliment Ohridski”(索菲亚大学”圣克莱门特·欧赫里迪斯”研究所)
;
Nanjing University of Aeronautics and Astronautics(南京航空航天大学)
;
University of Pisa(比萨大学)
;
University of Trento(特伦特大学)
LLM-powered Query Expansion for Enhancing Boundary Prediction in Language-driven Action Localization
Zirui Shang, Xinxiao Wu, Shuo Yang
机构
*
Beijing Key Laboratory of Intelligent Information Technology, School of Computer Science & Technology(北京智能信息科技重点实验室,计算机科学与技术学院)
;
Beijing Institute of Technology(北京理工大学)
;
Guangdong Laboratory of Machine Perception and Intelligent Computing(广东机器感知与智能计算实验室)
;
Shenzhen MSU-BIT University(深圳MSU-BIT大学)
Enhancement-Driven Pretraining for Robust Fingerprint Representation Learning
Ekta Gavas, Kaustubh Olpadkar, Anoop Namboodiri
机构
*
Centre for Visual Information Technology, International Institute of Information Technology, Hyderabad, India(视觉信息技术中心,国际信息学院,印度海得拉巴)
;
Stony Brook University, USA(石溪大学)
Journal refProceedings of the 19th International Joint Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications (VISIGRAPP 2024) - Volume 2: VISAPP, ISBN 978-989-758-679-8, ISSN 2184-4321, pages 821-828