LTX-2: Efficient Joint Audio-Visual Foundation Model
LTX-2:高效的音频-视觉联合基础模型
Yoav HaCohen, Benny Brazowski, Nisan Chiprut, Yaki Bitterman, Andrew Kvochko, Avishai Berkowitz, Daniel Shalem, Daphna Lifschitz, Dudu Moshe, Eitan Porat, Eitan Richardson, Guy Shiran, Itay Chachy, Jonathan Chetboun, Michael Finkelson, Michael Kupchick, Nir Zabari, Nitzan Guetta, Noa Kotler, Ofir Bibi, Ori Gordon, Poriya Panet, Roi Benita, Shahar Armon, Victor Kulikov, Yaron Inger, Yonatan Shiftan, Zeev Melumian, Zeev Farbman
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
China Mobile (Jiangxi) Virtual Reality Technology Co., Ltd.(中国移动(江西)虚拟现实技术有限公司)
;
School of Computer and Information Technology, Beijing Jiaotong University(北京交通大学计算机与信息学院)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
Comments8 pages for main paper (exclude citation pages), 6 pages for appendix, totally 10 figures 7 tables and 2 algorithms. The paper is accepted by WACV 2026
Journal refIEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026
机构
*
Shanghai Key Laboratory of Navigation & Location Based Services, Shanghai Jiao Tong University(上海导航与基于位置的服务关键实验室,上海交通大学)
;
Beijing Aerospace Microsystem and Information Technology Research Institute(北京航天微系统与信息技术研究所)
CommentsThis manuscript contains substantive technical inaccuracies and an incomplete treatment of the stated topic. Subsequent developments and a reassessment of the problem indicate that the scope and framing of the work do not adequately reflect the current state of research, and the analysis is therefore incomplete and outdated
机构
*
University of Washington, Seattle(华盛顿大学)
;
Bar-Ilan University(巴伊兰大学)
;
University of California, Irvine(加州大学伊文斯顿分校)
;
Allen Institute of AI(人工智能研究院)
Uni-FinLLM: A Unified Multimodal Large Language Model with Modular Task Heads for Micro-Level Stock Prediction and Macro-Level Systemic Risk Assessment
机构
*
The University of Hong Kong(香港大学)
;
University of Michigan, Ann Arbor(密歇根大学安娜堡分校)
;
University of Edinburgh(爱丁堡大学)
;
University of California, Santa Barbara(加州大学圣巴巴拉分校)
;
Ohio State University(俄亥俄州立大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
X-RAFT: Cross-Modal Non-Rigid Registration of Blue and White Light Neurosurgical Hyperspectral Images
X-RAFT:跨模态非刚性注册蓝光和白光神经外科高光谱图像
Charlie Budd, Silvère Ségaud, Matthew Elliot, Graeme Stasiuk, Yijing Xie, Jonathan Shapey, Tom Vercauteren
机构
*
Department of Surgical \& Interventional Engineering, School of Biomedical Engineering \& Imaging Sciences, King's College London, London SE1 7EH Department of Neurosurgery, King's College London Hospital NHS Foundation Trust, London SE5 9RS, UK. Department of Imaging Chemistry \& Biology, School of Biomedical Engineering \& Imaging Sciences, King's College London, London SE1 7EH, UK.