Studying Illustrations in Manuscripts: An Efficient Deep-Learning Approach
研究手稿中的插图:一种高效的深度学习方法
Yoav Evron, Michal Bar-Asher Siegal, Michael Fire
机构
*
Faculty of Computer and Information Science, Ben-Gurion University of the Negev, Be’er Sheva, Israel(计算机与信息科学学院,内盖夫本· Gurion大学,贝尔谢巴,以色列)
;
The Goldstein-Goren Department of Jewish Thought, Ben-Gurion University of the Negev, Be’er Sheva, Israel(犹太思想Goldstein-Goren部门,内盖夫本· Gurion大学,贝尔谢巴,以色列)
机构
*
Institute of Information Science, Beijing Jiaotong University(北京交通大学信息科学学院)
;
National University of Singapore(新加坡国立大学)
;
Meituan(美团)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Picsart AI Research (PAIR)(Picsart AI研究(PAIR))
Ander Alvarez, Alessandro Genuardi, Nilotpal Sinha, Antonio Tiene, Mikail Okyay, Bakbergen Ryskulov, David Montero, Samuel Mugel, Román Orús
机构
*
Multiverse Computing
;
Donostia International Physics Center
;
Ikerbasque Foundation for Science
;
Multiverse Computing, Centre for Social Innovation
Steering Vision-Language Pre-trained Models for Incremental Face Presentation Attack Detection
引导视觉-语言预训练模型进行增量面部呈现攻击检测
Haoze Li, Jie Zhang, Guoying Zhao, Stephen Lin, Shiguang Shan
机构
*
School of Computer Science, China University of Geosciences(中国地质大学(北京)计算机科学学院)
;
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences (CAS)(中国科学院人工智能安全重点实验室)
;
University of China Academy of Sciences(中国科学院大学)
;
Center for Machine Vision and Signal Analysis, University of Oulu(奥卢大学机器视觉与信号分析中心)
;
Microsoft Research Asia(微软亚洲研究院)
Embodied4C: Measuring What Matters for Embodied Vision-Language Navigation
Embodied4C: 评估具身视觉-语言导航中至关重要的因素
Tin Stribor Sohn, Maximilian Dillitzer, Jason J. Corso, Eric Sax
机构
*
Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)
;
UAS Esslingen(埃森嫩大学)
;
TU Wien(维也纳技术大学)
;
Dr. Ing. h.c. F. Porsche AG(保时捷股份有限公司)
;
University of Michigan(密歇根大学)
;
Voxel51 Inc.(Voxel51公司)
On the dynamic evolution of CLIP texture-shape bias and its relationship to human alignment and model robustness
CLIP纹理-形状偏倚的动态演变及其与人类对齐和模型鲁棒性的关系
Pablo Hernández-Cámara, Jose Manuel Jaén-Lorites, Alexandra Gómez-Villa, Jorge Vila-Tomás, Valero Laparra, Jesus Malo
机构
*
Image Processing Lab, Universitat de Valencia, Spain(瓦伦西亚大学图像处理实验室)
;
Centro de Biomateriales e Ingenieria Tisular, Universitat Politecnica de Valencia, Spain(瓦伦西亚理工大学生物材料与组织工程中心)
;
Computer Vision Center, Spain(西班牙计算机视觉中心)
;
Universitat Autònoma de Barcelona, Spain(巴塞罗那自治大学)
Continual Learning on CLIP via Incremental Prompt Tuning with Intrinsic Textual Anchors
通过内在文本锚点进行增量提示调优实现CLIP的持续学习
Haodong Lu, Xinyu Zhang, Kristen Moore, Jason Xue, Lina Yao, Anton van den Hengel, Dong Gong
机构
*
School of Computer Science and Engineering, University of New South Wales(新南威尔士大学计算机科学与工程学院)
;
Data61, CSIRO(CSIRO数据61研究中心)
;
School of Computer Science, University of Auckland(奥克兰大学计算机科学学院)
;
Australian Institute for Machine Learning (AIML), The University of Adelaide(阿德莱德大学澳大利亚机器学习研究所)
R4: Retrieval-Augmented Reasoning for Vision-Language Models in 4D Spatio-Temporal Space
R4:在4D时空空间中为视觉语言模型引入检索增强推理
Tin Stribor Sohn, Maximilian Dillitzer, Jason J. Corso, Eric Sax
机构
*
Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)
;
Esslingen University of Applied Sciences(埃斯林根应用科学大学)
;
Dr. Ing. h.c. F. Porsche AG(德意志联邦汽车工业协会)
;
University of Michigan(密歇根大学)
;
Voxel51 Inc.(Voxel51公司)
An Efficient and Effective Encoder Model for Vision and Language Tasks in the Remote Sensing Domain
一种用于遥感领域的高效且有效的编码器模型
João Daniel Silva, Joao Magalhaes, Devis Tuia, Bruno Martins
机构
*
INESC-ID, Instituto Superior Técnico, University of Lisbon(INESC-ID,理工学院,里斯本大学)
;
Department of Computer Science, Faculty of Science and Technology, Universidade NOVA de Lisboa(计算机科学系,科学与技术学院,诺瓦大学)
;
School of Architecture, Civil and Environmental Engineering, EPFL(建筑、土木和环境工程学院,EPFL)