机构
*
Institute of Biopharmaceutical and Health Engineering, Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学生物医药与健康工程研究院,清华大学深圳国际研究生院)
;
Department of Engineering Science, University of Oxford(牛津大学工程科学系)
;
Department of Pathology, The First Affiliated Hospital of Sun Yat-sen University(中山大学附属第一医院病理科)
;
Medical Optical Technology R&D Center, Research Institute of Tsinghua, Pearl River Delta(清华珠三角研究院医学光学技术研发中心)
Parameter-Efficient Vision-Language Adaptation with Continuous Metadata Conditioning for Animal Re-Identification
基于连续元数据条件的参数高效视觉语言适配用于动物再识别
Anil Osman Tur, Tonje Knutsen Sordalen, Kim Tallaksen Halvorsen, Cigdem Beyan
机构
*
Department of Computer Science, University of Verona(维罗纳大学计算机科学系)
;
Institute of Marine Research(海洋研究所)
;
University of Agder, Centre for Coastal Research(阿格德大学海岸研究中心)
CommentsThis is the author's version of the paper accepted for publication in Expert Systems with Applications. The final authenticated version will be available from the publisher
Explainable Novel Category Discovery in Semantic Concept Space
语义概念空间中的可解释新类别发现
Ifrat Ikhtear Uddin, Yang Zhou, KC Santosh, Longwei Wang
机构
*
Department of Computer Science, University of South Dakota(南达科他大学计算机科学系)
;
Department of Computer Science and Software Engineering, Auburn University(奥本大学计算机科学与软件工程系)
UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving
UniDrive: 面向自动驾驶可解释风险理解的统一视觉-语言与定位框架
Xiaowei Gao, Pengxiang Li, Yitai Cheng, Ruihan Xu, James Haworth, Stephen Law, Yun Ye
机构
*
organization= Department of Earth Science \& Engineering, Imperial College London , city= London , postcode= SW7 2AZ , country= United Kingdom
;
organization= SpaceTimeLab, Department of Civil, Environmental
;
Geomatic Engineering, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Department of Computing, The Hong Kong Polytechnic University , city= Hong Kong , country= China
;
organization= Trinity College, University of Oxford , city= Oxford , postcode= OX1 3BH , country= United Kingdom
;
organization= Department of Geography, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Centre for Global Infrastructure Resilience, The Bartlett School of Sustainable Construction, University College London , city= London , postcode= WC1E 7HB , country= United Kingdom
Page image classifier fine-tuned on century-spanning archives of scanned documents for further content-specific processing
基于百年跨度扫描文档档案微调的页面图像分类器,用于进一步的内容特定处理
Kateryna Lutsai, Dana Křivánková, Pavel Straňák, David Novák
机构
*
Institute of Formal and Applied Linguistics, Charles University MFF(查尔斯大学数学与物理学院形式与应用语言学研究所)
;
Institute of Archaeology, Czech Academy of Sciences(捷克科学院考古研究所)
Scalable Training of Spatially Grounded 2D Vision-Language Models for Radiology
面向放射学的空间定位2D视觉-语言模型的可扩展训练
Yusuf Salcan, Simon Ging, Robin Tibor Schirrmeister, Philipp Arnold, Elmar Kotter, Behzad Bozorgtabar, Thomas Brox
机构
*
Computer Vision Group, University of Freiburg, Germany(德国弗莱堡大学计算机视觉组)
;
Department of Radiology, Medical Center -- University of Freiburg, Germany(德国弗莱堡大学医学中心放射科)
;
CRIION-AI Lab, Freiburg, Germany(德国弗莱堡CRIION-AI实验室)
机构
*
Space Applications Centre, ISRO(印度空间研究组织空间应用中心)
;
Indian Institute of Science Education and Research (IISER) Bhopal(印度科学教育与研究学院(博帕尔))
;
Indian Institute of Technology Bombay(印度理工学院孟买分校)
Test-Time Hallucination Control in Large Vision-Language Models
大型视觉语言模型的测试时幻觉控制
Mehran Tamjidi, Hamidreza Dastmalchi, Ali Cheraghian, Mohammadreza Alimoradijazi, Aijun An, Hossein Rahmani
机构
*
Australian National University(澳大利亚国立大学)
;
The University of New South Wales(新南威尔士大学)
;
University of Technology Sydney(悉尼科技大学)
;
York University(约克大学)
;
Lancaster University(兰卡斯特大学)