TextDS: Parameter-Efficient Representation Alignment for Scene Text Detection under Distribution Shifts
TextDS: 分布偏移下场景文本检测的参数高效表示对齐
Boyuan Chen, Zichen Dang, Chuang Yang, Lap-Pui Chau, Yi Wang
机构
*
School of Electrical Engineering, Xi’an Jiaotong University(西安交通大学电气工程学院)
;
Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(香港理工大学电机及电子工程学系)
LALM-as-a-Judge: Benchmarking Large Audio-Language Models for Safety Evaluation in Multi-Turn Spoken Dialogues
LALM-as-a-Judge:用于多轮口语对话安全评估的大型音频语言模型基准测试
Amir Ivry, Shinji Watanabe
机构
*
Computer Engineering, Technion--Israel Institute of Technology, Haifa, Israel(技术学院电子工程系,技术离子技术研究所,以色列海法)
;
Language Technologies Institute, Carnegie Mellon University, Pittsburgh, PA, USA(语言技术研究所,卡内基梅隆大学,美国匹兹堡)
CapCLIP: A Vision-Language Representation Alignment Approach for Wireless Capsule Endoscopy Analysis
CapCLIP:一种用于无线胶囊内镜分析的视觉-语言表示对齐方法
Haroon Wahab, Irfan Mehmood, Hassan Ugail
机构
*
School of Computer Science, AI and Electronics Faculty of Engineering and Digital Technologies(计算机科学与电子工程学院,工程与数字技术学院)
;
School of Management Faculty of Mgmt, Law & Social Sciences(管理学院,管理、法律与社会科学学院)
;
Centre for Visual Computing and Intelligent Systems(视觉计算与智能系统中心)
CommentsTo be published in IEEE International Workshop on Decentralized Physical Infrastructure Networks 2025, in conjunction with ICBC'25. 7 pages. 3 figures
CommentsThis is an earlier version of the work released in May 2025. The version accepted at CHI 2026 is available as a separate preprint at arXiv:2511.04366
Safe-LLaVA: A Privacy-Preserving Vision-Language Dataset and Benchmark for Biometric Safety
Younggun Kim, Sirnam Swetha, Fazil Kagdi, Mubarak Shah
机构
*
Center For Research in Computer Vision, University of Central Florida, USA(计算机视觉研究中心,中央佛罗里达大学)
;
Department of Civil Environmental and Construction Engineering, University of Central Florida, USA(土木环境与建设工程系,中央佛罗里达大学)
;
Department of Computer Science, University of Central Florida, USA(计算机科学系,中央佛罗里达大学)
Towards a Framework for Operationalizing the Specification of Trustworthy AI Requirements
Hugo Villamizar, Daniel Mendez, Marcos Kalinowski
专题命中
安全评测
:trustworthy(title)
CommentsThis paper has been accepted for presentation at the 2025 IEEE 33rd International Requirements Engineering Conference Workshops (REW-RETRAI 2025)
MeDSLIP: Medical Dual-Stream Language-Image Pre-training with Pathology-Anatomy Semantic Alignment
Wenrui Fan, Mohammod N. I. Suvon, Shuo Zhou, Xianyuan Liu, Samer Alabed, Venet Osmani, Andrew J. Swift, Chen Chen, Haiping Lu
机构
*
Centre for Machine Intelligence and School of Computer Science, University of Sheffield(智能中心和计算机科学学院,谢菲尔德大学)
;
School of Medicine and Population Health, and INSIGNEO, Institute for in Silico Medicine, University of Sheffield(医学与人口健康学院,INSIGNEO,虚拟医学研究所,谢菲尔德大学)
;
Digital Environment Research Institute, Queen Mary University of London(数字环境研究所,伦敦大学玛丽女王学院)
;
School of Computer Science, University of Sheffield(计算机科学学院,谢菲尔德大学)
;
Department of Computing, Imperial College London(计算系,伦敦帝国学院)