CommentsAccepted for publication at the 7th International Workshop on Deep Learning for Testing and Testing for Deep Learning (DeepTest 2026), co-located with ICSE 2026
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Zhejiang University(浙江大学)
;
Savannah College of Art and Design(萨凡纳艺术设计学院)
;
West China Hospital, Sichuan University(四川大学华西医院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract)
CommentsThis version introduces a major architectural shift to Local LLMs and NLI-based assignment, scaling the framework to O(1) generative complexity. Formerly titled 'Question-Driven Analysis and Synthesis'
CARE Drive A Framework for Evaluating Reason-Responsiveness of Vision Language Models in Automated Driving
CARE Drive:一种用于评估视觉语言模型在自动驾驶中推理响应性的框架
Lucas Elbert Suryana, Farah Bierenga, Sanne van Buuren, Pepijn Kooij, Elsefien Tulleners, Federico Scari, Simeon Calvert, Bart van Arem, Arkady Zgonnikov
机构
*
organization= Department of Transport \& Planning, Faculty of Civil Engineering
;
Geosciences, Delft University of Technology
;
organization= Department of Cognitive Robotics, Faculty of Mechanical Engineering, Delft University of Technology
;
organization= Centre for Meaningful Human Control, Delft University of Technology
;
organization= Faculty of Mechanical Engineering, Delft University of Technology , city= Delft , country= The Netherlands
Unforgeable Watermarks for Language Models via Robust Signatures
通过鲁棒签名实现语言模型的不可伪造水印
Huijia Lin, Kameron Shahabi, Min Jae Song
机构
*
Paul G. Allen School of Computer Science & Engineering, University of Washington(保罗·G·艾伦计算机科学与工程学院,华盛顿大学)
;
Data Science Institute, University of Chicago(数据科学研究所,芝加哥大学)
Measuring Social Integration Through Participation: Categorizing Organizations and Leisure Activities in the Displaced Karelians Interview Archive using LLMs
CommentsPresented at: The 10th Joint SIGHUM Workshop on Computational Linguistics for Cultural Heritage, Social Sciences, Humanities and Literature; EACL 2026 Workshop
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
PII-Bench:评估查询感知隐私保护系统
Hao Shen, Zhouhong Gu, Haokai Hong, Weili Han
机构
*
Institute of Fintech, Fudan University(复旦大学金融科技学院)
;
Shanghai Key Laboratory of Data Science, School of Computer Science, Fudan University(复旦大学数据科学上海重点实验室)
;
Laboratory of Data Analytics and Security, Fudan University(复旦大学数据安全分析实验室)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
EduResearchBench: A Hierarchical Atomic Task Decomposition Benchmark for Full-Lifecycle Educational Research
EduResearchBench: 一个用于全生命周期教育研究的分层原子任务分解基准
Houping Yue, Zixiang Di, Mei Jiang, Bingdong Li, Hao Hao, Yu Song, Bo Jiang, Aimin Zhou
机构
*
Shanghai Institute of AI for Education, East China Normal University(上海人工智能教育研究院,东华大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
School of Computer Science and Technology, East China Normal University(东华大学计算机科学与技术学院)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Sascha Diefenbacher, Anna Hallin, Gregor Kasieczka, Michael Krämer, Anne Lauscher, Tim Lukas
机构
*
Physics Division, Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室物理部)
;
Institut für Experimentalphysik, Universität Hamburg(汉堡大学实验物理研究所)
;
Institute for Theoretical Particle Physics and Cosmology, RWTH Aachen University(亚琛工业大学理论粒子物理与宇宙学研究所)
;
Data Science Group, Universität Hamburg(汉堡大学数据科学组)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Bayesian Optimization for Design Parameters of 3D Image Data Analysis
基于贝叶斯优化的3D图像数据分析设计参数优化
David Exler, Joaquin Eduardo Urrutia Gómez, Martin Krüger, Maike Schliephake, John Jbeily, Mario Vitacolonna, Rüdiger Rudolf, Markus Reischl
机构
*
Institute for Automation and Applied Informatics, Karlsruhe Institute of Technology(自动化与应用信息学院,卡尔斯鲁厄技术大学)
;
CeMOS Research and Transfer Center, Technische Hochschule Mannheim(CeMOS研究与转移中心,技术大学曼海姆)
;
Institute of Biological and Chemical Systems, Karlsruhe Institute of Technology(生物与化学系统研究所,卡尔斯鲁厄技术大学)
K. Iwasawa, R. Gilli, F. Vito, Y. Matsuoka, M. Onoue, M. A. Strauss, N. Kashikawa, Y. Toba, K. Shimasaku, K. Inayoshi, T. Nagao, N. Kawanaka, J. D. Silverman, T. Izumi, K. Kohno, Y. Ueda