AIVV: Neuro-Symbolic LLM Agent-Integrated Verification and Validation for Trustworthy Autonomous Systems
AIVV: 用于可信自主系统的神经符号LLM代理集成验证与验证
Jiyong Kwon, Ujin Jeon, Sooji Lee, Guang Lin
机构
*
School of Mechanical Engineering, Purdue University(普渡大学机械工程学院)
;
School of Electrical and Computer Engineering, Purdue University(普渡大学电气与计算机工程学院)
;
Department of Computer Science, Purdue University(普渡大学计算机科学系)
;
Department of Mathematics, Purdue University(普渡大学数学系)
Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models
对前沿视觉-语言模型进行审计以实现可信的医学视觉问答:定位失败、格式崩溃和领域适应
Xupeng Chen, Binbin Shi, Chenqian Le, Qifu Yin, Lang Lin, Haowei Ni, Ran Gong, Panfeng Li
机构
*
New York University, New York, USA(纽约大学)
;
Tsinghua University, Beijing, China(清华大学)
;
Columbia University, New York, USA(哥伦比亚大学)
;
University of Michigan, Ann Arbor, USA(密歇根大学)
Comments6 Pages, 2 Figures. Accepted and presented at the 30th International Conference on Optical Network Design and Modelling (ONDM 2026), Munich, Germany, 12-15 May 2026
Decoding the Multimodal Mind: Generalizable Brain-to-Text Translation via Multimodal Alignment and Adaptive Routing
解码多模态思维:通过多模态对齐和自适应路由实现可泛化的脑到文本翻译
Chunyu Ye, Yunhao Zhang, Jingyuan Sun, Chong Li, Yang Zhao, Shaonan Wang
机构
*
State Key Laboratory of Multimodal Artificial Intelligence System, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Department of Computer Science, The University of Manchester(曼彻斯特大学计算机科学系)
;
Department of Language Science and Technology, Hong Kong Polytechnic University(香港理工大学语言科学与技术系)
Trustworthy Predictive Distributions for Tail Events with Semiparametric Diagnostic Transport Maps
面向尾部事件的可信预测分布:基于半参数诊断传输图
Elizabeth Cucuzzella, Rafael Izbicki, Ann B. Lee
机构
*
Department of Statistics and Data Science, Carnegie Mellon University(统计与数据科学系,卡内基梅隆大学)
;
Department of Statistics, Federal University of Sao Carlos(统计系,圣卡洛斯联邦大学)
Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety
用于代理网络运维和AI运维的大型语言模型:架构、评估与安全
Muhammad Bilal, Jon Crowcroft, Ruizhi Wang, Xiaolong Xu, Schahram Dustdar
机构
*
School of Computing and Communications(计算与通信学院)
;
University of Cambridge(剑桥大学)
;
School of Software(软件学院)
;
Nanjing University of Information Science and Technology(南京信息科技大學)
;
TU Wien(维也纳技术大学)
;
ICREA
Comments6 pages, 3 figures. Accepted at the Security, Trust and Privacy for Software and Applications (STPSA) Workshop, IEEE COMPSAC 2026, Madrid, Spain, July 7-10, 2026
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
基于LLM的软件工程社区心理安全定性编码的提示工程策略:一项受控实证研究
Moaath Alshaikh, Tasneem Alshaher, Ricardo Vieira, Beatriz Santana, Clelio Xavier, Jose Amancio, Glauco Carneiro, Julio Leite, Savio Freire, Manoel Mendonca
机构
*
Federal University of Bahia(巴伊亚联邦大学)
;
State University of Feira de Santana(费拉德桑塔纳州立大学)
;
Federal University of Sergipe(塞格皮联邦大学)
;
Federal Institute of Ceara(塞阿拉联邦理工学院)
Comments9 pages, 5 figures. Accepted at the 1st International Workshop on Prompt Engineering for Software Engineering (PROMPT-SE 2026), co-located with the 30th International Conference on Evaluation and Assessment in Software Engineering (EASE 2026), Glasgow, Scotland, United Kingdom, June 9--12, 2026