AIVV: Neuro-Symbolic LLM Agent-Integrated Verification and Validation for Trustworthy Autonomous Systems
AIVV: 用于可信自主系统的神经符号LLM代理集成验证与验证
Jiyong Kwon, Ujin Jeon, Sooji Lee, Guang Lin
机构
*
School of Mechanical Engineering, Purdue University(普渡大学机械工程学院)
;
School of Electrical and Computer Engineering, Purdue University(普渡大学电气与计算机工程学院)
;
Department of Computer Science, Purdue University(普渡大学计算机科学系)
;
Department of Mathematics, Purdue University(普渡大学数学系)
Why Does Grounding Hurt Medical VQA? Benchmarking, Diagnosis, and Fine-Tuning of Vision-Language Models
对前沿视觉-语言模型进行审计以实现可信的医学视觉问答:定位失败、格式崩溃和领域适应
Xupeng Chen, Binbin Shi, Chenqian Le, Qifu Yin, Lang Lin, Haowei Ni, Ran Gong, Panfeng Li
机构
*
New York University, New York, USA(纽约大学)
;
Tsinghua University, Beijing, China(清华大学)
;
Columbia University, New York, USA(哥伦比亚大学)
;
University of Michigan, Ann Arbor, USA(密歇根大学)
Decoding the Multimodal Mind: Generalizable Brain-to-Text Translation via Multimodal Alignment and Adaptive Routing
解码多模态思维:通过多模态对齐和自适应路由实现可泛化的脑到文本翻译
Chunyu Ye, Yunhao Zhang, Jingyuan Sun, Chong Li, Yang Zhao, Shaonan Wang
机构
*
State Key Laboratory of Multimodal Artificial Intelligence System, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Department of Computer Science, The University of Manchester(曼彻斯特大学计算机科学系)
;
Department of Language Science and Technology, Hong Kong Polytechnic University(香港理工大学语言科学与技术系)
Trustworthy Predictive Distributions for Tail Events with Semiparametric Diagnostic Transport Maps
面向尾部事件的可信预测分布:基于半参数诊断传输图
Elizabeth Cucuzzella, Rafael Izbicki, Ann B. Lee
机构
*
Department of Statistics and Data Science, Carnegie Mellon University(统计与数据科学系,卡内基梅隆大学)
;
Department of Statistics, Federal University of Sao Carlos(统计系,圣卡洛斯联邦大学)
Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety
用于代理网络运维和AI运维的大型语言模型:架构、评估与安全
Muhammad Bilal, Jon Crowcroft, Ruizhi Wang, Xiaolong Xu, Schahram Dustdar
机构
*
School of Computing and Communications(计算与通信学院)
;
University of Cambridge(剑桥大学)
;
School of Software(软件学院)
;
Nanjing University of Information Science and Technology(南京信息科技大學)
;
TU Wien(维也纳技术大学)
;
ICREA
Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for Secure Network Electronic Health Record (EHR) Interoperability
Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness
Infra-Bayesian 强化学习智能体在最坏情况鲁棒性上优于经典强化学习
Manish Aryal, Faiyaz Azam, Agnivo Banerjee, Syed Mahir Ahamed, Sai Sidhanth Manoharan Jayanthi, Allegra Laro, Clément Legentilhomme, Andrew Lin, Florian Lorkowski, Marina Pérez del Valle, Radman Rakhshandehroo, Patric Rommel, Emanuel Ruzak, Nathan Theng, Paul Yushin Rapoport
机构
*
Purdue University(普渡大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
WorldQuant University(WorldQuant大学)
;
UC Berkeley(加州大学伯克利分校)
;
Aix-Marseille University(阿维尼翁-马赛大学)
;
MIT(麻省理工学院)
;
University of Zurich(苏黎世大学)
;
University of British Columbia(不列颠哥伦比亚大学)
;
University of Stuttgart(斯图加特大学)
;
University of Buenos Aires(布宜诺斯艾利斯大学)
;
California State University, Fresno(弗雷斯诺加州州立大学)
;
University of Chicago(芝加哥大学)
Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts
神经符号AI中的符号接地:推理快捷方式的入门介绍
Emanuele Marconato, Samuele Bortolotti, Emile van Krieken, Paolo Morettin, Elena Umili, Antonio Vergari, Efthymia Tsamoura, Andrea Passerini, Stefano Teso
机构
*
University of Trento(特伦托大学)
;
Vrije Universiteit Amsterdam(阿姆斯特丹自由大学)
;
Sapienza University of Rome(罗马大学)
;
University of Edinburgh(爱丁堡大学)
;
Huawei Labs(华为实验室)
Estimating Tail Risks in Language Model Output Distributions
语言模型输出分布中的尾部风险估计
Rico Angell, Raghav Singhal, Zachary Horvitz, Zhou Yu, Rajesh Ranganath, Kathleen McKeown, He He
机构
*
Columbia University(哥伦比亚大学)
;
Department of Computer Science, New York University(纽约大学计算机科学系)
;
Center for Data Science, New York University(纽约大学数据科学中心)
Don't Walk the Line: Boundary Guidance for Filtered Generation
不要走线:边界引导用于过滤生成
Sarah Ball, Andreas Haupt
机构
*
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
Stanford University, Department of Computer Science, Stanford, California, USA(斯坦福大学计算机科学系)
Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy
虚拟言语治疗师:一种临床医生参与的AI言语治疗代理,用于个性化和监督式治疗
Shakeel Sheikh, Patrick Marmaroli, MD Sahidullah, Slim Ouni, Fabrice Hirsch, Goncalo Leal, Bjorn W Schuller
机构
*
The Kashmir Hub for Artficial Intelligence(喀布尔人工智能中心)
;
Microsoft / Vocametrix(微软 / Vocametrix)
;
IAI, TCG CREST(IAI,TCG CREST)
;
Université de Lorraine, CNRS, Inria, LORIA(洛林大学,CNRS,Inria,LORIA)
;
Laboratoire Praxiling, UMR5267, CNRS et Université Paul-Valéry Montpellier 3(Praxiling实验室,UMR5267,CNRS及蒙彼利埃Paul-Valéry大学)
;
Speechcare iStutter, Portuguese Catholic University(Speechcare iStutter,葡萄牙天主教大学)
;
CHI – Chair of Health Informatics, TUM University Hospital(健康信息学系,TUM大学医院)
;
GLAM – Group on Language, Audio, & Music, Imperial College London(语言、音频与音乐小组,伦敦帝国理工学院)