arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1755 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 幻觉与事实性 1755 篇

2505.07437 2025-05-13 cs.LG cs.AI cs.DB 62%

LEAD: Iterative Data Selection for Efficient LLM Instruction Tuning

Xiaotian Lin, Yanlin Qi, Yizhang Zhu, Themis Palpanas, Chengliang Chai, Nan Tang, Yuyu Luo

机构 * HKUST (GZ)(香港科技大学(广州)) Université Paris Cité(巴黎Cité大学) BIT(北京理工大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06680 2025-05-13 cs.AI cs.HC cs.LG cs.SY eess.SY physics.soc-ph 62%

A Survey on Data-Driven Modeling of Human Drivers' Lane-Changing Decisions

Linxuan Huang, Dong-Fan Xie, Li Li, Zhengbing He

机构 * School of Systems Science, Beijing Jiaotong University(北京交通大学系统科学学院) Department of Automation, BNRist, Tsinghua University(清华大学自动化系) Laboratory for Information and Decision Systems, Massachusetts Institute of Technology(麻省理工学院信息与决策系统实验室)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04651 2025-05-09 cs.CL cs.LG 62%

Scientific Hypothesis Generation and Validation: Methods, Datasets, and Future Directions

Adithya Kulkarni, Fatimah Alotaibi, Xinyue Zeng, Longfeng Wu, Tong Zeng, Barry Menglong Yao, Minqian Liu, Shuaicheng Zhang, Lifu Huang, Dawei Zhou

机构 * Virginia Tech(弗吉尼亚理工大学) University of California, Davis(加州大学戴维斯分校)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02874 2025-05-07 cs.LG cs.AI 62%

Uncertainty Quantification for Machine Learning in Healthcare: A Survey

L. Julián Lechuga López, Shaza Elsharief, Dhiyaa Al Jorf, Firas Darwish, Congbo Ma, Farah E. Shamout

机构 * New York University(纽约大学) New York University Abu Dhabi(纽约大学阿布扎赫尔分校)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI、cs.LG

Comments 46 pages, 3 figures, 2 tables, AHLI Conference on Health, Inference, and Learning (CHIL)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00917 2025-05-05 stat.ME cs.AI cs.LG stat.ML 62%

Multivariate Conformal Selection

Tian Bai, Yue Zhao, Xiang Yu, Archer Y. Yang

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI、cs.LG

Comments 25 pages, 4 figures. Accepted to ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00557 2025-05-02 cs.CL cs.AI 62%

Triggering Hallucinations in LLMs: A Quantitative Study of Prompt-Induced Hallucination in Large Language Models

Makoto Sato

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09527 2025-04-23 cs.LG cs.CL 62%

Confidence Estimation for Error Detection in Text-to-SQL Systems

Oleg Somov, Elena Tutubalina

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.LG

Comments 15 pages, 11 figures, to be published in AAAI 2025 Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11986 2025-04-22 cs.CL cs.AI 62%

Large Language Models as Quasi-crystals: Coherence Without Repetition in Generative Text

Jose Manuel Guevara-Vela

机构 * School of Engineering and Physical Sciences, Heriot-Watt University(工程与物理科学学院,赫瑞斯泰大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

Comments The discussion was restructured to add limitations to the analogy and other clarifications

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13644 2025-04-21 cs.AI cs.CL 62%

Exploring the Potential for Large Language Models to Demonstrate Rational Probabilistic Beliefs

Gabriel Freedman, Francesca Toni

机构 * Department of Computing, Imperial College London, UK(计算系,帝国理工学院伦敦分校)

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.CL、cs.AI

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20826 2025-04-16 cs.CV cs.CL cs.LG eess.IV 62%

Exploring CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation

Zhiwei Yang, Yucong Meng, Kexue Fu, Feilong Tang, Shuo Wang, Zhijian Song

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.LG

Comments CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06465 2025-04-11 cs.CL cs.AI 62%

MedCT: A Clinical Terminology Graph for Generative AI Applications in Healthcare

Ye Chen, Dongdong Huang, Haoyun Xu, Cong Fu, Lin Sheng, Qingli Zhou, Yuqiang Shen, Kai Wang

专题命中 幻觉与事实性 :safety(abstract);分类 cs.CL、cs.AI

Comments Accepted into ICCS 2025 and published in Springer's LNCS Series

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13622 2025-04-09 cs.CL cs.AI 62%

REFIND at SemEval-2025 Task 3: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models

DongGeon Lee, Hwanjo Yu

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.CL、cs.AI

Comments Accepted to SemEval@ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03640 2025-04-07 cs.CL cs.AI cs.CV 62%

Bonsai: Interpretable Tree-Adaptive Grounded Reasoning

Kate Sanders, Benjamin Van Durme

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

Comments 9 pages, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03211 2025-04-07 cs.LG cs.AI cs.GT econ.TH 62%

Persuasive Calibration

Yiding Feng, Wei Tang

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14582 2025-03-31 cs.AI cs.CL 62%

Do LLMs estimate uncertainty well in instruction-following?

Juyeon Heo, Miao Xiong, Christina Heinze-Deml, Jaya Narain

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09283 2025-03-20 cs.CL cs.AI 62%

DAHRS: Divergence-Aware Hallucination-Remediated SRL Projection

Sangpil Youm, Brodie Mather, Chathuri Jayaweera, Juliana Prada, Bonnie Dorr

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

Comments 15 pages, 6 figures, Accepted to The 29th International Conference on Natural Language & Information Systems (NLDB 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14675 2025-03-18 cs.CL cs.AI 62%

To Trust or Not to Trust? Enhancing Large Language Models' Situated Faithfulness to External Contexts

Yukun Huang, Sanxing Chen, Hongyi Cai, Bhuwan Dhingra

专题命中 幻觉与事实性 :DPO(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11283 2025-03-10 cs.CL cs.AI 62%

Zero-resource Hallucination Detection for Text Generation via Graph-based Contextual Knowledge Triples Modeling

Xinyue Fang, Zhen Huang, Zhiliang Tian, Minghui Fang, Ziyi Pan, Quntian Fang, Zhihua Wen, Hengyue Pan, Dongsheng Li

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

Comments Accepted by AAAI25

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10709 2025-03-04 cs.CL cs.AI 62%

An Empirical Analysis of Uncertainty in Large Language Model Evaluations

Qiujie Xie, Qingqiu Li, Zhuohao Yu, Yuejie Zhang, Yue Zhang, Linyi Yang

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16896 2025-02-25 cs.LG cs.AI 62%

Zero-shot Load Forecasting for Integrated Energy Systems: A Large Language Model-based Framework with Multi-task Learning

Jiaheng Li, Donghe Li, Ye Yang, Huan Xi, Yu Xiao, Li Sun, Dou An, Qingyu Yang

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03349 2025-02-18 physics.geo-ph cs.AI cs.LG physics.ao-ph 62%

When Geoscience Meets Generative AI and Large Language Models: Foundations, Trends, and Future Challenges

Abdenour Hadid, Tanujit Chakraborty, Daniel Busby

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI、cs.LG

Journal ref Expert Systems, 2024, Volume: 41, Issue: 10

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08909 2025-02-14 cs.CL cs.AI 62%

Towards Automated Fact-Checking of Real-World Claims: Exploring Task Formulation and Assessment with LLMs

Premtim Sahitaj, Iffat Maab, Junichi Yamagishi, Jawan Kolanowski, Sebastian Möller, Vera Schmitt

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05244 2025-02-11 cs.AI cs.LG 62%

Probabilistic Artificial Intelligence

Andreas Krause, Jonas Hübotter

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01691 2025-02-05 cs.CL cs.AI 62%

Agent-Based Uncertainty Awareness Improves Automated Radiology Report Labeling with an Open-Source Large Language Model

Hadas Ben-Atya, Naama Gavrielov, Zvi Badash, Gili Focht, Ruth Cytter-Kuint, Talar Hagopian, Dan Turner, Moti Freiman

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.15264 2025-02-03 cs.CL cs.AI 62%

ReXTrust: A Model for Fine-Grained Hallucination Detection in AI-Generated Radiology Reports

Romain Hardy, Sung Eun Kim, Du Hyun Ro, Pranav Rajpurkar

专题命中 幻觉与事实性 :safety(abstract);分类 cs.CL、cs.AI

Comments Accepted to AIMedHealth 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14719 2025-01-27 cs.CL cs.AI cs.HC cs.IR 62%

Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?

Ipek Baris Schlicht, Zhixue Zhao, Burcu Sayin, Lucie Flek, Paolo Rosso

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL、cs.AI

Comments 9 pages. Short paper appeared at 47th European Conference on Information Retrieval (ECIR 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13412 2025-01-24 cs.LG cs.AI 62%

Load and Renewable Energy Forecasting Using Deep Learning for Grid Stability

Kamal Sarkar

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI、cs.LG

Journal ref Proceedings of How AI can help in Green Energy, 24 Dec 2024, Science City , Kolkata, India

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04577 2025-01-24 cs.AR cs.AI cs.LG cs.RO 62%

A 65 nm Bayesian Neural Network Accelerator with 360 fJ/Sample In-Word GRNG for AI Uncertainty Estimation

Zephan M. Enciso, Boyang Cheng, Likai Pei, Jianbo Liu, Steven Davis, Michael Niemier, Ningyuan Cao

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI、cs.LG

Comments 7 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08188 2025-01-15 cs.CV cs.AI cs.LG 62%

A Critical Synthesis of Uncertainty Quantification and Foundation Models in Monocular Depth Estimation

Steven Landgraf, Rongjun Qin, Markus Ulrich

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03295 2025-01-09 cs.LG cs.AI eess.SP 62%

A Soft Sensor Method with Uncertainty-Awareness and Self-Explanation Based on Large Language Models Enhanced by Domain Knowledge Retrieval

Shuo Tong, Han Liu, Runyuan Guo, Wenqing Wang, Xueqiong Tian, Lingyun Wei, Lin Zhang, Huayong Wu, Ding Liu, Youmin Zhang

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏