arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-08-05 至 2025-08-05 共收录 4 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 幻觉与事实性 4 篇

2508.02419 2025-08-05 cs.CV cs.CL 57%

Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens

Haohan Zheng, Zhenguo Zhang

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01678 2025-08-05 cs.CV cs.AI 57%

Cure or Poison? Embedding Instructions Visually Alters Hallucination in Vision-Language Models

Zhaochen Wang, Yiwei Wang, Yujun Cai

机构 * The University of Queensland(昆士兰大学) University of California, Merced(加州大学默塞德分校)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06593 2025-08-05 cs.CV 50%

SAGI: Semantically Aligned and Uncertainty Guided AI Image Inpainting

Paschalis Giakoumoglou, Dimitrios Karageorgiou, Symeon Papadopoulos, Panagiotis C. Petrantonakis

机构 * Department of Electrical and Computer Engineering, Aristotle University of Thessaloniki(阿尔伯塔大学电气与计算机工程系) Information Technologies Institute, CERTH(信息科技研究所)

专题命中 幻觉与事实性 :alignment(abstract)

Comments ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14369 2025-08-05 eess.SP 50%

Uncertainty Awareness in Wireless Communications and Sensing

Shixiong Wang, Wei Dai, Jianyong Sun, Zongben Xu, Geoffrey Ye Li

专题命中 幻觉与事实性 :trustworthy(abstract)

Journal ref IEEE Communications Magazine, July 2025

详情

展开后加载摘要…

URL PDF HTML 收藏