arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2026-01-22 至 2026-01-22 共收录 40 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 14 篇

2601.14777 2026-01-22 cs.CV cs.AI 57%

FunCineForge: A Unified Dataset Toolkit and Model for Zero-Shot Movie Dubbing in Diverse Cinematic Scenes

FunCineForge: 一个统一的数据集工具包和模型,用于多样的影视场景零样本配音

Jiaxuan Liu, Yang Xiang, Han Zhao, Xiangang Li, Zhenhua Ling

机构 * Alibaba Group(阿里巴巴集团) Tongyi Lab(通义实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 FunCineForge提出了一种统一的数据集工具包和模型,用于多样的影视场景零样本配音,通过构建大规模数据集和基于大规模语言模型的模型,提升了音频质量、唇形同步和情绪表达性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14700 2026-01-22 cs.CL 57%

DARL: Encouraging Diverse Answers for General Reasoning without Verifiers

DARL: 促进无验证者的一般推理中的多样化答案

Chongxuan Huang, Lei Lin, Xiaodong Shi, Wenping Hu, Ruiming Tang

机构 * School of Informatics, Xiamen University(厦门大学信息学院) Kuaishou Technology(快手科技) Key Laboratory of Digital Protection and Intelligent Processing of Intangible Cultural Heritage of Fujian and Taiwan (Xiamen University), Ministry of Culture and Tourism(福建省和台湾非物质文化遗产数字化保护与智能处理重点实验室(厦门大学),文化和旅游部)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 DARL通过鼓励在参考答案可控偏差范围内生成多样化答案,提升大语言模型在一般推理任务中的性能和输出多样性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14679 2026-01-22 cs.MM cs.AI 57%

HCVR Scene Generation: High Compatibility Virtual Reality Environment Generation for Extended Redirected Walking

HCVR场景生成:为扩展定向行走设计的高兼容性虚拟现实环境生成

Yiran Zhang, Xingpeng Sun, Aniket Bera

机构 * Purdue University(普渡大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 HCVR通过优化虚拟场景与物理空间的兼容性,显著减少虚拟现实中的物理碰撞,提升扩展定向行走的沉浸体验。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14560 2026-01-22 cs.CL 57%

Rewarding How Models Think Pedagogically: Integrating Pedagogical Reasoning and Thinking Rewards for LLMs in Education

奖励模型如何思考:在教育中整合教学推理和思考奖励

Unggi Lee, Jiyeong Bae, Jaehyeon Park, Haeun Park, Taejun Park, Younghoon Jeon, Sungmin Cho, Junbo Koh, Yeil Jeong, Gyeonggeon Lee

机构 * Chosun University(昌原大学) Korea University(韩国大学) Seoul National University(首尔国立大学) Korea Institute for Curriculum and Evaluation(韩国课程与评价研究院) Upstage Indiana University Bloomington(印第安纳大学布卢明顿分校) Nanyang Technological University(南洋理工大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出PedagogicalRL-Thinking框架,通过教学推理提示和思考奖励方法,提升LLM在教育场景中的教学推理能力和结构化决策能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14261 2026-01-22 cs.CV cs.HC cs.LG 57%

Intelligent Power Grid Design Review via Active Perception-Enabled Multimodal Large Language Models

通过主动感知多模态大语言模型实现智能电网设计审查

Taoliang Tan, Chengwei Ma, Zhen Tian, Zhao Lin, Dongdong Li, Si Shi

机构 * Yangjiang Yangxi Power Supply Bureau, Guangdong Power Grid Co., Ltd., Yangjiang, China(阳江阳西供电局,广东电网公司) Guangdong Laboratory of Artificial Intelligence(广东人工智能实验室)

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出基于主动感知多模态大语言模型的智能电网设计审查方法,通过三阶段框架提升对设计错误的识别准确性和审查可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19699 2026-01-22 cs.NI cs.AI cs.MA 57%

A Layered Protocol Architecture for the Internet of Agents

面向智能体的分层协议架构

Charles Fleming, Luca Muscariello, Vijoy Pandey, Ramana Kompella

机构 * Cisco Research(思科研究)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出分层协议架构,通过智能体通信层和语义层实现智能体协作,解决LLMs在内存和计算能力上的限制,推动多智能体系统的扩展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07301 2026-01-22 cs.CV cs.AI 57%

Beyond Boundaries: Leveraging Vision Foundation Models for Source-Free Object Detection

超越边界:利用视觉基础模型进行无源目标检测

Huizai Yao, Sicheng Zhao, Pengteng Li, Yi Cui, Shuo Lu, Weiyu Guo, Yunfan Lu, Yijie Xu, Hui Xiong

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出一种利用视觉基础模型提升无源目标检测性能的新框架,通过增强特征对齐和标签质量,实现跨领域迁移和辨别性提升。

Comments Accepted to AAAI 2026. Extended version with full Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22950 2026-01-22 cs.CL 57%

StrucSum: Graph-Structured Reasoning for Long Document Extractive Summarization with LLMs

StrucSum: 图结构推理用于基于LLM的长文档提取式摘要

Haohan Yuan, Sukhwa Hong, Haopeng Zhang

机构 * ALOHA Lab, University of Hawaii at Manoa(夏威夷大学曼纳亚分校ALOHA实验室) University of Hawaii at Hilo(夏威夷大学希洛分校)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 StrucSum通过图结构推理提升基于LLM的长文档提取式摘要质量,有效增强文档结构建模与关键信息识别能力。

Comments 14 pages. Accepted by the findings of EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11910 2026-01-22 cs.CV 50%

A Training-Free Guess What Vision Language Model from Snippets to Open-Vocabulary Object Detection

无需训练的Guess What视觉语言模型:从片段到开放词汇物体检测

Guiying Zhu, Bowen Yang, Yin Zhuang, Tong Zhang, Guanqun Wang, Zhihao Che, He Chen, Lianlin Li

机构 * Aerospace and Informatics Domain(航空航天与信息领域) National Key Laboratory of Science and Technology on Space-Born Intelligent Information Processing(空间智能信息处理国家级重点实验室) School of Electronic(电子学院)

专题命中 其他安全 :alignment(abstract)

AI总结 本文提出无需训练的Guess What视觉语言模型GW-VLM,通过多尺度视觉语言搜索与上下文概念提示实现开放词汇物体检测的高效检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10958 2026-01-22 cs.IT cs.NI math.IT 50%

Fundamental Limits of Quantum Semantic Communication via Sheaf Cohomology

量子语义通信的fundamental limits via sheaf cohomology

Christo Kurisummoottil Thomas, Mingzhe Chen

专题命中 其他安全 :alignment(abstract)

AI总结 本文提出基于sheaf cohomology的量子语义通信框架,揭示语义模糊性的信息论限制,并通过量子纠缠和情境性降低上同调障碍,为自主系统提供新的通信理论基础。

详情

展开后加载摘要…

URL PDF HTML 收藏