arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-12-16 至 2025-12-16 共收录 73 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 14 篇

2512.13169 2025-12-16 cs.MM 78%

Integrated Semantic and Temporal Alignment for Interactive Video Retrieval

集成语义与时间对齐用于交互式视频检索

Thanh-Danh Luu, Le-Vu Nguyen Dinh, Duc-Thien Tran, Duy-Bao Bui, Nam-Tien Le, Tinh-Anh Nguyen Nhu

专题命中 其他安全 :alignment(title,abstract)

AI总结 本文提出集成语义与时间对齐的交互式视频检索框架,通过QUEST和DANTE组件解决复杂现实查询问题,提升视频检索性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07128 2025-12-16 cs.CV 78%

MulCLIP: A Multi-level Alignment Framework for Enhancing Fine-grained Long-context CLIP

MulCLIP: 一种多级对齐框架,用于增强细粒度长上下文CLIP

Chau Truong, Hieu Ta Quang, Dung D. Le

机构 * FPT Software AI Center(FPT软件人工智能中心) VinUniversity(文大学)

专题命中 其他安全 :alignment(title,abstract)

AI总结 MulCLIP通过多级对齐框架提升细粒度长上下文CLIP的性能,采用标记重建和子描述聚合策略增强语义连接与上下文提取。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12950 2025-12-16 cs.CL cs.AI 62%

Building from Scratch: A Multi-Agent Framework with Human-in-the-Loop for Multilingual Legal Terminology Mapping

从零开始构建:一种带有人在回路的多智能体框架用于多语言法律术语映射

Lingyi Meng, Maolin Liu, Hao Wang, Yilan Cheng, Qi Yang, Idlkaid Mohanmmed

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种人机协作的多智能体框架,用于构建多语言法律术语数据库,通过整合AI和人类专家,提升多语言法律术语映射的精度和可扩展性。

Comments 43 pages, 6 fingures, accepted in Artificial Intelligence and Law (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19929 2025-12-16 physics.soc-ph cs.AI 57%

DynamiX: Large-Scale Dynamic Social Network Simulator

DynamiX: 大规模动态社交网络模拟器

Yanhui Sun, Wu Liu, Wentao Wang, Hantao Yao, Jiebo Luo, Yongdong Zhang

机构 * School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学) Department of Computer Science, University of Rochester(计算机科学系,罗切斯特大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 DynamiX通过动态层级模块和不同用户类型的社交关系建模策略,提升了大规模动态社交网络模拟的准确性与实用性。

Comments Social and Information Networks

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13074 2025-12-16 cs.IR cs.AI 57%

A Simple and Effective Framework for Symmetric Consistent Indexing in Large-Scale Dense Retrieval

一种用于大规模密集检索中对称一致索引的有效框架

Huimu Wang, Yiming Qiu, Xingzhi Yao, Zhiguo Chen, Guoyu Tang, Songlin Wang, Sulong Xu, Mingming Li

机构 * JD.com China(京东中国) Institute of Information Engineering, Chinese Academy of Sciences China(中国科学院信息工程研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出SCI框架,通过双塔协同和对称表示对齐,解决大规模密集检索中表示空间错位和索引不一致问题,提升检索精度与稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12907 2025-12-16 cs.LG 57%

Machine Learning Architectures for the Estimation of Predicted Occupancy Grids in Road Traffic

用于道路交通中预测占用网格估计的机器学习架构

Parthasarathy Nadarajan, Michael Botsch, Sebastian Sardina

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出了一种新颖的机器学习架构,用于高效估计道路交通中的预测占用网格,通过模拟验证其在准确性和计算时间上的性能。

Comments Journal of Advances in Information Technology

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12824 2025-12-16 cs.CV cs.AI 57%

Adapting Multimodal Foundation Models for Few-Shot Learning: A Comprehensive Study on Contrastive Captioners

为少样本学习适应多模态基础模型:对比captioners的全面研究

N. K. B. M. P. K. B. Narasinghe, Uthayasanker Thayasivam

机构 * Department of Computer Science and Engineering, University of Moratuwa, Sri Lanka(计算机科学与工程系,穆塔瓦大学,斯里兰卡)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文研究了如何通过对比captioners适应少样本学习,探讨了参数高效微调策略及生成-对比基础模型的高效适应方法。

Comments 9 pages, 3 figures. Accepted to VISAPP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12444 2025-12-16 cs.CL 57%

Can GPT replace human raters? Validity and reliability of machine-generated norms for metaphors

GPT能否替代人类评分者?机器生成的隐喻规范的有效性和可靠性

Veronica Mangiaterra, Hamad Al-Azary, Chiara Barattieri di San Pietro, Paolo Canal, Valentina Bambini

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文研究了GPT在隐喻评分中的有效性与可靠性,发现较大模型能有效替代人类评分,但需注意隐喻惯例性和多模态因素的影响。

Comments 30 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08772 2025-12-16 cs.LG 57%

De novo generation of functional terpene synthases using TpsGPT

利用 TpsGPT 从头设计功能性萜类合成酶

Hamsini Ramanathan, Roman Bushuiev, Matouš Soldát, Jirí Kohout, Téo Hebra, Joshua David Smith, Josef Sivic, Tomáš Pluskal

机构 * Seattle Academy of Arts and Sciences (SAAS)(西雅图艺术与科学学院) Czech Institute of Informatics, Robotics and Cybernetics (CIIRC)(捷克信息学、机器人学与自动控制研究所) Czech Technical University(捷克技术大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 TpsGPT 通过微调蛋白质语言模型生成功能性萜类合成酶,验证了从头设计酶的可行性。

Comments 11 pages, 8 figures, Accepted at the NeurIPS 2025 AI for Science and Machine Learning for Structural Biology 2025 workshops Fixed incorrect threshold in Fig 1

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13444 2025-12-16 cs.SE 50%

A Data Annotation Requirements Representation and Specification (DARS)

数据标注需求的表示与规范(DARS)

Yi Peng, Hina Saeeda, Hans-Martin Heyn, Jennifer Horkoff, Eric Knauss, Fredrick Warg

专题命中 其他安全 :safety(abstract)

AI总结 本文提出DARS,一种专门用于数据标注需求的表示与规范方法,通过标注协商卡片和基于场景的规范,提升数据标注的可靠性与准确性。

Comments 17 pages, 3 figures, currently submitted and under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13313 2025-12-16 cs.CV 50%

KlingAvatar 2.0 Technical Report

KlingAvatar 2.0 技术报告

Kling Team, Jialu Chen, Yikang Ding, Zhixue Fang, Kun Gai, Yuan Gao, Kang He, Jingyun Hua, Boyuan Jiang, Mingming Lao, Xiaohan Li, Hui Liu, Jiwen Liu, Xiaoqiang Liu, Yuan Liu, Shun Lu, Yongsen Mao, Yingchao Shao, Huafeng Shi, Xiaoyu Shi, Peiqin Sun, Songlin Tang, Pengfei Wan, Chao Wang, Xuebo Wang, Haoxian Zhang, Yuanxing Zhang, Yan Zhou

机构 * Kuaishou Technology(快手科技)

专题命中 其他安全 :alignment(abstract)

AI总结 KlingAvatar 2.0 通过时空级联框架和多模态指令融合技术,实现了高效且高质量的长视频生成,提升了视觉清晰度和多模态对齐能力。

Comments 14 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13004 2025-12-16 eess.SY cs.SY 50%

Large Language Models for Power System Applications: A Comprehensive Literature Survey

大语言模型在电力系统应用中的综述:全面的文献调查

Muhammad Sarwar, Muhammad Rizwan, Mubushra Aziz, Abdul Rehman Sudais

专题命中 其他安全 :safety(abstract)

AI总结 本文综述了大语言模型在电力系统中的应用,探讨了其在故障诊断、负荷预测等领域的潜力及面临的挑战,提出未来研究方向。

Comments 17 pages, 1 table, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08648 2025-12-16 cs.CV 50%

Repulsor: Accelerating Generative Modeling with a Contrastive Memory Bank

Repulsor:利用对比记忆库加速生成建模

Shaofeng Zhang, Xuanqi Chen, Ning Liao, Haoxiang Zhao, Xiaoxing Wang, Haoru Tan, Sitong Wu, Xiaosong Jia, Qi Fan, Junchi Yan

机构 * School of Artificial Intelligence and Data Science, University of Science and Technology of China(人工智能与数据科学学院,中国科学技术大学) Shanghai Jiao Tong University(上海交通大学) HKU(香港大学) CUHK(香港大学) Fudan University(复旦大学) Nanjing University(南京大学)

专题命中 其他安全 :alignment(abstract)

AI总结 Repulsor通过对比记忆库机制,无需外部编码器,实现高效的生成建模,显著提升收敛速度和生成质量。

Comments 19 pages, 19 figures

详情

展开后加载摘要…

URL PDF HTML 收藏