arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2512.15614 2025-12-18 cs.LG 57%

Behavior Tokens Speak Louder: Disentangled Explainable Recommendation with Behavior Vocabulary

行为令牌发声更大:解耦可解释推荐与行为词汇

Xinshun Feng, Mingzhe Liu, Yi Qiao, Tongyu Zhu, Leilei Sun, Shuai Wang

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 BEAT通过解耦行为词汇与语义,提升推荐系统的可解释性和零样本性能。

Comments accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12875 2025-12-17 cs.AI 57%

LTA-thinker: Latent Thought-Augmented Training Framework for Large Language Models on Complex Reasoning

LTA-thinker: 用于复杂推理的大语言模型的潜在思维增强训练框架

Jiaqi Wang, Binquan Ji, Haibo Luo, Yiyang Qi, Ruiting Li, Huiyan Wang, Yuantao Han, Cangyi Yang, jiaxu Zhang, Feiliang Ren

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 LTA-Thinker通过提升潜在思维分布方差和信息效率,优化大语言模型的复杂推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14130 2025-12-17 cs.CR cs.AI 57%

UIXPOSE: Mobile Malware Detection via Intention-Behaviour Discrepancy Analysis

基于意图-行为不一致分析的移动恶意软件检测:UIXPOSE

Amirmohammad Pasdar, Toby Murray, Van-Thuan Pham

机构 * School of Computing and Information Systems(计算与信息系统学院) The University of Melbourne(墨尔本大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 UIXPOSE通过意图-行为不一致分析提升移动恶意软件的动态检测能力

Comments 15 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13716 2025-12-17 cs.AI 57%

ValuePilot: A Two-Phase Framework for Value-Driven Decision-Making

ValuePilot:一种面向价值驱动决策的双阶段框架

Yitong Luo, Ziang Chen, Hou Hei Lam, Jiayu zhan, Junqi Wang, Zhenliang Zhang, Xue Feng

机构 * State Key Laboratory of General Artificial Intelligence, BIGAI(1 通用人工智能国家重点实验室,BIGAI) Tsinghua University(2 清华大学) Peking University(3 北京大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 ValuePilot通过双阶段框架实现价值驱动的个性化决策,提升AI在复杂场景中的可解释性和适应性。

Comments Accepted at LAW Workshop, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19929 2025-12-16 physics.soc-ph cs.AI 57%

DynamiX: Large-Scale Dynamic Social Network Simulator

DynamiX: 大规模动态社交网络模拟器

Yanhui Sun, Wu Liu, Wentao Wang, Hantao Yao, Jiebo Luo, Yongdong Zhang

机构 * School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学) Department of Computer Science, University of Rochester(计算机科学系,罗切斯特大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 DynamiX通过动态层级模块和不同用户类型的社交关系建模策略,提升了大规模动态社交网络模拟的准确性与实用性。

Comments Social and Information Networks

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13074 2025-12-16 cs.IR cs.AI 57%

A Simple and Effective Framework for Symmetric Consistent Indexing in Large-Scale Dense Retrieval

一种用于大规模密集检索中对称一致索引的有效框架

Huimu Wang, Yiming Qiu, Xingzhi Yao, Zhiguo Chen, Guoyu Tang, Songlin Wang, Sulong Xu, Mingming Li

机构 * JD.com China(京东中国) Institute of Information Engineering, Chinese Academy of Sciences China(中国科学院信息工程研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出SCI框架,通过双塔协同和对称表示对齐,解决大规模密集检索中表示空间错位和索引不一致问题,提升检索精度与稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12907 2025-12-16 cs.LG 57%

Machine Learning Architectures for the Estimation of Predicted Occupancy Grids in Road Traffic

用于道路交通中预测占用网格估计的机器学习架构

Parthasarathy Nadarajan, Michael Botsch, Sebastian Sardina

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出了一种新颖的机器学习架构,用于高效估计道路交通中的预测占用网格,通过模拟验证其在准确性和计算时间上的性能。

Comments Journal of Advances in Information Technology

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12824 2025-12-16 cs.CV cs.AI 57%

Adapting Multimodal Foundation Models for Few-Shot Learning: A Comprehensive Study on Contrastive Captioners

为少样本学习适应多模态基础模型:对比captioners的全面研究

N. K. B. M. P. K. B. Narasinghe, Uthayasanker Thayasivam

机构 * Department of Computer Science and Engineering, University of Moratuwa, Sri Lanka(计算机科学与工程系,穆塔瓦大学,斯里兰卡)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文研究了如何通过对比captioners适应少样本学习,探讨了参数高效微调策略及生成-对比基础模型的高效适应方法。

Comments 9 pages, 3 figures. Accepted to VISAPP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12444 2025-12-16 cs.CL 57%

Can GPT replace human raters? Validity and reliability of machine-generated norms for metaphors

GPT能否替代人类评分者?机器生成的隐喻规范的有效性和可靠性

Veronica Mangiaterra, Hamad Al-Azary, Chiara Barattieri di San Pietro, Paolo Canal, Valentina Bambini

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文研究了GPT在隐喻评分中的有效性与可靠性,发现较大模型能有效替代人类评分,但需注意隐喻惯例性和多模态因素的影响。

Comments 30 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08772 2025-12-16 cs.LG 57%

De novo generation of functional terpene synthases using TpsGPT

利用 TpsGPT 从头设计功能性萜类合成酶

Hamsini Ramanathan, Roman Bushuiev, Matouš Soldát, Jirí Kohout, Téo Hebra, Joshua David Smith, Josef Sivic, Tomáš Pluskal

机构 * Seattle Academy of Arts and Sciences (SAAS)(西雅图艺术与科学学院) Czech Institute of Informatics, Robotics and Cybernetics (CIIRC)(捷克信息学、机器人学与自动控制研究所) Czech Technical University(捷克技术大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 TpsGPT 通过微调蛋白质语言模型生成功能性萜类合成酶,验证了从头设计酶的可行性。

Comments 11 pages, 8 figures, Accepted at the NeurIPS 2025 AI for Science and Machine Learning for Structural Biology 2025 workshops Fixed incorrect threshold in Fig 1

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11661 2025-12-15 cs.HC cs.AI 57%

From Verification Burden to Trusted Collaboration: Design Goals for LLM-Assisted Literature Reviews

从验证负担到可信协作:LLM辅助文献综述的设计目标

Brenda Nogueira, Werner Geyer, Andrew Anderson, Toby Jia-Jun Li, Dongwhi Kim, Nuno Moniz, Nitesh V. Chawla

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文针对LLM辅助文献综述中的信任和协作问题,提出六个设计目标和框架,通过可视化、验证和反馈对齐提升可信度与协作效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11296 2025-12-15 cs.CV cs.AI cs.HC 57%

Few-Shot VLM-Based G-Code and HMI Verification in CNC Machining

少样本基于视觉语言模型的CNC加工G代码和人机界面验证

Yasaman Hashem Pour, Nazanin Mahjourian, Vinh Nguyen

机构 * Department of Mechanical and Aerospace Engineering, Michigan Technological University(机械与航空航天工程系,密歇根技术大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文提出基于视觉语言模型的少样本方法,用于验证CNC加工中G代码与人机界面的错误和安全状态。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10435 2025-12-12 cs.CL 57%

Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature

对抗性剽窃的语义重建:一种基于上下文的框架,用于检测和恢复科学文献中的‘扭曲短语’

Agniva Maiti, Prajwal Panth, Suresh Chandra Satapathy

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出SRAP框架,通过双阶段架构检测并恢复科学文献中的对抗性剽窃,显著提升恢复准确率。

Comments 10 pages, 5 figures; unpublished manuscript; submitted to arXiv for dissemination

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09898 2025-12-11 cs.RO cs.AI cs.CV cs.MA cs.SY eess.SY 57%

Visual Heading Prediction for Autonomous Aerial Vehicles

自主航空器的视觉航向预测

Reza Ahmari, Ahmad Mohammadi, Vahid Hemmati, Mohammed Mynuddin, Parham Kebria, Mahmoud Nabil Mahmoud, Xiaohong Yuan, Abdollah Homaifar

机构 * Department of Computer Science at North Carolina A&T State University(北卡罗来纳A&T州立大学计算机科学系) Department of Electrical and Computer Engineering at North Carolina A&T State University(北卡罗来纳A&T州立大学电气与计算机工程系)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出基于视觉的数据驱动框架,实现无人机与无人地面车辆的实时整合,通过YOLOv5检测UGV并利用轻量级ANN预测航向角,实现高精度的导航与协调。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09895 2025-12-11 cs.AI cs.DL 57%

Human-in-the-Loop and AI: Crowdsourcing Metadata Vocabulary for Materials Science

人机协同与AI:为材料科学众包元数据词汇

Jane Greenberg, Scott McClellan, Addy Ireland, Robert Sammarco, Colton Gerber, Christopher B. Rauch, Mat Kelly, John Kunze, Yuan An, Eric Toberer

机构 * Metadata Research Center, College of Computing and Informatics, Drexel University(元数据研究中心、计算与信息学院、德雷塞尔大学) Penn State University(宾夕法尼亚州立大学) Colorado School of Mines(科罗拉多矿业学院)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出MatSci-YAMZ平台,通过结合人工智能和人机协同方法,实现材料科学元数据词汇表的众包开发,验证了AI-HILT模型在跨学科领域的可行性与可扩展性。

Comments Metadata and Semantics Research Conference 2025, 14 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08143 2025-12-11 cs.LG 57%

PolyLingua: Margin-based Inter-class Transformer for Robust Cross-domain Language Detection

PolyLingua: 基于边界的跨领域语言识别变换器

Ali Lotfi Rezaabad, Bikram Khanal, Shashwat Chaurasia, Lu Zeng, Dezhi Hong, Hossein Bashashati, Thomas Butler, Megan Ganji

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 PolyLingua是一种轻量级变换器模型,通过两级对比学习框架实现精确的跨领域语言识别,以高准确率和低资源消耗应对复杂语言识别挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.05664 2025-12-11 cs.RO cs.AI 57%

Altruistic Maneuver Planning for Cooperative Autonomous Vehicles Using Multi-agent Advantage Actor-Critic

为合作自主车辆的利他性动作规划使用多智能体优势Actor-Critic

Behrad Toghi, Rodolfo Valiente, Dorsa Sadigh, Ramtin Pedarsani, Yaser P. Fallah

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文提出一种多智能体优势Actor-Critic算法,用于自动驾驶车辆在混合交通环境中的利他性动作规划,以提升交通效率与安全。

Comments Accepted to 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2021) - Workshop on Autonomous Driving: Perception, Prediction and Planning

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08524 2025-12-10 cs.CV cs.CL 57%

Beyond Real Weights: Hypercomplex Representations for Stable Quantization

超越真实权重:用于稳定量化 的超复数表示

Jawad Ibn Ahad, Maisha Rahman, Amrijit Biswas, Muhammad Rafsan Kabir, Robin Krambroeckers, Sifat Momen, Nabeel Mohammed, Shafin Rahman

机构 * Artificial Intelligence Department, RobotBulls Labs(机器人bulls实验室人工智能部门) Machine Intelligence Lab (MILab), North South University(北南大学机器智能实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出了一种基于超复数乘法的渐进式重新参数化策略,用于压缩多模态语言模型,实现参数和计算量的显著减少,同时保持模型性能。

Comments Accepted in Winter Conference on Applications of Computer Vision (WACV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00218 2025-12-10 cs.AI cs.CR 57%

Reasoning Under Pressure: How do Training Incentives Influence Chain-of-Thought Monitorability?

压力下的推理:训练激励如何影响推理链的可监控性?

Matt MacDermott, Qiyao Wei, Rada Djoneva, Francis Rhys Ward

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文研究了训练激励对推理链可监控性的影响,发现对抗性优化降低监控性能,而直接优化可监控性未显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22884 2025-12-10 cs.DC cs.AI cs.ET cs.NI cs.SY eess.SY 57%

Performance Measurements in the AI-Centric Computing Continuum Systems

面向AI导向计算连续体系统的性能测量

Praveen Kumar Donta, Qiyang Zhang, Schahram Dustdar

机构 * Department of Computer Systems and Sciences(计算机系统科学系) Stockholm University(斯德哥尔摩大学) Computer Science School(计算机科学学院) Peking University(北京大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文探讨了AI导向计算连续体系统中性能测量的挑战与方法,提出新的性能维度以适应不断变化的计算需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07694 2025-12-09 cs.CL 57%

Automated Generation of Custom MedDRA Queries Using SafeTerm Medical Map

基于SafeTerm医学地图的自动生成定制MedDRA查询

Francois Vandenhende, Anna Georgiou, Michalis Georgiou, Theodoros Psaras, Ellie Karekla, Elena Hadjicosta

机构 * ClinBAY Limited(ClinBAY有限公司)

专题命中 其他安全 :safety(abstract);分类 cs.CL

AI总结 SafeTerm系统通过多维向量空间和相似度计算,实现自动生成MedDRA查询,优化查询生成的精度与召回率。

Comments 12 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07552 2025-12-09 cs.CL 57%

Performance of the SafeTerm AI-Based MedDRA Query System Against Standardised MedDRA Queries

SafeTerm基于AI的MedDRA查询系统在标准化MedDRA查询中的表现

Francois Vandenhende, Anna Georgiou, Michalis Georgiou, Theodoros Psaras, Ellie Karekla, Elena Hadjicosta

机构 * ClinBAY Limited(ClinBAY有限公司)

专题命中 其他安全 :safety(abstract);分类 cs.CL

AI总结 SafeTerm基于AI的MedDRA查询系统在标准化MedDRA查询中表现出良好性能,通过多标准统计方法实现高召回率和精度平衡。

Comments 8 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06837 2025-12-09 cs.LG 57%

Neural Factorization-based Bearing Fault Diagnosis

基于神经分解的轴承故障诊断

Zhenhao Li, Xu Cheng, Yi Zhou

机构 * College of Computer and Information Science, Southwest University, Chongqing, China(计算机与信息科学学院,西南大学,重庆,中国) College of Vehicle Engineering, Chongqing Industry and Trade Polytechnic, Chongqing, China(车辆工程学院,重庆工业贸易职业技术学院,重庆,中国)

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出基于神经分解的轴承故障诊断框架,通过多模式特征嵌入和神经分解融合,提升复杂条件下故障诊断性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06787 2025-12-09 cs.CL 57%

LLM4SFC: Sequential Function Chart Generation via Large Language Models

LLM4SFC: 通过大语言模型生成顺序功能图

Ofek Glick, Vladimir Tchuiev, Marah Ghoummaid, Michal Moshkovitz, Dotan Di-Castro

机构 * Bosch Research(博世研究)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 LLM4SFC通过大语言模型生成可执行的顺序功能图,实现图形与文本PLC语言的高效转换。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06688 2025-12-09 cs.CL 57%

PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User Personas and Agentic Memory

PersonaMem-v2:通过学习隐式用户人设和代理记忆实现个性化智能

Bowen Jiang, Yuan Yuan, Maohao Shen, Zhuoqun Hao, Zhangchen Xu, Zichen Chen, Ziyi Liu, Anvesh Rao Vijjini, Jiashu He, Hanchao Yu, Radha Poovendran, Gregory Wornell, Lyle Ungar, Dan Roth, Sihao Chen, Camillo Jose Taylor

机构 * University of Pennsylvania(宾夕法尼亚大学) Massachusetts Institute of Technology(麻省理工学院) University of Washington(华盛顿大学) University of California Santa Barbara(加州大学圣巴巴拉分校) Meta University of Southern California(南加州大学) University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) Microsoft Corporation(微软公司)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 PersonaMem-v2通过学习隐式用户人设和代理记忆提升LLM个性化能力,实验显示强化微调使模型在隐式个性化任务中准确率达53%。

Comments Data is available at https://huggingface.co/datasets/bowen-upenn/PersonaMem-v2

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05502 2025-12-08 cs.LG 57%

GRASP: Graph Reasoning Agents for Systems Pharmacology with Human-in-the-Loop

GRASP:具有人机交互循环的图推理代理用于系统药理学

Omid Bazgir, Vineeth Manthapuri, Ilia Rattsev, Mohammad Jafarnejad

机构 * Clinical Pharmacology, Genentech(基因泰克临床药理部) Preclinical & Translational PKPD, Genentech Inc.(基因泰克预临床与转化药代动力学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 GRASP通过图推理代理实现系统药理学模型开发的自动化与严谨性,提升生物医学建模的效率和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05365 2025-12-08 cs.AI q-bio.QM 57%

MCP-AI: Protocol-Driven Intelligence Framework for Autonomous Reasoning in Healthcare

MCP-AI:面向医疗领域自主推理的协议驱动智能框架

Zag ElSayed, Craig Erickson, Ernest Pedapati

机构 * School of Information Technology University of Cincinnati Ohio, USA(信息科技学院 俄亥俄州立大学 奥哈伊俄州) Adolescent Psychiatry Cincinnati Children’s Hospital Medical Center Ohio, USA(青少年精神病学 奥克兰儿童医院医疗中心 奥哈伊俄州)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 MCP-AI是一种基于模型上下文协议的医疗自主推理框架,通过整合临床逻辑和安全协作,提升医疗决策的可解释性和适应性。

Comments 6 pages, 4 figures

Journal ref IEEE ICMLA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05226 2025-12-08 cs.LG 57%

Variance Matters: Improving Domain Adaptation via Stratified Sampling

方差至关重要:通过分层抽样改进领域适应

Andrea Napoli, Paul White

机构 * Institute of Sound and Vibration Research(声学与振动研究所) University of Southampton(南安普顿大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 本文提出VaRDASS方法,通过分层抽样减少领域适应中的方差,提升领域差异估计精度和目标领域性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20395 2025-12-05 cs.LG 57%

Identifying environmental factors associated with tetrodotoxin contamination in bivalve mollusks using eXplainable AI

利用可解释AI识别与双壳类软体动物中四氢大麻酚污染相关的环境因素

M. C. Schoppema, B. H. M. van der Velden, A. Hürriyetoğlu, M. D. Klijnstra, E. J. Faassen, A. Gerssen, H. J. van der Fels-Klerx

机构 * Wageningen Food Safety Research(瓦赫宁根食品安全研究所)

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本研究利用可解释深度学习模型识别出太阳小时数、全球辐射、水温和水氯离子浓度是影响双壳类软体动物中TTX污染的关键环境因素。

Comments 18 pages, 6 figures, submitted to npj Science of Food

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03976 2025-12-04 cs.CL 57%

Adapting Large Language Models to Low-Resource Tibetan: A Two-Stage Continual and Supervised Fine-Tuning Study

将大型语言模型适应于低资源藏语:一种两阶段持续和监督微调研究

Lifeng Chen, Ryan Lai, Tianming Liu

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本研究通过两阶段方法将Qwen2.5-3B适应到藏语,通过持续预训练和监督微调提升翻译质量,实现低资源语言的模型适应。

详情

展开后加载摘要…

URL PDF HTML 收藏