arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2508.11404 2025-08-18 cs.RO cs.AI cs.HC 57%

An Exploratory Study on Crack Detection in Concrete through Human-Robot Collaboration

Junyeon Kim, Tianshu Ruan, Cesar Alan Contreras, Manolis Chiou

机构 * Extreme Robotics Lab (ERL) and National Center for Nuclear Robotics (NCNR)(极端机器人实验室(ERL)和核机器人国家中心(NCNR)) University of Birmingham(伯明翰大学) National Center for Nuclear Robotics (NCNR)(核机器人国家中心) Queen Mary University of London(伦敦女王大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10919 2025-08-18 cs.HC cs.AI 57%

Human-AI collaboration or obedient and often clueless AI in instruct, serve, repeat dynamics?

Mohammed Saqr, Kamila Misiejuk, Sonsoles López-Pernas

机构 * University of Eastern Finland(东方芬兰大学) FernUniversität in Hagen(哈根弗伦大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02269 2025-08-18 cs.AI 57%

AirTrafficGen: Configurable Air Traffic Scenario Generation with Large Language Models

Dewi Sid William Gould, George De Ath, Ben Carvell, Nick Pepper

机构 * The Alan Turing Institute(阿尔ัน图灵研究院) NATS University of Exeter(埃克塞特大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments 9 pages and appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10349 2025-08-15 cs.DC cs.LG 57%

Flexible Personalized Split Federated Learning for On-Device Fine-Tuning of Foundation Models

Tianjun Yuan, Jiaxiang Geng, Pengchao Han, Xianhao Chen, Bing Luo

机构 * Duke Kunshan University(杜克大学昆山学院) The University of Hong Kong(香港大学) Guangdong University of Technology(广东技术大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments 10 pages, Submitted to INFOCOM2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10104 2025-08-15 cs.CV cs.LG 57%

DINOv3

Oriane Siméoni, Huy V. Vo, Maximilian Seitzer, Federico Baldassarre, Maxime Oquab, Cijo Jose, Vasil Khalidov, Marc Szafraniec, Seungeun Yi, Michaël Ramamonjisoa, Francisco Massa, Daniel Haziza, Luca Wehrstedt, Jianyuan Wang, Timothée Darcet, Théo Moutakanni, Leonel Sentana, Claire Roberts, Andrea Vedaldi, Jamie Tolan, John Brandt, Camille Couprie, Julien Mairal, Hervé Jégou, Patrick Labatut, Piotr Bojanowski

机构 * Meta AI Research(Meta AI研究院) WRI(世界知识产权组织) Inria(法国国家信息与自动化研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10012 2025-08-15 cs.CL 57%

Guided Navigation in Knowledge-Dense Environments: Structured Semantic Exploration with Guidance Graphs

Dehao Tao, Guangjie Liu, Weizheng, Yongfeng Huang, Minghu jiang

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09919 2025-08-14 eess.IV cs.AI cs.CV 57%

T-CACE: A Time-Conditioned Autoregressive Contrast Enhancement Multi-Task Framework for Contrast-Free Liver MRI Synthesis, Segmentation, and Diagnosis

Xiaojiao Xiao, Jianfeng Zhao, Qinmin Vivian Hu, Guanghui Wang

机构 * Department of Computer Science, Toronto Metropolitan University(计算机科学系,多伦多 Metropolitan 大学) School of Biomedical Engineering, Western University(生物医学工程学院,西部大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments IEEE Journal of Biomedical and Health Informatics, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23077 2025-08-14 cs.CL 57%

Efficient Inference for Large Reasoning Models: A Survey

Yue Liu, Jiaying Wu, Yufei He, Ruihan Gong, Jun Xia, Liang Li, Hongcheng Gao, Hongyu Chen, Baolong Bi, Jiaheng Zhang, Zhiqi Huang, Bryan Hooi, Stan Z. Li, Keqin Li

专题命中 其他安全 :safety(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09453 2025-08-14 cs.CV cs.LG 57%

HyperKD: Distilling Cross-Spectral Knowledge in Masked Autoencoders via Inverse Domain Shift with Spatial-Aware Masking and Specialized Loss

Abdul Matin, Tanjim Bin Faruk, Shrideep Pallickara, Sangmi Lee Pallickara

机构 * Colorado State University(科罗拉多州立大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08678 2025-08-13 cs.CY 57%

Exploring Large Language Model Agents for Piloting Social Experiments

Jinghua Piao, Yuwei Yan, Nian Li, Jun Zhang, Yong Li

专题命中 其他安全 :alignment(abstract);分类 cs.CY

Comments Accepted by COLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08011 2025-08-12 cs.CL 57%

Progressive Depth Up-scaling via Optimal Transport

Mingzi Cao, Xi Wang, Nikolaos Aletras

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07885 2025-08-12 cs.RO cs.AI cs.CV cs.SY eess.SY 57%

Autonomous Navigation of Cloud-Controlled Quadcopters in Confined Spaces Using Multi-Modal Perception and LLM-Driven High Semantic Reasoning

Shoaib Ahmmad, Zubayer Ahmed Aditto, Md Mehrab Hossain, Noushin Yeasmin, Shorower Hossain

机构 * Department of Mechanical Engineering(机械工程系) Rajshahi University of Engineering and Technology(拉贾沙希工程与技术大学) Department of Industrial and Production Engineering(工业与生产工程系) Shahjalal University of Science and Technology(沙赫jalal科学与技术大学) Bangladesh University of Engineering and Technology(孟加拉工程与技术大学) Department of Urban and Regional Planning(城市与区域规划系) Department of Computer Science Engineering(计算机科学与工程系) United International University(联合国际大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09969 2025-08-12 cs.HC cs.AI 57%

Steering AI-Driven Personalization of Scientific Text for General Audiences

Taewook Kim, Dhruv Agarwal, Jordan Ackerman, Manaswi Saha

机构 * Northwestern University(西北大学) Cornell University(康奈尔大学) Center for Advanced AI, Accenture(Accenture高级人工智能中心) Accenture Labs(Accenture实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 28 pages, 7 figures, 1 table. Accepted to PACM HCI (CSCW 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06765 2025-08-12 cs.LG 57%

Fed MobiLLM: Efficient Federated LLM Fine-Tuning over Heterogeneous Mobile Devices via Server Assisted Side-Tuning

Xingke Yang, Liang Li, Sicong Li, Liwei Guan, Hao Wang, Xiaoqi Qi, Jiang Liu, Xin Fu, Miao Pan

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06103 2025-08-11 cs.CL cs.IR 57%

Few-Shot Prompting for Extractive Quranic QA with Instruction-Tuned LLMs

Mohamed Basem, Islam Oshallah, Ali Hamdi, Ammar Mohammed

机构 * Faculty of Computer Science MSA University Giza, Egypt(计算机科学学院 MSA大学 埃及吉扎) Faculty of Computer Science MSA University Egypt(计算机科学学院 MSA大学 埃及)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments 6 pages , 2 figures , Accepted in IMSA 2025,Egypt , https://imsa.msa.edu.eg/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23673 2025-08-11 cs.AI 57%

HASD: Hierarchical Adaption for pathology Slide-level Domain-shift

Jingsong Liu, Han Li, Chen Yang, Michael Deutges, Ario Sadafi, Xin You, Katharina Breininger, Nassir Navab, Peter J. Schüffler

机构 * Institute of Pathology, Technical University of Munich, TUM School of Medicine and Health(病理研究所,慕尼黑技术大学,TUM医学院和健康学院) Computer Aided Medical Procedures (CAMP), TU Munich(计算机辅助医疗程序(CAMP),慕尼黑技术大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML)) Institute of AI for Health, Computational Health Center, Helmholtz Munich(健康人工智能研究所,计算健康中心,海德堡慕尼黑) Center for AI and Data Science (CAIDAS), Julius-Maximilians-Universität Würzburg(人工智能与数据科学中心(CAIDAS),维尔茨堡约纳斯-马克斯-迈克尔-大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Accepted by MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08501 2025-08-11 cs.LG 57%

Learning to Match Unpaired Data with Minimum Entropy Coupling

Mustapha Bounoua, Giulio Franzese, Pietro Michiardi

机构 * Department of Data Science, EURECOM, France(数据科学系,EURECOM,法国)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01979 2025-08-11 cs.CL 57%

Gradient-Regularized Latent Space Modulation in Large Language Models for Structured Contextual Synthesis

Derek Yotheringhay, Beatrix Nightingale, Maximilian Featherstone, Edmund Worthington, Hugo Ashdown

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00301 2025-08-11 cs.CL 57%

Contextual Morphogenesis in Large Language Models: A Novel Approach to Self-Organizing Token Representations

Alistair Dombrowski, Beatrix Engelhardt, Dimitri Fairbrother, Henry Evidail

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18826 2025-08-11 cs.CL 57%

Structural Embedding Projection for Contextual Large Language Model Inference

Vincent Enoasmo, Cedric Featherstonehaugh, Xavier Konstantinopoulos, Zacharias Huntington

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05427 2025-08-08 cs.AI 57%

Large Language Models Transform Organic Synthesis From Reaction Prediction to Automation

Kartar Kumar Lohana Tharwani, Rajesh Kumar, Sumita, Numan Ahmed, Yong Tang

机构 * International Research Center for Complexity Sciences, Hangzhou International Innovation Institute, Beihang University(国际复杂科学研究中心,杭州创新研究院,北京航空航天大学) School of Computer Science and Engineering, University of Electronic Science and Technology of China(计算机科学与工程学院,电子科学与技术大学) Yangtze Delta Region Institute (Huzhou), University of Electronic Science and Technology of China(长江三角洲地区研究所(湖州),电子科学与技术大学) Government Boys Higher Secondary School, Bukera Sharif, Tando Allahyar, Affiliated with BISE Hyderabad, Sindh, Pakistan(政府男生高级中学,布克里·沙里夫,塔ndo阿勒·耶尔,隶属于Hyderabad BISE,Sindh,巴基斯坦) School of Environment and Architecture, University of Shanghai for Science and Technology(环境与建筑学院,上海科学技术大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05388 2025-08-08 cs.AI 57%

An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal

Silvia García-Méndez, Francisco de Arriba-Pérez, Fátima Leal, Bruno Veloso, Benedita Malheiro, Juan Carlos Burguillo-Rial

机构 * Information Technologies Group, atlanTTic, University of Vigo, Spain(信息科技组,atlanTTic,维戈大学,西班牙) REMIT, Universidade Portucalense, Portugal(REMIT,葡萄牙普拉亚恩斯大学) Faculty of Economics, University of Porto, Portugal(经济学院,波尔图大学,葡萄牙) INESC TEC, Porto, Portugal(INESC TEC,波尔图,葡萄牙) ISEP, Porto, Portugal(ISEP,波尔图,葡萄牙)

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05232 2025-08-08 cs.LG 57%

Cross-LoRA: A Data-Free LoRA Transfer Framework across Heterogeneous LLMs

Feifan Xia, Mingyang Liao, Yuyang Fang, Defang Li, Yantong Xie, Weikang Li, Yang Li, Deguo Xia, Jizhou Huang

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05116 2025-08-08 cs.AI cs.MA 57%

Beyond Automation: Socratic AI, Epistemic Agency, and the Implications of the Emergence of Orchestrated Multi-Agent Learning Architectures

Peer-Benedikt Degen, Igor Asanov

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10479 2025-08-08 cs.AI 57%

DeclareAligner: A Leap Towards Efficient Optimal Alignments for Declarative Process Model Conformance Checking

Jacobo Casas-Ramos, Manuel Lama, Manuel Mucientes

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08800 2025-08-08 cs.CL 57%

Data Processing for the OpenGPT-X Model Family

Nicolo' Brandizzi, Hammam Abdelwahab, Anirban Bhowmick, Lennard Helmer, Benny Jörg Stein, Pavel Denisov, Qasid Saleem, Michael Fromm, Mehdi Ali, Richard Rutmann, Farzad Naderi, Mohamad Saif Agy, Alexander Schwirjow, Fabian Küch, Luzian Hahn, Malte Ostendorff, Pedro Ortiz Suarez, Georg Rehm, Dennis Wegener, Nicolas Flores-Herr, Joachim Köhler, Johannes Leveling

机构 * Fraunhofer IAIS(弗劳恩霍夫人工智能研究所) Fraunhofer IIS(弗劳恩霍夫信息处理研究所) DFKI(德意志国防科研机构)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14805 2025-08-07 cs.CL 57%

How Well Do LLMs Represent Values Across Cultures? Empirical Analysis of LLM Responses Based on Hofstede Cultural Dimensions

Julia Kharchenko, Tanya Roosta, Aman Chadha, Chirag Shah

机构 * University of Washington(华盛顿大学) UC Berkeley, Amazon(伯克利大学、亚马逊) Stanford University, Amazon GenAI(斯坦福大学、亚马逊生成人工智能)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments KDD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03651 2025-08-06 cs.HC cs.AI 57%

Probing the Gaps in ChatGPT Live Video Chat for Real-World Assistance for People who are Blind or Visually Impaired

Ruei-Che Chang, Rosiana Natalie, Wenqian Xu, Jovan Zheng Feng Yap, Anhong Guo

机构 * University of Michigan(密歇根大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments ACM ASSETS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03393 2025-08-06 cs.SE cs.AI 57%

Agentic AI in 6G Software Businesses: A Layered Maturity Model

Muhammad Zohaib, Muhammad Azeem Akbar, Sami Hyrynsalmi, Arif Ali Khan

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 6 pages, 3 figures and FIT'25 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03053 2025-08-06 cs.RO cs.AI 57%

SkeNa: Learning to Navigate Unseen Environments Based on Abstract Hand-Drawn Maps

Haojun Xu, Jiaqi Xiang, Wu Wei, Jinyu Chen, Linqing Zhong, Linjiang Huang, Hongyu Yang, Si Liu

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏