arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2509.05358 2025-09-09 cs.CY 57%

Mobile Phone Sensor-based Nigerian Driving Dataset to Detect Alcohol-influenced Behaviours

Iniakpokeikiye Peter Thompson, Yi Dewei, Reiter Ehud

专题命中 其他安全 :safety(abstract);分类 cs.CY

Comments The article has 4 tables and 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14096 2025-09-09 cs.CV cs.AI 57%

Image Segmentation with Large Language Models: A Survey with Perspectives for Intelligent Transportation Systems

Sanjeda Akter, Ibne Farabi Shihab, Anuj Sharma

机构 * Department of Computer Science, Iowa State University, Ames, IA, USA(Iowa State University) Department of Civil, Construction and Environmental Engineering, Iowa State University, Ames, IA, USA(Iowa State University)

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05051 2025-09-08 quant-ph cs.LG 57%

QCA-MolGAN: Quantum Circuit Associative Molecular GAN with Multi-Agent Reinforcement Learning

Aaron Mark Thomas, Yu-Cheng Chen, Hubert Okadome Valencia, Sharu Theresa Jose, Ronin Wu

机构 * Department of Computer Science(计算机科学系) University of Birmingham(伯明翰大学) QunaSys Europe(QunaSys欧洲分公司) Hon Hai Research Institute(鸿海研究机构) Departement of Computer Science(计算机科学系)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Accepted to the proceedings of IEEE Quantum Artificial Intelligence, 6 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04942 2025-09-08 cs.LG 57%

Ontology-Aligned Embeddings for Data-Driven Labour Market Analytics

Heinke Hihn, Dennis A. V. Dittrich, Carl Jeske, Cayo Costa Sobral, Helio Pais, Timm Lochmann

机构 * IU International University of Applied Sciences, Department of Computer Science and Engineering(国际应用科学大学计算机科学与工程系) The Stepstone Group, AI Labs(Stepstone集团人工智能实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Workshop SIG Knowledge Management (FG WM) at KI2025, Potsdam, Germany

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04751 2025-09-08 cs.IR cs.LG 57%

Multimodal Foundation Model-Driven User Interest Modeling and Behavior Analysis on Short Video Platforms

Yushang Zhao, Yike Peng, Li Zhang, Qianyi Sun, Zhihui Zhang, Yingying Zhuang

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07173 2025-09-08 cs.AI 57%

Translating Federated Learning Algorithms in Python into CSP Processes Using ChatGPT

Miroslav Popovic, Marko Popovic, Miodrag Djukic, Ilija Basicevic

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments 6 pages, 4 tables; Published by IEEE Xplore

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03680 2025-09-05 cs.GR cs.AI cs.CV 57%

LuxDiT: Lighting Estimation with Video Diffusion Transformer

Ruofan Liang, Kai He, Zan Gojcic, Igor Gilitschenski, Sanja Fidler, Nandita Vijaykumar, Zian Wang

机构 * NVIDIA University of Toronto(多伦多大学) Vector Institute(向量研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Project page: https://research.nvidia.com/labs/toronto-ai/LuxDiT/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08537 2025-09-05 cs.LG math.CT 57%

Recursive Reward Aggregation

Yuting Tang, Yivan Zhang, Johannes Ackermann, Yu-Jie Zhang, Soichiro Nishimori, Masashi Sugiyama

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Reinforcement Learning Conference 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10118 2025-09-05 cs.CV cs.AI 57%

Image Embedding Sampling Method for Diverse Captioning

Sania Waheed, Na Min An

机构 * University of Southhampton(南安普顿大学) KAIST(韩国科学技术院)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 17 pages, 5 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03409 2025-09-04 cs.SD cs.AI cs.MM 57%

Multi-level SSL Feature Gating for Audio Deepfake Detection

Hoan My Tran, Damien Lolive, Aghilas Sini, Arnaud Delhay, Pierre-François Marteau, David Guennec

机构 * Univ Rennes, IRISA, CNRS(里昂大学、IRISA、CNRS) Univ Bretagne Sud, IRISA, CNRS(布列塔尼南大学、IRISA、CNRS)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments This paper has been accepted by ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02918 2025-09-04 cs.CV cs.AI 57%

Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach

Midhat Urooj, Ayan Banerjee, Farhat Shaikh, Kuntal Thakur, Sandeep Gupta

机构 * Impact Lab, Arizona State University(Impact实验室,亚利桑那州立大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Accepted in ANSyA 2025: 1st International Workshop on Advanced Neuro-Symbolic Applications

Journal ref ANSyA 2025: 1st International Workshop on Advanced Neuro-Symbolic Applications

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02610 2025-09-04 q-bio.QM cs.AI 57%

Resilient Biosecurity in the Era of AI-Enabled Bioweapons

Jonathan Feldman, Tal Feldman

机构 * Georgia Institute of Technology(佐治亚理工学院) Yale Law School(耶鲁法学院)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02017 2025-09-03 cs.IR cs.AI 57%

Empowering Large Language Model for Sequential Recommendation via Multimodal Embeddings and Semantic IDs

Yuhao Wang, Junwei Pan, Xinhang Li, Maolin Wang, Yuan Wang, Yue Liu, Dapeng Liu, Jie Jiang, Xiangyu Zhao

机构 * City University of Hong Kong(香港城市大学) Tencent Inc.(腾讯公司) Tsinghua University(清华大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments CIKM 2025 Full Research Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01360 2025-09-03 cs.CV cs.LG 57%

M3Ret: Unleashing Zero-shot Multimodal Medical Image Retrieval via Self-Supervision

Che Liu, Zheng Jiang, Chengyu Fang, Heng Guo, Yan-Jie Zhou, Jiaqi Qu, Le Lu, Minfeng Xu

机构 * DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) Imperial College London(帝国理工学院) Tsinghua University(清华大学) Hupan Lab(华潘实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01147 2025-09-03 cs.CL 57%

Zero-shot Cross-lingual NER via Mitigating Language Difference: An Entity-aligned Translation Perspective

Zhihao Zhang, Sophia Yat Mei Lee, Dong Zhang, Shoushan Li, Guodong Zhou

机构 * School of Computer Science & Technology, NLP Lab, Soochow University, China(计算机科学与技术学院、自然语言处理实验室、苏州大学) Department of Chinese and Bilingual Studies, The Hong Kong Polytechnic University(中文与双语研究系、香港理工大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05025 2025-09-03 cs.LG cs.HC 57%

Will You Be Aware? Eye Tracking-Based Modeling of Situational Awareness in Augmented Reality

Zhehan Qu, Tianyi Hu, Christian Fronk, Maria Gorlatova

机构 * Department of Computer Science Duke University(计算机科学系 哥伦比亚大学) Department of Electrical and Computer Engineering Duke University(电气与计算机工程系 哥伦比亚大学)

专题命中 其他安全 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00802 2025-09-03 cs.LG 57%

XAI-Driven Machine Learning System for Driving Style Recognition and Personalized Recommendations

Feriel Amel Sellal, Ahmed Ayoub Bellachia, Meryem Malak Dif, Enguerrand De Rautlin De La Roy, Mouhamed Amine Bouchiha, Yacine Ghamri-Doudane

机构 * L3i - La Rochelle University(拉罗什-拉科什大学L3i)

专题命中 其他安全 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00664 2025-09-03 cs.CV cs.AI 57%

Fusion to Enhance: Fusion Visual Encoder to Enhance Multimodal Language Model

Yifei She, Huangxuan Wu

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00325 2025-09-03 cs.CL cs.IR 57%

GIER: Gap-Driven Self-Refinement for Large Language Models

Rinku Dewri

机构 * University of Denver(丹佛大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00284 2025-09-03 cs.CV cs.AI 57%

Generative AI for Industrial Contour Detection: A Language-Guided Vision System

Liang Gong, Tommy, Wang, Sara Chaker, Yanchen Dong, Fouad Bousetouane, Brenden Morton, Mark Mendez

机构 * The University of Chicago(芝加哥大学) FabTrack

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 20 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20778 2025-09-03 cs.IR cs.LG 57%

SEAL: Structure and Element Aware Learning to Improve Long Structured Document Retrieval

Xinhao Huang, Zhibo Ren, Yipeng Yu, Ying Zhou, Zulong Chen, Zeyi Wen

机构 * HKUST (Guangzhou)(香港科技大学(广州)) HKUST(香港科技大学) Alibaba Group(阿里巴巴集团) Zhejiang Lab(浙江实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Accepted at EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21291 2025-09-01 cs.AI 57%

Complex System Diagnostics Using a Knowledge Graph-Informed and Large Language Model-Enhanced Framework

Saman Marandi, Yu-Shu Hu, Mohammad Modarres

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments 22 Pages, 11 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20722 2025-08-29 cs.CL 57%

rStar2-Agent: Agentic Reasoning Technical Report

Ning Shang, Yifei Liu, Yi Zhu, Li Lyna Zhang, Weijiang Xu, Xinyu Guan, Buze Zhang, Bingcheng Dong, Xudong Zhou, Bowen Zhang, Ying Xin, Ziming Miao, Scarlett Li, Fan Yang, Mao Yang

机构 * Microsoft Research(微软研究院)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20374 2025-08-29 cs.AI 57%

TCIA: A Task-Centric Instruction Augmentation Method for Instruction Finetuning

Simin Ma, Shujian Liu, Jun Tan, Yebowen Hu, Song Wang, Sathish Reddy Indurthi, Sanqiang Zhao, Liwei Wu, Jianbing Han, Kaiqiang Song

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15576 2025-08-29 cs.CV cs.LG 57%

Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models

Xin Huang, Ruibin Li, Tong Jia, Wei Zheng, Ya Wang

机构 * School of Artificial Intelligence and Software Engineering, Nanyang Normal University, Henan, China(人工智能与软件工程学院,南阳师范学院,河南) Institute for Artificial Intelligence, Peking University, Beijing, China(人工智能研究院,北京大学,北京) Collaborative Innovation Center of Intelligent Explosion-proof Equipment, Henan, China(智能防爆设备协同创新中心,河南)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Accepted at the International Joint Conference on Artificial Intelligence (IJCAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.00258 2025-08-29 cs.AI q-bio.NC 57%

Possible Principles for Aligned Structure Learning Agents

Lancelot Da Costa, Tomáš Gavenčiak, David Hyland, Mandana Samiei, Cristian Dragos-Manta, Candice Pattisapu, Adeel Razi, Karl Friston

机构 * VERSES AI Research Lab(VERSES AI研究实验室) Charles University(查尔斯大学) University of Oxford(牛津大学) Mila, Quebec AI Institute(魁北克人工智能研究院) McGill University(麦吉尔大学) University of Montreal(蒙特利尔大学) University College London(伦敦大学学院) Monash University(莫纳什大学) CIFAR Azrieli Global Scholars Program(CIFAR阿兹里埃利全球学者计划)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 24 pages of content, 33 with references; accepted version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19896 2025-08-28 cs.LG cs.CV 57%

NM-Hebb: Coupling Local Hebbian Plasticity with Metric Learning for More Accurate and Interpretable CNNs

Davorin Miličević, Ratko Grbić

机构 * Faculty of Electrical Engineering, Computer Science and Information Technology Osijek(电子工程、计算机科学与信息技术学院奥斯杰克)

专题命中 其他安全 :safety(abstract);分类 cs.LG

Comments 13 pages, 4 figures. Submitted to Elsevier Neurocomputing, under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19638 2025-08-28 cs.CV cs.AI 57%

Beyond BEV: Optimizing Point-Level Tokens for Collaborative Perception

Yang Li, Quan Yuan, Guiyang Luo, Xiaoyuan Fu, Rui Pan, Yujia Yang, Congzhang Shao, Yuewen Liu, Jinglin Li

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19517 2025-08-28 cs.HC cs.AI 57%

Orchid: Orchestrating Context Across Creative Workflows with Generative AI

Srishti Palani, Gonzalo Ramos

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19289 2025-08-28 cs.CV cs.AI 57%

Seeing Like a Designer Without One: A Study on Unsupervised Slide Quality Assessment via Designer Cue Augmentation

Tai Inui, Steven Oh, Magdeline Kuan

机构 * Waseda University(早稻田大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏