arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-07-24 至 2025-07-24 共收录 10 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 10 篇

2507.17118 2025-07-24 cs.AI 83%

HySafe-AI: Hybrid Safety Architectural Analysis Framework for AI Systems: A Case Study

Mandar Pitale, Jelena Frtunikj, Abhinaw Priyadershi, Vasu Singh, Maria Spence

机构 * Nvidia Corporation(英伟达公司) Nvidia GmbH(英伟达德国公司)

专题命中 其他安全 :safety(title,abstract);AI safety(abstract);分类 cs.AI

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16450 2025-07-24 cs.LG eess.SP 74%

RIS-aided Latent Space Alignment for Semantic Channel Equalization

Tomás Hüttebräucker, Mario Edoardo Pandolfo, Simone Fiorellino, Emilio Calvanese Strinati, Paolo Di Lorenzo

机构 * CEA Leti, University Grenoble Alpes(CEA Leti,格勒诺布尔大学) DIAG Department, Sapienza University of Rome(罗马萨皮恩扎大学DIAG系) Consorzio Nazionale Interuniversitario per le Telecomunicazioni (CNIT)(全国大学电信联合体) DIET Department, Sapienza University of Rome(罗马萨皮恩扎大学DIET系)

专题命中 其他安全 :alignment(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17080 2025-07-24 cs.IR cs.AI cs.CV 57%

VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings

Ramin Giahi, Kehui Yao, Sriram Kollipara, Kai Zhao, Vahid Mirjalili, Jianpeng Xu, Topojoy Biswas, Evren Korpeoglu, Kannan Achan

机构 * Walmart Global Tech(沃尔玛全球技术)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Accepted at RecSys 2025; DOI:https://doi.org/10.1145/3705328.3748064

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17055 2025-07-24 cs.RO cs.LG 57%

Shared Control of Holonomic Wheelchairs through Reinforcement Learning

Jannis Bähler, Diego Paez-Granados, Jorge Peña-Queralta

机构 * Swiss Paraplegic Research, SPF(瑞士瘫痪研究机构) SCAI Lab, D-HEST, Swiss Federal School of Technology in Zurich - ETH Zurich(SCAI实验室,瑞士联邦理工学院-苏黎世-ETH Zurich) Centre for Artificial Ingelligece, Zurich University of Applied Sciences - ZHAW. Switzerland(人工智能中心,瑞士应用科学大学-ZHAW)

专题命中 其他安全 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16947 2025-07-24 cs.CL 57%

AI-based Clinical Decision Support for Primary Care: A Real-World Study

Robert Korom, Sarah Kiptinness, Najib Adan, Kassim Said, Catherine Ithuli, Oliver Rotich, Boniface Kimani, Irene King'ori, Stellah Kamau, Elizabeth Atemba, Muna Aden, Preston Bowman, Michael Sharman, Rebecca Soskin Hicks, Rebecca Distler, Johannes Heidecke, Rahul K. Arora, Karan Singhal

机构 * Penda Health(Penda健康) Nairobi County(内罗毕县) OpenAI

专题命中 其他安全 :safety(abstract);分类 cs.CL

Comments Blog: https://openai.com/index/ai-clinical-copilot-penda-health/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12122 2025-07-24 cs.RO cs.AI 57%

ICCO: Learning an Instruction-conditioned Coordinator for Language-guided Task-aligned Multi-robot Control

Yoshiki Yano, Kazuki Shibata, Maarten Kokshoorn, Takamitsu Matsubara

机构 * Division of Information Science, Graduate School of Science and Technology, Nara Institute of Science and Technology (NAIST)(信息科学系,科学技术研究生学校,科学与技术国立研究所) Department of Cognitive Robotics, Faculty of Mechanical Engineering, Delft University of Technology(认知机器人系,机械工程学院,代尔夫特理工大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 8 pages, 9 figures, to be published in the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17456 2025-07-24 cs.CV 50%

Dynamic Scoring with Enhanced Semantics for Training-Free Human-Object Interaction Detection

Francesco Tonini, Lorenzo Vaquero, Alessandro Conti, Cigdem Beyan, Elisa Ricci

机构 * University of Trento(特伦托大学) University of Verona(威尼斯大学) Department of Computer Science(计算机科学系)

专题命中 其他安全 :alignment(abstract)

Comments Accepted to ACM Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10054 2025-07-24 cs.SE 50%

Explicit Vulnerability Generation with LLMs: An Investigation Beyond Adversarial Attacks

Emir Bosnak, Sahand Moslemi, Mayasah Lami, Anil Koyuncu

专题命中 其他安全 :safety(abstract)

Comments Accepted to ICSME 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17165 2025-07-24 cs.SE 50%

Can LLMs Write CI? A Study on Automatic Generation of GitHub Actions Configurations

Taher A. Ghaleb, Dulina Rathnayake

专题命中 其他安全 :alignment(abstract)

Comments Accepted at the 41st IEEE International Conference on Software Maintenance and Evolution 2025 (ICSME'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17050 2025-07-24 cs.CV 50%

Toward Scalable Video Narration: A Training-free Approach Using Multimodal Large Language Models

Tz-Ying Wu, Tahani Trigui, Sharath Nittur Sridhar, Anand Bodas, Subarna Tripathi

机构 * Intel(英特尔)

专题命中 其他安全 :alignment(abstract)

Comments Accepted to CVAM Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏