arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-09-25 至 2025-09-25 共收录 37 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 15 篇

2509.19789 2025-09-25 cs.LG cs.AI cs.RO 62%

RDAR: Reward-Driven Agent Relevance Estimation for Autonomous Driving

Carlo Bosio, Greg Woelki, Noureldin Hendy, Nicholas Roy, Byungsoo Kim

机构 * UC Berkeley(伯克利大学) Zoox Inc.(Zoox公司)

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19628 2025-09-25 cs.CE cs.CL q-fin.CP 57%

Multimodal Language Models with Modality-Specific Experts for Financial Forecasting from Interleaved Sequences of Text and Time Series

Ross Koval, Nicholas Andrews, Xifeng Yan

机构 * University of California, Santa Barbara(加州大学圣巴巴拉分校) Johns Hopkins University(约翰霍普金斯大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19453 2025-09-25 astro-ph.IM cs.LG 57%

The Platonic Universe: Do Foundation Models See the Same Sky?

UniverseTBD, :, Kshitij Duraphe, Michael J. Smith, Shashwat Sourav, John F. Wu

机构 * Independent Researcher(独立研究者) AstroAI Harvard-Smithsonian CfA(哈佛-史密松天体物理中心) University of Hertfordshire(赫特福德郡大学) Washington University St. Louis(圣路易斯华盛顿大学) Space Telescope Science Institute(空间望远镜科学研究所) Johns Hopkins University(约翰霍普金斯大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments 9 pages, 3 tables, 1 figure. Accepted as a workshop paper to Machine Learning and the Physical Sciences at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20343 2025-09-25 cs.CV 50%

Efficient Encoder-Free Pose Conditioning and Pose Control for Virtual Try-On

Qi Li, Shuwen Qiu, Julien Han, Xingzi Xu, Mehmet Saygin Seyfioglu, Kee Kiat Koo, Karim Bouyarmane

机构 * Amazon(亚马逊) University of California, Los Angeles(加州大学洛杉矶分校) Duke University(杜克大学)

专题命中 其他安全 :alignment(abstract)

Comments Submitted to CVPR 2025 and Published at CVPR 2025 AI for Content Creation workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19790 2025-09-25 cs.AR 50%

Open-source Stand-Alone Versatile Tensor Accelerator

Anthony Faure-Gignoux, Kevin Delmas, Adrien Gauffriau, Claire Pagetti

专题命中 其他安全 :safety(abstract)

Journal ref 44th Digital Avionics Systems Conference (DASC), Sep 2025, Montreal, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12587 2025-09-25 cs.CV 50%

Multimodal Chain of Continuous Thought for Latent-Space Reasoning in Vision-Language Models

Tan-Hanh Pham, Chris Ngo

机构 * Harvard Medical School, Harvard University(哈佛医学院、哈佛大学) Athinoula A. Martinos Center for Biomedical Imaging, Massachusetts General Hospital(阿提尼乌拉A.马丁努斯生物医学成像中心、麻省总医院) Knovel Engineering Lab, Singapore(Knovel工程实验室、新加坡)

专题命中 其他安全 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16803 2025-09-25 cs.RO cs.CV cs.NI eess.IV 50%

RG-Attn: Radian Glue Attention for Multi-modality Multi-agent Cooperative Perception

Lantao Li, Kang Yang, Wenqi Zhang, Xiaoxue Wang, Chen Sun

机构 * Sony (China) Limited(索尼(中国)有限公司) Renmin University of China(中国人民大学)

专题命中 其他安全 :alignment(abstract)

Comments Accepted by ICCV 2025 DriveX workshop (Final Version)

详情

展开后加载摘要…

URL PDF HTML 收藏