arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2026-03-03 至 2026-03-03 共收录 136 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 40 篇

2602.04369 2026-03-03 cs.LG 57%

Multi-scale hypergraph meets LLMs: Aligning large language models for time series analysis

多尺度超图与大语言模型:面向时间序列分析的对齐方法

Zongjiang Shang, Dongliang Cui, Binqing Wu, Ling Chen

机构 * State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学) College of Computer Science and Technology, Zhejiang University(计算机科学与技术学院,浙江大学)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 本文提出MSH-LLM方法,通过多尺度超图机制和跨模态对齐模块,提升大语言模型在时间序列分析中的表现。

Comments Accepted by ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00350 2026-03-03 cs.AI 57%

Monotropic Artificial Intelligence: Toward a Cognitive Taxonomy of Domain-Specialized Language Models

单调人工智能:面向领域专用语言模型的认知分类

Antonio de Sousa Leitão Filho, Allan Kardec Duailibe Barros Filho, Fabrício Saul Lima, Selby Mykael Lima dos Santos, Rejani Bandeira Vieira Sousa

机构 * Aia Context Federal University of Maranhão(马那瓜联邦大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 单调人工智能通过牺牲通用性实现领域内高精度,挑战通用智能主导的AI研究范式,提出专门化与通用系统互补共存的认知生态。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05064 2026-03-03 cs.LG 57%

Boomerang Distillation Enables Zero-Shot Model Size Interpolation

回环蒸馏使零样本模型大小插值

Sara Kangaslahti, Nihal V. Nayak, Jonathan Geuter, Marco Fumero, Francesco Locatello, David Alvarez-Melis

机构 * Harvard University(哈佛大学) Kempner Institute(凯普纳研究所) IST Austria(IST奥地利研究院)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 回环蒸馏通过蒸馏和重构实现零样本模型大小插值,生成细粒度模型家族,降低训练成本并提升适应性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12490 2026-03-03 physics.ao-ph cs.LG 57%

SamudrACE: Fast and Accurate Coupled Climate Modeling with 3D Ocean and Atmosphere Emulators

SamudrACE:基于3D海洋和大气模拟器的快速准确耦合气候建模

James P. C. Duncan, Elynn Wu, Surya Dheeshjith, Adam Subel, Troy Arcomano, Spencer K. Clark, Brian Henn, Anna Kwa, Jeremy McGibbon, W. Andre Perkins, William Gregory, Carlos Fernandez-Granda, Julius Busecke, Oliver Watt-Meyer, William J. Hurlin, Alistair Adcroft, Laure Zanna, Christopher Bretherton

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 SamudrACE通过3D海洋和大气模拟器实现快速准确的耦合气候建模,能够模拟数百年长的高分辨率气候现象。

Comments 29 pages, 26 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18564 2026-03-03 cs.LG cs.CE 57%

Efficient Aircraft Design Optimization Using Multi-Fidelity Models and Multi-fidelity Physics Informed Neural Networks

利用多保真模型和多保真物理指导神经网络实现高效飞机设计优化

Apurba Sarker

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 本研究利用多保真物理指导神经网络和生成对抗网络,实现高效飞机设计优化,提升设计迭代速度与经济性。

Comments 7 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00144 2026-03-03 cs.CV cs.AI 57%

Disentangled Hierarchical VAE for 3D Human-Human Interaction Generation

解耦层次变分自编码器用于3D人-人交互生成

Zichen Geng, Zeeshan Hayder, Bo Miao, Jian Liu, Wei Liu, Ajmal Mian

机构 * Department of CSSE, The University of Western Australia(西澳大学计算机科学与工程系) Data61, CSIRO(澳大利亚联邦科学工业研究组织Data61部门) Australian Institute for Machine Learning, The University of Adelaide(澳大利亚阿德莱德大学人工智能研究所) NERC-RVC, Hunan University(湖南大学NERC-RVC部门)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出DHVAE,通过解耦层次变分自编码器生成结构化且可控的3D人-人交互,提升运动保真度和物理合理性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01552 2026-03-03 cs.CV 50%

Align-cDAE: Alzheimer's Disease Progression Modeling with Attention-Aligned Conditional Diffusion Auto-Encoder

Align-cDAE: 利用注意力对齐的条件扩散自编码器进行阿尔茨海默病进展建模

Ayantika Das, Keerthi Ram, Mohanasankar Sivaprakasam

机构 * Department of Electrical Engineering, Indian Institute of Technology Madras, Chennai, India(电子工程系,印度理工学院马德拉斯,钦奈,印度) Sudha Gopalakrishnan Brain Centre, Indian Institute of Technology Madras, Chennai, India(苏达·戈帕拉克里希南脑中心,印度理工学院马德拉斯,钦奈,印度) Department of Electrical Engineering, Indian Institute of Technology Madras Chennai, India(电子工程系,印度理工学院马德拉斯钦奈,印度)

专题命中 其他安全 :alignment(abstract)

AI总结 Align-cDAE通过引入注意力对齐和结构化潜在空间,提升扩散自编码器在阿尔茨海默病进展建模中的精度和可控性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01417 2026-03-03 cs.IR 50%

ReFeed: Retrieval Feedback-Guided Dataset Construction for Style-Aware Query Rewriting

ReFeed: 基于检索反馈的领域感知查询重写数据集构建

Jiyoon Myung, Jungki Son, Kyungro Lee, Jihyeon Park, Joohyung Han

专题命中 其他安全 :alignment(abstract)

AI总结 ReFeed通过检索反馈驱动数据集构建,提升查询重写模型对文档风格的感知能力,改进领域特定场景下的检索效果。

Comments Accepted at the Workshop on New Frontiers in Information Retrieval (AAAI 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01319 2026-03-03 cs.HC 50%

Caught in a Mafia Romance: How Users Explore Intimate Roleplay and Narrative Exploration with Chatbots

陷入黑手党浪漫:用户如何通过聊天机器人探索亲密角色扮演与叙事探索

Julia Kieserman, Cat Mai, Sara Lignell, Lucy Qin, Athanasios Andreou, Damon McCoy, Rosanna Bellini

专题命中 其他安全 :safety(abstract)

AI总结 研究探讨用户通过聊天机器人进行亲密角色扮演和幻想探索的行为,发现用户偏好特定角色设定并对其内容的性化程度提出安全需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11021 2026-03-03 econ.GN q-fin.EC 50%

AI and Worker Well-Being: Differential Impacts Across Generational Cohorts and Genders

人工智能与工人福祉:不同世代和性别群体的差异影响

Voraprapa Nakavachara

专题命中 其他安全 :safety(abstract)

AI总结 本文研究人工智能对不同世代和性别群体工人福祉的影响,发现AI使用在心理健康、工作满意度和身体健康方面的影响存在显著差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00682 2026-03-03 cs.CV 50%

CoLC: Communication-Efficient Collaborative Perception with LiDAR Completion

CoLC: 通信高效协同感知与激光雷达补全

Yushan Han, Hui Zhang, Qiming Xia, Yi Jin, Yidong Li

机构 * Key Laboratory of Big Data & Artificial Intelligence in Transportation, Ministry of Education(交通运输大数据与人工智能重点实验室,教育部) School of Computer Science and Technology, Beijing Jiaotong University(北京交通大学计算机科学与技术学院) Xiamen University(厦门大学)

专题命中 其他安全 :alignment(abstract)

AI总结 CoLC通过激光雷达补全和早期融合技术,在稀疏传输下实现高效协同感知,提升感知与通信的平衡性能。

Comments Accepted by CVPR'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00482 2026-03-03 cs.CV cs.IT math.IT 50%

TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications

TokenCom: 一种用于多模态和多任务令牌通信的视觉-语言模型

Feibo Jiang, Siwei Tu, Li Dong, Xiaolong Li, Kezhi Wang, Cunhua Pan, Zhu Han, Jiangzhou Wang

机构 * Hunan Provincial Key Laboratory of Intelligent Computing and Language Information Processing, Hunan Normal University(湖南省级智能计算与语言信息处理重点实验室,湖南师范大学) School of Information Science and Engineering, Hunan Normal University(信息科学与工程学院,湖南师范大学) Changsha Social Laboratory of Artificial Intelligence, Hunan University of Technology and Business(长沙人工智能社会实验室,湖南工业大学) School of Computer Science, Hunan University of Technology and Business(计算机科学学院,湖南工业大学) Department of Computer Science, Brunel University London(伦敦布鲁内尔大学计算机科学系) National Mobile Communications Research Laboratory, Southeast University(东南大学国家移动通信研究中心) Department of Electrical and Computer Engineering, University of Houston(电子与计算机工程系,休斯顿大学)

专题命中 其他安全 :alignment(abstract)

AI总结 TokenCom提出了一种新的视觉-语言模型框架TaiChi,通过双视觉分词器和双向注意力网络提升多模态和多任务令牌通信的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00466 2026-03-03 cs.CV 50%

DreamWorld: Unified World Modeling in Video Generation

DreamWorld: 视频生成中的统一世界建模

Boming Tan, Xiangdong Zhang, Ning Liao, Yuqing Zhang, Shaofeng Zhang, Xue Yang, Qi Fan, Yanyong Zhang

机构 * University of Science(科学技术大学) Shanghai Jiao Tong University, Shanghai, China.(上海交通大学) Nanjing University, Suzhou, China.(南京大学)

专题命中 其他安全 :alignment(abstract)

AI总结 DreamWorld通过联合世界建模范式和约束退火技术,提升视频生成的世界一致性,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00409 2026-03-03 cs.CV 50%

SSR: Pushing the Limit of Spatial Intelligence with Structured Scene Reasoning

SSR:通过结构化场景推理推动空间智能的极限

Yi Zhang, Youya Xia, Yong Wang, Meng Song, Xin Wu, Wenjun Wan, Bingbing Liu, AiXue Ye, Hongbo Zhang, Feng Wen

机构 * Foundation Model Department, Huawei(华为基础模型部门)

专题命中 其他安全 :alignment(abstract)

AI总结 SSR通过结构化场景推理框架,在减少对齐成本的同时,实现了高效的空间智能,优于更大规模模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04395 2026-03-03 cs.CV 50%

Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models

傅里叶-注意力表示学习:一种引导少样本泛化的视觉-语言模型框架

Hieu Dinh Trung Pham, Huy Minh Nhat Nguyen, Cuong Tuan Nguyen

机构 * Vietnamese German University(越南德语大学)

专题命中 其他安全 :alignment(abstract)

AI总结 本文提出FARL框架,通过傅里叶分析解缠视觉表示,提升视觉-语言模型在少样本场景下的泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00114 2026-03-03 cs.CV 50%

Automated Quality Check of Sensor Data Annotations

传感器数据标注的自动化质量检查

Niklas Freund, Zekiye Ilknur-Öz, Tobias Klockau, Patrick Naumann, Philipp Neumaier, Martin Köppel

专题命中 其他安全 :safety(abstract)

AI总结 本文提出了一种开源工具,用于自动检测铁路车辆多传感器数据集中的九种常见错误,以提高训练数据质量并加速自动驾驶系统的发展。

Journal ref Proceeding of 4th IEEE International Conference on Consumer Electronics (ICCE), Berlin, Germany, September, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏