arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

NeurIPS

Conference on Neural Information Processing Systems · 会议 · Machine Learning

至 收录 45
2606.24650 2026-06-24 cs.CL cs.LG 新提交

Harmonic: Hierarchical State Space Models for Efficient Long-Context Language Modeling

Harmonic: 用于高效长上下文语言建模的分层状态空间模型

Petr Nyoma

机构 * Independent Researcher(独立研究员)

AI总结 提出分层状态空间模型Harmonic,通过堆叠三个不同时间尺度的循环层,每层接收下层预测误差而非原始隐藏状态,在长序列上显著优于Transformer和Mamba,并消除了RoPE位置编码限制。

Comments 12 pages, 8 figures. NeurIPS 2024 format

URL PDF HTML 收藏
2606.22858 2026-06-23 cs.LG cs.AI 新提交

The Unseen Hand: Manipulating Model Fairness and SHAP with Targeted Identity Re-Association Attacks

无形之手:通过目标身份重新关联攻击操纵模型公平性和SHAP

Sannaan Khan, Muhammad U. S. Khan

机构 * National University of Sciences and Technology (NUST)(国立科技大学(NUST))

AI总结 提出目标身份重新关联(TIRA)攻击,通过概率性微扰操纵模型输出,在不留痕迹的情况下扭曲公平性指标和SHAP解释。

Comments Accepted at NeurIPS Workshops 2025

URL PDF HTML 收藏
2606.22182 2026-06-23 cs.CV cs.AI 新提交

Dual-Stream EEG Decoding for 3D Visual Perception

用于3D视觉感知的双流脑电解码

Ninon Lizé Masclef, Taisija Demcenko, Antonella Catanzaro, Nataliya Kosmyna

机构 * Massachusetts Institute of Technology(麻省理工学院)

AI总结 提出一种模仿生物视觉腹侧和背侧通路的双流脑电解码模型,通过圆形回归预测角度和EEG条件多视图扩散实现3D重建,揭示了时间动态的通道参与模式。

Comments 17 pages, 4 figures. Accepted at the Symmetry and Geometry in Neural Representations Workshop (NeurReps), NeurIPS 2025. To appear in Proceedings of Machine Learning Research (PMLR)

URL PDF HTML 收藏
2606.19882 2026-06-19 cs.CV cs.LG 新提交

Multimodal Concept Bottleneck Models

多模态概念瓶颈模型

Tongqing Shi, Ge Yan, Tuomas Oikarinen, Tsui-Wei Weng

机构 * UC San Diego(加州大学圣地亚哥分校)

AI总结 提出多模态概念瓶颈模型(MM-CBM),利用双概念瓶颈层对齐图像和文本嵌入,实现可解释的零样本分类和图像检索,在四个基准上平均准确率提升高达51.26%。

Comments Present at NeurIPS 2025 Mechanistic Interpretability Workshop

URL PDF HTML 收藏
2606.18338 2026-06-18 cs.LG astro-ph.EP astro-ph.IM 新提交

ThousandWorlds: A benchmark for climate emulation of potentially habitable exoplanets

ThousandWorlds: 一个用于潜在宜居系外行星气候模拟的基准数据集

Edward T. Stevenson, Mei Ting Mak, Eric Wolf, Denis E. Sergeev, Tobi Hammond, N. J. Mayne, Miles Cranmer

机构 * University of Cambridge(剑桥大学) University of Oxford(牛津大学) University of Colorado Boulder(科罗拉多大学博尔德分校) University of Bristol(布里斯托大学) Purdue University(普渡大学) University of Exeter(埃克塞特大学)

AI总结 为加速系外行星气候模拟,提出ThousandWorlds基准数据集,包含五个全球气候模型的约1800次模拟,用于评估机器学习模拟器在低数据、多模拟器参数到场回归任务中的性能。

Comments 10 pages main text, 26 pages references/appendix, plus NeurIPS checklist. Data at https://doi.org/10.57967/hf/8695. Code at https://github.com/edstevenson/ThousandWorlds

URL PDF HTML 收藏
2606.17639 2026-06-18 cs.RO cs.CV 新提交

ERQA-Plus: A Diagnostic Benchmark for Reasoning in Embodied AI

ERQA-Plus:具身AI推理的诊断基准

Hong Yang, Basura Fernando

机构 * Centre for Frontier AI Research, Agency for Science, Technology and Research(新加坡科技研究局前沿人工智能研究中心) College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)

AI总结 提出ERQA-Plus基准,包含1766个基于机器人中心图像的问答实例,覆盖感知、动作、社交、导航和常识推理,用于诊断具身AI的推理能力。

Comments under review at NeurIPS

URL PDF HTML 收藏
2606.17526 2026-06-17 cs.LG 新提交

MGUP: A Momentum-Gradient Alignment Update Policy for Stochastic Optimization

MGUP:一种用于随机优化的动量-梯度对齐更新策略

Da Chang, Ganzhao Yuan

机构 * Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院) Shenzhen University of Advanced Technology(深圳理工大学) Pengcheng Laboratory(鹏城实验室) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 提出MGUP机制,通过按固定比例选择参数施加大步长、其余参数用小步长,增强动量优化器,理论保证收敛,实验表明提升训练效率与稳定性。

Comments Published in NeurIPS 2025

URL PDF HTML 收藏
2606.15899 2026-06-16 cs.CR cs.AI cs.HC cs.LG cs.MA 新提交

SkillVetBench: LLM-as-Judge for Multi-Dimensional Security Risk Evaluation in Open-Source LLM Agent Skills

SkillVetBench: 基于LLM评判的多维安全风险评估开源LLM智能体技能

Ismail Hossain, Sai Puppala, Md Jahangir Alam, Tanzim Ahad, Sajedul Talukder

机构 * SUPREME Lab, University of Texas at El Paso, Texas, USA(SUPREME实验室,德克萨斯理工大学埃尔帕索分校,德克萨斯州,美国)

AI总结 提出SkillVetBench,利用LLM作为评判器对开源LLM智能体技能进行多维安全风险评估,引入五维技能智能体风险评分(SARS)和CVSS v4.0向量分解,在78个恶意技能上实现零假阴性,22个良性技能上零假阳性。

Comments The main research paper is submitted to NeurIPS 2027, it is in under review

URL PDF HTML 收藏
2606.14673 2026-06-15 cs.LG 新提交

Compressed Computation is (probably) not Computation in Superposition

压缩计算(可能)不是叠加计算

Jai Bhagat, Sara Molas-Medina, Giorgi Giglemiani, Stefan Heimersheim

机构 * Metamorphic Independent(独立研究者) UK AI Security Institute(英国人工智能安全研究所) Apollo Research

AI总结 通过分析压缩计算(CC)模型,发现其性能提升源于标签中的混合矩阵,而非真正的叠加计算,SNMF基线可复现其损失特征。

Comments Presented at the Mechanistic Interpretability Workshop at NeurIPS 2025

URL PDF HTML 收藏
2606.14108 2026-06-15 cs.LG cs.AI 新提交

Numbers Already Carry Their Own Embeddings

数字本身已携带其嵌入

Suhyun Bae, Donghun Lee

机构 * Department of Mathematics, Korea University(高丽大学数学系)

AI总结 提出无训练嵌入方法AOE,同时保留数字的实数值与p-adic模签名,实现即插即用并在代数组合基准上首次达到完美精度。

Comments Presented at the MATH-AI Workshop at NeurIPS 2025

URL PDF HTML 收藏
2606.12747 2026-06-12 cs.AI 新提交

Prefill Awareness in Large Language Models

大型语言模型中的预填充感知

Andy Wang, Parv Mahajan, David Demitri Africa, Alexandra Souly, Jordan Taylor, Robert Kirk

机构 * Constellation University of Wisconsin-Madison(威斯康星大学麦迪逊分校星座研究所) Constellation Georgia Institute of Technology(佐治亚理工学院星座研究所) UK AI Security Institute(英国人工智能安全研究所)

AI总结 研究大型语言模型能否识别并响应其助手消息被预填充或篡改,发现前沿模型具有显著预填充感知能力,可能影响安全评估方法。

Comments Submitted to NeurIPS 2026

URL PDF HTML 收藏
2606.11199 2026-06-11 cs.CL cs.AI cs.IR cs.LG 新提交

NightFeats @ MMU-RAGent NeurIPS 2025: A Context-Optimized Multi-Agent RAG System for the Text-to-Text Track

NightFeats @ MMU-RAGent NeurIPS 2025: 面向文本到文本轨道的上下文优化多智能体RAG系统

Quentin Fever, Naziha Aslam

机构 * NightFeats

AI总结 提出一种结构化多智能体RAG系统NightFeats,通过检索、策展和组合三阶段分解知识合成,引入时序语义重排序、矛盾协调和引用保留架构,在MMU-RAGent竞赛中超越商业基线。

Comments 5 pages, 1 figure, 1 table. NeurIPS 2025 Competition Track (MMU-RAGent). System developed October 2025

URL PDF HTML 收藏
2606.11130 2026-06-10 cs.LG 新提交

Robust Regression of General ReLUs with Queries

一般ReLU的鲁棒回归与查询

Ilias Diakonikolas, Daniel M. Kane, Mingchen Ma

机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校) University of California, San Diego(加利福尼亚大学圣迭戈分校)

AI总结 针对高斯分布下一般ReLU的平方损失鲁棒回归,提出首个高效查询算法,使用d polylog(1/ε)+Õ(min{1/p,1/ε})个标签查询达到O(opt)+ε误差,并证明查询复杂度近最优。

Comments Appeared at NeurIPS 2025

URL PDF HTML 收藏
2606.09940 2026-06-10 cs.LG cs.AI 新提交

Interactions Between Crosscoder Features: A Compact Proofs Perspective

交叉编码器特征间的交互:一个紧凑证明的视角

Dmitry Manning-Coe, Thomas Read, Anna Soligo, Oliver Clive-Griffin, Chun-Hei Yip, Rajashree Agrawal, Jason Gross

机构 * Anthony J. Leggett Institute for Condensed Matter Theory(安东尼·J·莱格特凝聚态理论研究所) MATS UK AI Security Institute (AISI)(英国人工智能安全研究所) Imperial College London(帝国理工学院伦敦分校) Goodfire University of Cambridge(剑桥大学) Theorem Labs(定理实验室)

AI总结 本文从紧凑证明角度形式化交叉编码器特征交互,提出交互度量并应用于计算稀疏性、语义聚类和检测休眠代理。

Comments Accepted at the NeurIPS 2025 Workshop on Mechanistic Interpretability

URL PDF HTML 收藏
2606.09156 2026-06-09 cs.CV 新提交

OmniGen-AR: AutoRegressive Any-to-Image Generation

OmniGen-AR: 自回归任意到图像生成

Junke Wang, Xun Wang, Qiushan Guo, Peize Sun, Weilin Huang, Zuxuan Wu, Yu-Gang Jiang

机构 * Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究所) Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心) Bytedance Seed(字节跳动Seed) The University of Hong Kong(香港大学)

AI总结 提出统一自回归框架OmniGen-AR,通过共享视觉分词器和解耦因果注意力,支持文本、空间信号和视觉上下文等多种条件输入,在多项基准上达到最优或竞争性能。

Comments Accepted by NeurIPS

URL PDF HTML 收藏