arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

NeurIPS

Conference on Neural Information Processing Systems · 会议 · Machine Learning

2026-01-13 至 2026-01-13 共收录 21
2601.07748 2026-01-13 cs.LG cs.AI

Improving Domain Generalization in Contrastive Learning using Adaptive Temperature Control

通过自适应温度控制提升对比学习中的领域泛化能力

Robert Lewis, Katie Matton, Rosalind W. Picard, John Guttag

机构 * Massachusetts Institute of Technology(麻省理工学院)

AI总结 本文提出通过自适应温度控制提升对比学习中的领域泛化能力,通过引入领域标签增强表示的领域不变性,从而提升模型在分布外场景下的性能。

Comments NeurIPS SSL Workshop 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07440 2026-01-13 cs.LG

Variational Autoencoder with Normalizing flow for X-ray spectral fitting

带有归一化流的变分自编码器用于X射线光谱拟合

Fiona Redmen, Ethan Tregidga, James F. Steiner, Cecilia Garraffo

机构 * Department of Physics & Astronomy University of Southampton(物理与天文学系 英国南安普顿大学) Laboratoire d’astrophysique EPFL(天体物理学实验室 EPFL) Harvard-Smithsonian Center for Astrophysics(哈佛-史密松天体物理中心)

AI总结 本文提出一种基于变分自编码器和归一化流的模型,用于高效X射线光谱拟合,实现快速且精确的光谱重建。

Comments 7 pages, 1 table, 3 figures. Accepted as a workshop paper to Machine Learning and the Physical Sciences at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06913 2026-01-13 cs.LG stat.ML

Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities

可计算的多名义逻辑 logit 上下文老虎机与非线性效用

Taehyun Hwang, Dahngoon Kim, Min-hwan Oh

机构 * Seoul National University(首尔国立大学)

AI总结 本文提出了一种基于上置信界原则的算法,用于解决具有非线性效用函数的多名义逻辑上下文老虎机问题,实现了 O(√T) 的遗憾界。

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20928 2026-01-13 cs.CV

Smooth regularization for efficient video recognition

用于高效视频识别的平滑正则化

Gil Goldman, Raja Giryes, Mahadev Satyanarayanan

机构 * Computer Science Department(计算机科学系) Carnegie Mellon University(卡内基梅隆大学) School of Electrical and Computer Engineering(电气与计算机工程学院) Tel-Aviv University(特拉维夫大学)

AI总结 本文提出了一种平滑正则化方法,通过建模连续帧的变化为高斯随机游走,提升轻量级视频识别模型的准确率。

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18559 2026-01-13 cs.CV

C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction

C3Po:通过点图预测实现跨视角跨模态对应

Kuan Wei Huang, Brandon Li, Bharath Hariharan, Noah Snavely

AI总结 C3Po通过点图预测实现跨视角和跨模态的对应关系,提升了几何推理的性能和准确性。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02109 2026-01-13 cs.AI cs.CL cs.CY

Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences

深度价值基准:衡量模型是泛化深度价值还是浅层偏好

Joshua Ashkinaze, Hua Shen, Saipranav Avula, Eric Gilbert, Ceren Budak

机构 * University of Michigan Ann Arbor(密歇根大学安娜堡分校) New York University Shanghai(纽约大学上海)

AI总结 深度价值基准通过测试模型泛化深度价值还是浅层偏好,评估AI对齐的核心能力。

Comments NeurIPS 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25263 2026-01-13 cs.CV

LangHOPS: Language Grounded Hierarchical Open-Vocabulary Part Segmentation

LangHOPS: 语言引导的层次开放词汇部件分割

Yang Miao, Jan-Nico Zaech, Xi Wang, Fabien Despinoy, Danda Pani Paudel, Luc Van Gool

机构 * INSAIT Sofia University "St. Kliment Ohridski"(索菲亚大学"圣克莱门特·欧里迪斯基") ETH Zurich(苏黎世联邦理工学院) TU Munich(慕尼黑技术大学) Toyota Motor Europe(丰田欧洲公司)

AI总结 LangHOPS通过多模态大语言模型实现开放词汇物体-部件实例分割,取得领域内和跨数据集的优异性能。

Comments 10 pages, 5 figures, 14 tables, Neurips 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23013 2026-01-13 cs.LG cs.AI

MoEMeta: Mixture-of-Experts Meta Learning for Few-Shot Relational Learning

MoEMeta:基于少量样本关系学习的专家混合元学习

Han Wu, Jie Yin

机构 * The University of Sydney, Australia(悉尼大学) Peking University, China(北京大学)

AI总结 MoEMeta通过混合专家模型和任务定制适应机制,提升少样本关系学习的泛化与适应能力。

Comments Appear in NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25031 2026-01-13 cs.LG

Bayesian Surrogates for Risk-Aware Pre-Assessment of Aging Bridge Portfolios

贝叶斯代理用于风险意识的桥梁资产老化预评估

Sophia V. Kuhn, Rafael Bischof, Marius Weber, Antoine Binggeli, Michael A. Kraus, Walter Kaufmann, Fernando Pérez-Cruz

机构 * Institute of Structural Engineering, ETH Zurich, Switzerland(瑞士苏黎世联邦理工学院结构工程研究所) Computational Design Lab, ETH Zurich, Switzerland(瑞士苏黎世联邦理工学院计算设计实验室) Design ++, ETH Zurich, Switzerland(瑞士苏黎世联邦理工学院设计++) Institute of Structural Mechanics and Design, TU Darmstadt, Germany(德国达姆施塔特技术大学结构力学与设计研究所) Department of Computer Science, ETH Zurich, Switzerland(瑞士苏黎世联邦理工学院计算机科学系) Bank for International Settlements, Switzerland(瑞士国际清算银行)

AI总结 本文提出贝叶斯神经网络代理用于快速评估桥梁资产老化风险,通过校准不确定性实现高效预评估,减少整体成本和排放。

Comments Accepted at the NeurIPS 2025 Workshop on MLxOR: Mathematical Foundations and Operational Integration of Machine Learning for Uncertainty-Aware Decision-Making

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20211 2026-01-13 cs.LG

Practical do-Shapley Explanations with Estimand-Agnostic Causal Inference

实用的do-Shapley解释与估量无关的因果推理

Álvaro Parafita, Tomas Garriga, Axel Brando, Francisco J. Cazorla

机构 * Barcelona Supercomputing Center(巴塞罗那超级计算中心) Novartis(诺华)

AI总结 本文提出估量量无关方法,使do-Shapley解释在复杂图中可行,并通过加速计算和解释不可接触的数据生成过程,提升解释的可靠性。

Comments Published at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20779 2026-01-13 stat.ML cs.LG

Stable Minima of ReLU Neural Networks Suffer from the Curse of Dimensionality: The Neural Shattering Phenomenon

ReLU神经网络的稳定极小值遭受维度诅咒:神经碎裂现象

Tongtong Liang, Dan Qiao, Yu-Xiang Wang, Rahul Parhi

机构 * UC San Diego(UC圣地亚哥大学)

AI总结 本文研究了ReLU神经网络中稳定极小值的泛化能力问题,揭示了平坦解在高维情况下因维度诅咒而表现不佳的机制。

Comments Camera Ready Version. Accepted by Neurips 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11700 2026-01-13 cs.LG

Geometry-Aware Edge Pooling for Graph Neural Networks

面向几何的边池化用于图神经网络

Katharina Limbeck, Lydia Mezrag, Guy Wolf, Bastian Rieck

机构 * Helmholtz Munich(海德堡-穆恩大学) Technical University of Munich(慕尼黑技术大学) Université de Montréal(蒙特利尔大学) Mila - Quebec AI Institute(魁北克人工智能研究所) Université de Fribourg(弗里堡大学)

AI总结 本文提出面向几何的边池化方法,通过扩散几何和迭代缩减图结构,提升图神经网络在多样任务中的性能和可解释性。

Comments Accepted at the 39th Conference on Neural Information Processing Systems (NeurIPS) 2025. Our code is available at https://github.com/aidos-lab/mag_edge_pool

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10947 2026-01-13 cs.LG cs.RO cs.SY eess.SY math.OC

Certifying Stability of Reinforcement Learning Policies using Generalized Lyapunov Functions

利用广义李雅普诺夫函数验证强化学习策略的稳定性

Kehan Long, Jorge Cortés, Nikolay Atanasov

机构 * Contextual Robotics Institute University of California San Diego(情境机器人研究所 卡罗来纳大学圣地亚哥分校)

AI总结 本文提出利用广义李雅普诺夫函数验证强化学习策略稳定性,通过结合价值函数与神经网络残差项,提升稳定性认证效率。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09572 2026-01-13 cs.LG math.LO math.OC stat.ML

SAD Neural Networks: Divergent Gradient Flows and Asymptotic Optimality via o-minimal Structures

SAD神经网络:通过o-最小结构实现的分歧梯度流和渐近最优性

Julian Kranz, Davide Gallon, Steffen Dereich, Arnulf Jentzen

机构 * Department of Information Systems, University of Münster, Germany(慕尼黑大学信息系统系) Applied Mathematics: Institute for Analysis and Numerics, University of Münster, Germany(慕尼黑大学应用数学系) RiskLab Switzerland, ETH Zürich, Switzerland(苏黎世联邦理工学院瑞士风险实验室) Applied Mathematics: Institute for Mathematical Stochastics, University of Münster, Germany(慕尼黑大学应用数学系) School of Data Science and School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen (CUHK-Shenzhen), China(香港中文大学(深圳)数据科学学院和人工智能学院)

AI总结 SAD神经网络通过o-最小结构的几何特性,证明了梯度流在特定条件下会发散到无穷大,揭示了神经网络损失优化的渐近最优性。

Comments Accepted for NeurIPS 2025, 30 pages, 6 figures. The result about continuous data distributions now has an additional assumption since a gap was identified in a previous version of the proof

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06425 2026-01-13 cs.CL cs.AI cs.LG

Tensor Product Attention Is All You Need

张量积注意力是所有你所需要的

Yifan Zhang, Yifeng Liu, Huizhuo Yuan, Zhen Qin, Yang Yuan, Quanquan Gu, Andrew Chi-Chih Yao

机构 * IIIS, Tsinghua University(清华大学信息科学技术学院) Shanghai Qi Zhi Institute(上海启智研究院) University of California, Los Angeles(加州大学洛杉矶分校)

AI总结 TPA通过张量分解实现高效注意力机制,T6模型在语言建模任务中超越传统基线,提升性能与内存效率。

Comments Published in NeurIPS 2025 (Spotlight); Project Page: https://github.com/tensorgi/TPA

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06109 2026-01-13 cs.AI cs.LG

CBMAS: Cognitive Behavioral Modeling via Activation Steering

通过激活引导的认知行为建模:CBMAS

Ahmed H. Ismail, Anthony Kuang, Ayo Akinkugbe, Kevin Zhu, Sean O'Brien

AI总结 CBMAS通过连续激活引导技术,提升大型语言模型的认知行为可解释性,提供诊断框架和数据集以分析模型行为演变。

Comments Accepted to CogInterp @ NeurIPS 2025. Equal contribution by Ahmed H. Ismail and Anthony Kuang

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06092 2026-01-13 cs.CY cs.AI

Islamic Chatbots in the Age of Large Language Models

大型语言模型时代下的伊斯兰聊天机器人

Muhammad Aurangzeb Ahmad

机构 * Department of Computer Science & Software Engineering University of Washington Bothell(计算机科学与软件工程系华盛顿大学Bothell分校)

AI总结 本文探讨了大型语言模型驱动的伊斯兰聊天机器人对宗教实践的影响,分析了其在知识获取民主化与权威侵蚀之间的矛盾,并提出负责任设计的建议。

Comments Muslim in ML Workshop at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06039 2026-01-13 cs.CL

Operation Veja: Fixing Fundamental Concepts Missing from Modern Roleplaying Training Paradigms

Operation Veja: 修复现代角色扮演训练范式中缺失的根本概念

Yueze Liu, Ajay Nagi Reddy Kumdam, Ronit Kanjilal, Hao Yang, Yichi Zhang

机构 * Divergence 2% LLC Department of Electrical and Computer Engineering University of Illinois Urbana-Champaign(电气与计算机工程系伊利诺伊大学厄巴纳-香槟分校) Department of Computer Science University of Illinois Urbana-Champaign(计算机科学系伊利诺伊大学厄巴纳-香槟分校)

AI总结 Operation Veja提出VEJA框架,通过整合价值观、经历、判断和能力,改进角色扮演模型的数据整理方法,以提升角色真实性和叙述连续性。

Comments Accepted to NeurIPS 2025 PeronaLLM workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06870 2026-01-13 cs.CV

Towards Robust Pseudo-Label Learning in Semantic Segmentation: An Encoding Perspective

面向语义分割的鲁棒伪标签学习:一种编码视角

Wangkai Li, Rui Sun, Zhaoyang Li, Tianzhu Zhang

机构 * University of Science and Technology of China(中国科学技术大学) National Key Laboratory of Deep Space Exploration, Deep Space Exploration Laboratory(国家空间科学探测重点实验室,深空探测实验室)

AI总结 ECOCSeg通过引入基于ECOC的分类器和位级标签去噪机制,提升语义分割中伪标签学习的鲁棒性和稳定性,适用于无监督领域自适应和半监督学习任务。

Comments Accepted by Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08211 2026-01-13 cs.CL cs.AI cs.LG

SAEMark: Steering Personalized Multilingual LLM Watermarks with Sparse Autoencoders

SAEMark: 通过稀疏自编码器引导个性化多语言大语言模型水印

Zhuohao Yu, Xingru Jiang, Weizheng Gu, Yidong Wang, Qingsong Wen, Shikun Zhang, Wei Ye

机构 * Peking University(北京大学)

AI总结 SAEMark通过稀疏自编码器实现多语言大语言模型的高效水印标记,无需修改模型参数,保持文本质量,适用于多种语言和领域。

Comments 24 pages, 12 figures, NeurIPS 2025, code available: https://zhuohaoyu.github.io/SAEMark

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06472 2026-01-13 cs.CL cs.AI cs.CE cs.DL

KARMA: Leveraging Multi-Agent LLMs for Automated Knowledge Graph Enrichment

KARMA:利用多智能体大语言模型实现知识图谱的自动化丰富

Yuxing Lu, Wei Wu, Xukai Zhao, Rui Peng, Jinzhuo Wang

机构 * Department of Big Data and Biomedical AI, Peking University(北京大学大数据与生物医学人工智能系) Wallace H. Coulter Department of Biomedical Engineering, Georgia Institute of Technology(佐治亚理工学院生物医学工程系) School of Architecture, Tsinghua University(清华大学建筑学院)

AI总结 KARMA通过多智能体大语言模型自动丰富知识图谱,有效识别新实体并提高正确性与一致性

Comments 24 pages, 3 figures, 2 tables

Journal ref Spotlight paper of NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏