Shared Parameter Subspaces and Cross-Task Linearity in Emergently Misaligned Behavior
Daniel Aarao Reis Arturi, Eric Zhang, Andrew Ansah, Kevin Zhu, Ashwinee Panda, Aishwarya Balwani
机构
*
McGill University(麦吉尔大学)
;
McMaster University(麦马斯特大学)
;
University of Alberta(阿尔伯塔大学)
;
Algoverse AI Research(Algoverse AI研究)
;
St. Jude Children’s Research Hospital(圣犹大儿童研究医院)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Edit Less, Achieve More: Dynamic Sparse Neuron Masking for Lifelong Knowledge Editing in LLMs
Jinzhe Liu, Junshu Sun, Shufan Shen, Chenxue Yang, Shuhui Wang
机构
*
Key Lab of Intell. Info. Process., Inst. of Comput. Tech., CAS(智能信息处理重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Agriculture Information Institute, CAAS(农业信息研究所,中国农业科学院)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.LG
Comments19 pages, 11 figures, Accepted by NeurIPS 2025
Methodological Insights into Structural Causal Modelling and Uncertainty-Aware Forecasting for Economic Indicators
Federico Cerutti
机构
*
University of Brescia, Italy(意大利布雷西亚大学)
;
Imperial College London, UK(伦敦帝国理工学院)
;
Cardiff University, UK(卡迪夫大学)
;
University of Southampton, UK(南安普顿大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
CommentsAccepted at the 2nd edition of the Workshop in AI and Finance at ECAI-2025
机构
*
Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
;
Peng Cheng Laboratory(鹏城实验室)
;
Huazhong University of Science and Technology(华中科技大学)
;
Xiamen University(厦门大学)
;
The Hong Kong University of Science and Technology, Guangzhou(香港科技大学(广州))
;
School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Concept-Guided Interpretability via Neural Chunking
Shuchen Wu, Stephan Alaniz, Shyamgopal Karthik, Peter Dayan, Eric Schulz, Zeynep Akata
机构
*
Allen Institute(阿伦研究所)
;
University of Washington(华盛顿大学)
;
Télécom Paris, Institut Polytechnique de Paris(巴黎高等电信学院)
;
Institute of Explainable Machine Learning, Helmholtz Munich(可解释机器学习研究所,海德堡大学)
;
Department of Computational Neuroscience, Max Planck Institute for Biological Cybernetics(生物信息学研究所,马克斯·普朗克研究院)
;
Institute for Human-Centered AI, Helmholtz Munich(以人为中心的人工智能研究所,海德堡大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
机构
*
Department of Computer Science, The University of Texas at Dallas(德克萨斯大学达拉斯分校计算机科学系)
;
Department of Computer Science, University of California, Santa Barbara(加州大学圣巴巴拉分校计算机科学系)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
Key Laboratory of Aerospace Information Security and Trusted Computing, Ministry of Education, School of Cyber Science and Engineering, Wuhan University(航空信息安全与可信计算重点实验室,教育部,网络安全与工程学院,武汉大学)
;
Zhejiang University(浙江大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI