Operationalising the Superficial Alignment Hypothesis via Task Complexity
通过任务复杂度操作化浅层对齐假设
Tomás Vergara-Browne, Darshan Patil, Ivan Titov, Siva Reddy, Tiago Pimentel, Marius Mosbach
机构
*
University of Maryland(马里兰大学)
;
University of California, Berkeley(加州大学伯克利分校)
;
University of Washington(华盛顿大学)
;
University of Toronto(多伦多大学)
;
University of Edinburgh(爱丁堡大学)
专题命中
指令微调
:large language model(abstract);language model(abstract);post-training(abstract);分类 cs.LG
Role Steering of Language Models for Social Simulations
用于社会模拟的语言模型角色引导
Isaac Song, Mohammed Rehan Parwani, Glenn Matlin, Emile Anand, Akhil Theerthala, Arjun Chatterjee, Anthony Wen-Ming Zang, Maria Kostylew, Yonadav G. Shavit, Sebastien Krier, Mark Riedl
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
ML Alignment & Theory Scholars (MATS)(ML对齐与理论学者组织(MATS))
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Oxford(牛津大学)
;
Google DeepMind(谷歌DeepMind)
;
OpenAI(开放人工智能公司)
FiLoRA: Focus-and-Ignore LoRA for Controllable Feature Reliance
FiLoRA: 专注与忽略LoRA用于可控的特征依赖
Hyunsuk Chung, Soyeon Caren Han, Seungyeon Ji, Jinwoo Kim, Eun-Jung Holden, Kyungreem Han
机构
*
University of Melbourne, Melbourne, Australia
;
Brain Science Institute, Korea Institute of Science
;
Department of Computer Science
;
Engineering, Korea University, Seoul, Republic of Korea
;
Division of Bio-Medical Science \& Technology, University of Science
;
Technology KIST School, Seoul, Republic of Korea
Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for Secure Network Electronic Health Record (EHR) Interoperability
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
预条件测试时适应用于叙事生成中的分布外去偏
Hanwen Shen, Ting Ying, Jiajie Lu, Shanshan Wang
机构
*
Laboratory for Artificial Intelligence in Mathematics Education, Stevens Institute of Technology(数学教育中的人工智能实验室,史蒂文斯理工学院)
;
Independent Researcher(独立研究者)
;
NLP2CT Lab, Department of Computer and Information Science, University of Macau(NLP2CT实验室,澳门大学计算机与信息科学系)
专题命中
指令微调
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Decoupled Alignment for Robust Plug-and-Play Adaptation
用于鲁棒即插即用适应的解耦对齐
Haozheng Luo, Jiahao Yu, Wenxin Zhang, Jialong Li, Chenghao Qiu, Yimin Wang, Eric Hanchen Jiang, Jerry Yao-Chieh Hu, Yan Chen, Binghui Wang, Xinyu Xing, Han Liu
机构
*
Northwestern University(西北大学)
;
New York University Abu Dhabi(纽约大学阿布扎克分校)
;
Stanford University(斯坦福大学)
;
Texas A&M University(德克萨斯农工大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Illinois Institute of Technology(伊利诺伊理工学院)
专题命中
指令微调
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
CommentsRevised to correct the Acknowledgments section. Previous versions inadvertently included acknowledgments of NSF and NIH awards that did not support this work. Those funding acknowledgments have been removed. The technical content, results, and conclusions are unchanged
IRC-Bench: Recognizing Entities from Contextual Cues in First-Person Reminiscences
IRC-Bench: 从第一人称回忆中的上下文线索识别实体
Yehudit Aperstein, Eden Moran, Alexander Apartsin
机构
*
Intelligent Systems, Afeka Academic College of Engineering(阿法卡学术工程学院智能系统)
;
School of Computer Science, Faculty of Sciences, Holon Institute of Technology(霍隆理工学院计算机科学学院)