REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing
REDDIT:基于回放的分布编辑在无遗忘条件下校正ASR的模型生成时间戳漂移
Cheng-Kang Chou, Ming-To Chuang, Ke-Han Lu, Chan-Jan Hsu, Hung-yi Lee
机构
*
National Taiwan University(台湾大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
NTU Artificial Intelligence Center of Research Excellence (NTU AI-CoRE)(国立清华大学人工智能研究中心(NTU AI-CoRE))
机构
*
City University of Hong Kong(香港城市大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
The Chinese University of Hong Kong(香港中文大学)
PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails
PolicyShiftGuard:基准测试与改进策略自适应图像护栏
Mingyang Song, Luxin Xu, Haoyu Sun, Minzhou Pan, Yu Cheng, Bo Li
机构
*
Fudan University(复旦大学)
;
Tongji University(同济大学)
;
Virtue AI(美德人工智能公司)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Chicago(芝加哥大学)
;
University of Illinois, Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
Priority-Aware Learning-Unlearning Correction for Dynamic Decentralized LoRA Fine-Tuning
面向动态去中心化LoRA微调的优先级感知学习-遗忘校正
Nuocheng Yang, Yechen He, Sihua Wang, Zihan Chen, Tony Q. S. Quek, Changchuan Yin
机构
*
Beijing Key Laboratory of Network System Architecture and Convergence, Beijing University of Posts and Telecommunications(网络系统架构与融合北京市重点实验室,北京邮电大学)
;
Information Systems Technology and Design Pillar, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计系)
专题命中
指令微调
:large language model(abstract);language model(abstract);post-training(abstract);分类 cs.AI、cs.LG
Sparsity Curse: Understanding RLVR Model Parameter Space from Model Merging
稀疏性诅咒:从模型合并理解RLVR模型参数空间
Chenrui Wu, Zexi Li, Jiajun Bu, Jiangchuan Liu, Haishuai Wang
机构
*
Zhejiang University(浙江大学)
;
Simon Fraser University(西蒙菲莎大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Zhejiang Key Lab of Accessible Perception and Intelligent Systems(浙江省可感知智能系统重点实验室)
机构
*
MoE Key Lab of BIPC, University of Science and Technology of China(中科院大学科学技术大学MoE关键实验室)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai AI Laboratory(上海人工智能实验室)
Comments9 pages, 5 figures. Empirical study of in-context learning and LoRA fine-tuning for synthetic tabular data generation, introducing the phenomenon of categorical prior lock-in. Under review
Cognitive Fatigue in Autoregressive Transformers: Formalization and Measurement
自回归Transformer中的认知疲劳:形式化与测量
Riju Marwah, Ritvik Garimella, Vishal Pallagani, Atishay Jain, Michael Stewart, Amit Sheth
机构
*
Guru Gobind Singh Indraprastha University, India(古鲁·戈宾德·辛格·印度普拉斯塔大学)
;
Artificial Intelligence Institute, University of South Carolina, USA(人工智能研究所,南卡罗来纳大学)
;
Indian Institute of Technology, Kanpur, India(印度理工学院,坎浦尔)
;
Indian AI Research Organization, India(印度人工智能研究组织)
机构
*
Department of Computing, The Hong Kong Polytechnic University, Hong Kong, China(香港理工大学计算机系)
;
Department of Control Science and Engineering, Zhejiang University(浙江大学控制科学与工程学院)
专题命中
指令微调
:LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations
CroCo: 基于自生成结果的跨语言对比偏好调优
Mike Zhang, Ali Basirat, Desmond Elliott
机构
*
Department of Computer Science (DIKU), University of Copenhagen(哥本哈根大学计算机科学系(DIKU))
;
Centre for Language Technology (CST), University of Copenhagen(哥本哈根大学语言技术中心)
;
Pioneer Centre for Artificial Intelligence(先锋人工智能中心)
专题命中
指令微调
:large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.CL、cs.AI
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
Inner Mongolia University of Technology(内蒙古科技大学)
;
Tsinghua University(清华大学)
;
Shandong University(山东大学)
RotMoLE: Enhancing Mixture of Low-Rank Experts through Rotational Gating Mechanism
RotMoLE:通过旋转门控机制增强混合低秩专家
Mengyang Sun, Maochuan Dou, Tao Feng, Dan Zhang, Yihao Wang, Junpeng Liu, Yifan Zhu, Jie Tang
机构
*
Tsinghua University(清华大学)
;
Beijing Information Science and Technology University(北京信息科技大学)
;
National University of Singapore(新加坡国立大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Beijing University of Posts and Telecommunications(北京邮电大学)
专题命中
指令微调
:LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG