Comments2 pages, industry paper, to appear in proceedings of the 35th ACM International Conference on Information and Knowledge Management (CIKM '26), 2026
机构
*
School of Electrical and Computer Engineering, College of Engineering, University of Tehran(德黑兰大学工程学院电气与计算机工程学院)
;
Sharif University of Technology(谢里夫理工大学)
;
Missouri University of Science and Technology(密苏里科技大学)
;
Tehran Institute for Advanced Studies, Khatam University(哈塔姆大学德黑兰高等研究院)
Comments22 pages total (12-page main paper + 10-page supplementary material). Revised after peer review. Accepted for oral presentation and publication in the BICA 2026 proceedings, Springer Lecture Notes in Electrical Engineering (LNEE)
M2K: Making the Model-Kernel Interface Explicit for Reliable CUDA Kernel Verification
Model2Kernel: 用于安全CUDA内核的模型感知符号执行
Mengting He, Shihao Xia, Haomin Jia, Wenfei Wu, Linhai Song
机构
*
The Pennsylvania State University(宾夕法尼亚州立大学)
;
State Key Lab of Processors, Institute of Computing Technology, CAS(中国科学院计算技术研究所处理器实验室)
;
Peking University(北京大学)
专题命中
长上下文与记忆
:LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.AI
ShardMemo: Scope-Before-Routing for Agentic Memory Retrieval
ShardMemo: 用于分片代理LLM内存的遮蔽MoE路由
Yang Zhao, Chengxiao Dai, Mengying Kou, Yue Xiu, Dusit Niyato
机构
*
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
LongGuard: Mechanistic Analysis and Training-Free Mitigation of Long-Context Failure in Safety Guardrails
LongGuard:针对安全护栏长上下文失效的机制分析与无训练缓解方案
Ziyang Chen, Xing Wu, Songlin Hu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)
专题命中
长上下文与记忆
:LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI
GRKV: Global Regression for Training-Free KV Cache Compression in Long-Context LLMs
GRKV: 长上下文LLM中免训练的KV缓存压缩的全局回归
Junjie Peng, You Wu, Haoyi Wu, Jialong Han, Xiaohua Xie, Kewei Tu, Jianhuang Lai
机构
*
Sun Yat-sen University(中山大学)
;
ShanghaiTech University(上海科技大学)
;
Guangdong Province Key Laboratory of Information Security Technology(广东省信息安全技术重点实验室)
专题命中
长上下文与记忆
:LLM(title_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
Can Retrieval Heads See Images? Multimodal Retrieval Heads in Long-Context Vision-Language Models
检索头能看见图像吗?长上下文视觉语言模型中的多模态检索头
Aaron Branson Cigres Li, Zhaowei Wang, Yu Zhao, Yiming Du, Haobo Li, Xiyu Ren, Ginny Wong, Simon See, Lishu Luo, Haodong Duan, Pasquale Minervini, Yangqiu Song
机构
*
HKUST(香港科技大学)
;
University of Edinburgh(爱丁堡大学)
;
CUHK(香港中文大学)
;
NVAITC, NVIDIA, Santa Clara, USA(NVIDIA Santa Clara 分公司)
;
Tsinghua University(清华大学)
专题命中
长上下文与记忆
:language model(title,abstract);large language model(abstract)
机构
*
Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
专题命中
长上下文与记忆
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI
Hindsight Memory-PRM: Supervising Memory Management with Auditable Hindsight Credit
事后记忆-PRM:通过可审计的事后信用监督记忆管理
Haoxuan Jia, Yang Liu, Yingguang Yang, Yancheng Chen, Chongyang Zhang, Hao Zheng, Qian Li, Yulin Huang, Jianshen Zhang, Yongzhi Qi, Shang Luo, Kefu Xu, Hao Peng, Junyu Lu, Du Cheng, Philip S. Yu, Bin Chong
机构
*
Peking University(北京大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Fullive-AI
;
Nanyang Technological University(南洋理工大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Beihang University(北京航空航天大学)
;
Beijing Institute of Technology, Zhuhai(北京理工大学珠海学院)
;
Northeastern University(东北大学)
;
University of Illinois Chicago(芝加哥大学伊利诺伊分校)
DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory Architectures
DCC: 面向处理-内存架构的机器学习内核数据驱动编译
Peiming Yang, Sankeerth Durvasula, Ivan Fernandez, Mohammad Sadrosadati, Onur Mutlu, Gennady Pekhimenko, Christina Giannoula
机构
*
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
;
Barcelona Supercomputing Center(巴塞罗那超级计算中心)
;
ETH Zürich(苏黎世联邦理工学院)
;
Nvidia(英伟达)
;
Max Planck Institute for Software Systems(马克斯·普朗克软件系统研究所)
专题命中
长上下文与记忆
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
LLMs Can't Play Hangman: On the Necessity of a Private Working Memory for Language Agents
LLMs无法玩井字游戏:关于语言代理所需私人工作记忆的必要性
Davide Baldelli, Ali Parviz, Amal Zouaq, Sarath Chandar
机构
*
Mila – Quebec AI Institute(魁北克人工智能研究所)
;
Polytechnique Montréal(蒙特利尔大学)
;
University of California, San Diego(加州大学圣地亚哥分校)
;
LAMA-WeST Lab(LAMA-WeST实验室)
;
Chandar Research Lab(Chandar研究实验室)