Foundations of Top-$k$ Decoding For Language Models
语言模型中Top-k解码的基础理论
专题命中 其他LLM :language model(title);LLM(abstract);分类 cs.AI、cs.LG
AI总结 本研究提出了一种理论框架,解释并推广了Top-k解码,展示了其在稀疏分布恢复中的有效性,并提出了新的解码策略。
AI 大模型
大语言模型、预训练、指令微调、后训练和语言模型应用。
语言模型中Top-k解码的基础理论
专题命中 其他LLM :language model(title);LLM(abstract);分类 cs.AI、cs.LG
AI总结 本研究提出了一种理论框架,解释并推广了Top-k解码,展示了其在稀疏分布恢复中的有效性,并提出了新的解码策略。
汽车系统工程中可信生成式人工智能的工作流级设计原则
机构 * Carl von Ossietzky University of Oldenburg(奥尔登堡卡尔·冯·奥西特齐克大学) ; DENSO AUTOMOTIVE
专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.LG
AI总结 本文提出可信生成式人工智能在汽车系统工程中工作流级设计原则,通过需求增量识别、SysML架构更新及可追溯测试保障安全关键系统工程的可信度。
Botson:一种易于获取且低成本的社会机器人研究平台
机构 * University of Michigan-Dearborn(密歇根大学迪尔伯恩分校)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 Botson是一种基于大型语言模型的人形社会机器人,旨在为社会机器人研究提供低成本且易于获取的平台。
Comments 5 pages, 7 figures
promptolution: 一种统一的、模块化的提示优化框架
机构 * ELLIS Institute(ELLIS研究所) ; University of Freiburg(弗赖堡大学) ; LMU Munich(慕尼黑大学) ; Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) ; Technical University of Munich(慕尼黑技术大学) ; TU Dortmund University(多特蒙德技术大学) ; Lamarr Institute for Machine Learning and Artificial Intelligence(Lamarr机器学习与人工智能研究所)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
AI总结 promptolution 提供了一种统一的模块化框架,整合多种提示优化器,支持系统化的基准测试,并返回与框架无关的提示字符串,以提升大型语言模型在各种任务中的性能。
GLaDiGAtor: 基于语言模型的多关系图学习用于预测疾病-基因关联
机构 * Biological Data Science Lab, Dept. of Computer Engineering, Hacettepe University(生物数据科学实验室,计算机工程系,哈切泰佩大学) ; Dept. of Bioinformatics, Graduate School of Health Sciences, Hacettepe University(生物信息学系,健康科学研究生院,哈切泰佩大学) ; Dept. of Health Informatics, Institute of Informatics, Hacettepe University(健康信息学系,信息学院,哈切泰佩大学)
专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.AI、cs.LG
AI总结 GLaDiGAtor通过整合语言模型特征的异构图学习方法,提升了疾病-基因关联预测的准确性和泛化能力。
通过基于LLM的查询规范化实现OLAP的语义缓存
专题命中 其他LLM :LLM(title)
AI总结 本文提出基于LLM的查询规范化方法,通过统一的OLAP意图签名提升OLAP缓存命中率,实现82%的高命中率,显著优于传统方法。
Comments 12 pages, 2 figures, 5 tables. Extended version of the short paper published at DOLAP 2026 (co-located with EDBT/ICDT 2026)
SAMAS:一种基于频谱的多智能体系统,用于实现文学翻译中的风格保真
机构 * Beijing Normal University(北京师范大学) ; University of Science and Technology of China(中国科学技术大学) ; University of Coimbra(科英布拉大学) ; Beijing University of Posts and Telecommunications(北京邮电大学) ; Northwestern Polytechnical University(西北工业大学)
专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL
AI总结 SAMAS通过将风格保真视为信号处理任务,利用小波包变换生成风格特征频谱,动态组装翻译智能体工作流程,从而提升文学翻译中的风格保真度。
通过用户-助手模拟开发前瞻性与个性化的人工智能助手
机构 * KAIST(韩国科学技术院)
专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL
AI总结 ProPerSim通过用户-助手模拟框架开发了能够主动和个性化推荐的AI助手,实验显示其在多样化的用户场景中有效提升了用户满意度。
Comments Accepted at ICLR 2026
好奇心胜过喧嚣:通过建模动机语言来理解选择性量子轨迹中的早期成果
专题命中 其他LLM :language model(abstract);small language model(abstract)
AI总结 研究通过分析申请人的动机语言,探讨其在早期量子计算课程中的表现预测,发现好奇心相关主题与学业成绩相关,但推断测试效果有限,需进一步研究。
Comments Published in the Proceedings of IEEE ICALTER 2025. 5 pages, 7 figures
Journal ref Proceedings of the IEEE International Conference on Advanced Learning Technologies (ICALTER), 2025
Tele-Omni: 一种用于视频生成与编辑的统一多模态框架
机构 * TeleAI
专题命中 其他LLM :large language model(abstract);language model(abstract)
AI总结 Tele-Omni是一种统一多模态框架,通过解析文本、图像和参考视频指令,实现视频生成与编辑的灵活控制,提升时间一致性和视觉一致性。
文化沟通中的巴别塔:#Give Me a Chinese Name# 在“TikTok难民”事件中的案例研究
专题命中 其他LLM :large language model(abstract);language model(abstract)
AI总结 研究通过分析TikTok难民请求中文名字的跨文化沟通事件,揭示了跨语言文化动态中的编码解码机制及影响参与度的策略。
Comments 21 pages, 6 figures, 6 tables
通过模拟语义翻译能否帮助LLMs进行代码翻译?基于伪代码的研究
专题命中 其他LLM :large language model(abstract);language model(abstract)
AI总结 本研究通过基于伪代码的翻译方法提升LLMs在代码翻译中的表现,发现其在复杂程序处理中具有优势,但受限于伪代码的准确性。
Comments Accepted by ACM Transactions on Software Engineering and Methodology (TOSEM)
为优化低技能用户策略的个性化帮助
专题命中 其他LLM :language agent(abstract);分类 cs.CL
AI总结 本文提出通过CICERO生成个性化建议,帮助低技能玩家在Diplomacy游戏中提升策略表现,即使玩家不遵循建议,其存在也具有优势。
Comments 9 pages, 3 figures
RDFC-GAN:基于RGB-深度融合的循环GAN用于室内深度补全
机构 * State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications, China(网络与交换技术国家重点实验室,北京邮电大学,中国) ; Midea Group, China(美的集团,中国) ; School of Computer Science, Beijing University of Posts and Telecommunications, China(计算机科学学院,北京邮电大学,中国)
专题命中 其他LLM :prompting(abstract);分类 cs.AI
AI总结 RDFC-GAN通过融合RGB和深度图像,利用循环GAN和自适应融合模块提升室内深度补全效果。
Comments Haowen Wang and Zhengping Che are with equal contributions. Paper accepted by IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI). An earlier version has been accepted by CVPR 2022 (arXiv:2203.10856). arXiv admin note: text overlap with arXiv:2203.10856
Journal ref IEEE Transactions on Pattern Analysis and Machine Intelligence (Volume: 46, Issue: 11, November 2024)
开放权重安全性的收敛速率控制极限
机构 * Dalhousie University(达尔豪斯大学) ; Vector Institute(向量研究所) ; Indian Institute of Management Bangalore(班加罗尔印度管理学院)
专题命中 其他LLM :foundation model(abstract);分类 cs.LG
AI总结 本文提出SpecDef算法,通过谱重参数化在非对抗性设置中减缓优化收敛速度,并揭示了对抗性环境下收敛速率控制方法的理论极限。
Comments Submitted to ICML 2026. 13 figures, 30 tables
人机交互:从视频演示中学习机器人模仿
专题命中 其他LLM :language model(abstract)
AI总结 本研究提出了一种基于视频演示的机器人模仿学习方法,通过模块化框架结合时间位移模块和深度强化学习,实现机器人从无结构视频中学习基本操作技能。
通过拼接实现自动、表达性强且可扩展的模糊测试
专题命中 其他LLM :LLM(abstract)
AI总结 STITCH通过拼接技术实现自动、表达性强且可扩展的模糊测试,发现更多真实bug并提高精度。