arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-24 至 2026-02-24 共收录 347 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 24 篇

2505.19371 2026-02-24 cs.AI cs.LG math.ST stat.TH 81%

Foundations of Top-$k$ Decoding For Language Models

语言模型中Top-k解码的基础理论

Georgy Noarov, Soham Mallick, Tao Wang, Sunay Joshi, Yan Sun, Yangxinyu Xie, Mengxin Yu, Edgar Dobriban

专题命中 其他LLM :language model(title);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出了一种理论框架,解释并推广了Top-k解码,展示了其在稀疏分布恢复中的有效性,并提出了新的解码策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19614 2026-02-24 cs.SE cs.LG 77%

Workflow-Level Design Principles for Trustworthy GenAI in Automotive System Engineering

汽车系统工程中可信生成式人工智能的工作流级设计原则

Chih-Hong Cheng, Brian Hsuan-Cheng Liao, Adam Molin, Hasan Esen

机构 * Carl von Ossietzky University of Oldenburg(奥尔登堡卡尔·冯·奥西特齐克大学) DENSO AUTOMOTIVE

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.LG

AI总结 本文提出可信生成式人工智能在汽车系统工程中工作流级设计原则,通过需求增量识别、SysML架构更新及可追溯测试保障安全关键系统工程的可信度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19491 2026-02-24 cs.RO cs.AI cs.HC 77%

Botson: An Accessible and Low-Cost Platform for Social Robotics Research

Botson:一种易于获取且低成本的社会机器人研究平台

Samuel Bellaire, Abdalmalek Abu-raddaha, Natalie Kim, Nathan Morhan, William Elliott, Samir Rawashdeh

机构 * University of Michigan-Dearborn(密歇根大学迪尔伯恩分校)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Botson是一种基于大型语言模型的人形社会机器人,旨在为社会机器人研究提供低成本且易于获取的平台。

Comments 5 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02840 2026-02-24 cs.CL 77%

promptolution: A Unified, Modular Framework for Prompt Optimization

promptolution: 一种统一的、模块化的提示优化框架

Tom Zehle, Timo Heiß, Moritz Schlager, Matthias Aßenmacher, Matthias Feurer

机构 * ELLIS Institute(ELLIS研究所) University of Freiburg(弗赖堡大学) LMU Munich(慕尼黑大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) Technical University of Munich(慕尼黑技术大学) TU Dortmund University(多特蒙德技术大学) Lamarr Institute for Machine Learning and Artificial Intelligence(Lamarr机器学习与人工智能研究所)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 promptolution 提供了一种统一的模块化框架,整合多种提示优化器,支持系统化的基准测试,并返回与框架无关的提示字符串,以提升大型语言模型在各种任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18769 2026-02-24 cs.LG cs.AI 73%

GLaDiGAtor: Language-Model-Augmented Multi-Relation Graph Learning for Predicting Disease-Gene Associations

GLaDiGAtor: 基于语言模型的多关系图学习用于预测疾病-基因关联

Osman Onur Kuzucu, Tunca Doğan

机构 * Biological Data Science Lab, Dept. of Computer Engineering, Hacettepe University(生物数据科学实验室,计算机工程系,哈切泰佩大学) Dept. of Bioinformatics, Graduate School of Health Sciences, Hacettepe University(生物信息学系,健康科学研究生院,哈切泰佩大学) Dept. of Health Informatics, Institute of Informatics, Hacettepe University(健康信息学系,信息学院,哈切泰佩大学)

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.AI、cs.LG

AI总结 GLaDiGAtor通过整合语言模型特征的异构图学习方法,提升了疾病-基因关联预测的准确性和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19811 2026-02-24 cs.DB 71%

Semantic Caching for OLAP via LLM-Based Query Canonicalization (Extended Version)

通过基于LLM的查询规范化实现OLAP的语义缓存

Laurent Bindschaedler

专题命中 其他LLM :LLM(title)

AI总结 本文提出基于LLM的查询规范化方法,通过统一的OLAP意图签名提升OLAP缓存命中率,实现82%的高命中率,显著优于传统方法。

Comments 12 pages, 2 figures, 5 tables. Extended version of the short paper published at DOLAP 2026 (co-located with EDBT/ICDT 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19840 2026-02-24 cs.CL 70%

SAMAS: A Spectrum-Guided Multi-Agent System for Achieving Style Fidelity in Literary Translation

SAMAS:一种基于频谱的多智能体系统,用于实现文学翻译中的风格保真

Jingzhuo Wu, Jiajun Zhang, Keyan Jin, Dehua Ma, Junbo Wang

机构 * Beijing Normal University(北京师范大学) University of Science and Technology of China(中国科学技术大学) University of Coimbra(科英布拉大学) Beijing University of Posts and Telecommunications(北京邮电大学) Northwestern Polytechnical University(西北工业大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SAMAS通过将风格保真视为信号处理任务,利用小波包变换生成风格特征频谱,动态组装翻译智能体工作流程,从而提升文学翻译中的风格保真度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21730 2026-02-24 cs.CL 70%

ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant Simulation

通过用户-助手模拟开发前瞻性与个性化的人工智能助手

Jiho Kim, Junseong Choi, Woosog Chay, Daeun Kyung, Yeonsu Kwon, Yohan Jo, Edward Choi

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 ProPerSim通过用户-助手模拟框架开发了能够主动和个性化推荐的AI助手,实验显示其在多样化的用户场景中有效提升了用户满意度。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19659 2026-02-24 physics.ed-ph quant-ph 67%

Curiosity Over Hype: Modeling Motivation Language to Understand Early Outcomes in a Selective Quantum Track

好奇心胜过喧嚣:通过建模动机语言来理解选择性量子轨迹中的早期成果

Daniella Alexandra Crysti Vargas Saldana, Freddy Herrera Cueva

专题命中 其他LLM :language model(abstract);small language model(abstract)

AI总结 研究通过分析申请人的动机语言,探讨其在早期量子计算课程中的表现预测,发现好奇心相关主题与学业成绩相关,但推断测试效果有限,需进一步研究。

Comments Published in the Proceedings of IEEE ICALTER 2025. 5 pages, 7 figures

Journal ref Proceedings of the IEEE International Conference on Advanced Learning Technologies (ICALTER), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09609 2026-02-24 cs.CV 67%

Tele-Omni: a Unified Multimodal Framework for Video Generation and Editing

Tele-Omni: 一种用于视频生成与编辑的统一多模态框架

Jialun Liu, Tian Li, Xiao Cao, Yukuo Ma, Gonghu Shang, Haibin Huang, Chi Zhang, Xiangzhen Chang, Zhiyong Huang, Jiakui Hu, Zuoxin Li, Yuanzhi Liang, Cong Liu, Junqi Liu, Robby T. Tan, Haitong Tang, Qizhen Weng, Yifan Xu, Liying Yang, Xiaoyan Yang, Peng Yu, Shiwen Zhang, Xuelong Li

机构 * TeleAI

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Tele-Omni是一种统一多模态框架,通过解析文本、图像和参考视频指令,实现视频生成与编辑的灵活控制,提升时间一致性和视觉一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18549 2026-02-24 cs.HC 67%

Tower of Babel in Cross-Cultural Communication: A Case Study of #Give Me a Chinese Name# Dialogues During the "TikTok Refugees'' Event

文化沟通中的巴别塔:#Give Me a Chinese Name# 在“TikTok难民”事件中的案例研究

Jielin Feng, Zhibo Yang, Jingyi Zhao, Yujia Li, Xinwu Ye, Xingyu Lan, Siming Chen

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究通过分析TikTok难民请求中文名字的跨文化沟通事件,揭示了跨语言文化动态中的编码解码机制及影响参与度的策略。

Comments 21 pages, 6 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00920 2026-02-24 cs.SE 67%

Can Emulating Semantic Translation Help LLMs with Code Translation? A Study Based on Pseudocode

通过模拟语义翻译能否帮助LLMs进行代码翻译?基于伪代码的研究

Songqiang Chen, Congying Xu, Jingyi Chen, Jialun Cao, Jiarong Wu, Shing-Chi Cheung

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过基于伪代码的翻译方法提升LLMs在代码翻译中的表现,发现其在复杂程序处理中具有优势,但受限于伪代码的准确性。

Comments Accepted by ACM Transactions on Software Engineering and Methodology (TOSEM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09109 2026-02-24 cs.CL 57%

Personalized Help for Optimizing Low-Skilled Users' Strategy

为优化低技能用户策略的个性化帮助

Feng Gu, Wichayaporn Wongkamjan, Jonathan K. Kummerfeld, Denis Peskoff, Jonathan May, Jordan Boyd-Graber

专题命中 其他LLM :language agent(abstract);分类 cs.CL

AI总结 本文提出通过CICERO生成个性化建议,帮助低技能玩家在Diplomacy游戏中提升策略表现,即使玩家不遵循建议,其存在也具有优势。

Comments 9 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.03584 2026-02-24 cs.CV cs.AI 57%

RDFC-GAN: RGB-Depth Fusion CycleGAN for Indoor Depth Completion

RDFC-GAN:基于RGB-深度融合的循环GAN用于室内深度补全

Haowen Wang, Zhengping Che, Yufan Yang, Mingyuan Wang, Zhiyuan Xu, Xiuquan Qiao, Mengshi Qi, Feifei Feng, Jian Tang

机构 * State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications, China(网络与交换技术国家重点实验室,北京邮电大学,中国) Midea Group, China(美的集团,中国) School of Computer Science, Beijing University of Posts and Telecommunications, China(计算机科学学院,北京邮电大学,中国)

专题命中 其他LLM :prompting(abstract);分类 cs.AI

AI总结 RDFC-GAN通过融合RGB和深度图像,利用循环GAN和自适应融合模块提升室内深度补全效果。

Comments Haowen Wang and Zhengping Che are with equal contributions. Paper accepted by IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI). An earlier version has been accepted by CVPR 2022 (arXiv:2203.10856). arXiv admin note: text overlap with arXiv:2203.10856

Journal ref IEEE Transactions on Pattern Analysis and Machine Intelligence (Volume: 46, Issue: 11, November 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18868 2026-02-24 math.OC cs.LG 57%

Limits of Convergence-Rate Control for Open-Weight Safety

开放权重安全性的收敛速率控制极限

Domenic Rosati, Xijie Zeng, Hong Huang, Sebastian Dionicio, Subhabrata Majumdar, Frank Rudzicz, Hassan Sajjad

机构 * Dalhousie University(达尔豪斯大学) Vector Institute(向量研究所) Indian Institute of Management Bangalore(班加罗尔印度管理学院)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 本文提出SpecDef算法,通过谱重参数化在非对抗性设置中减缓优化收敛速度,并揭示了对抗性环境下收敛速率控制方法的理论极限。

Comments Submitted to ICML 2026. 13 figures, 30 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19184 2026-02-24 cs.RO 50%

Human-to-Robot Interaction: Learning from Video Demonstration for Robot Imitation

人机交互:从视频演示中学习机器人模仿

Thanh Nguyen Canh, Thanh-Tuan Tran, Haolan Zhang, Ziyan Gao, Nak Young Chong, Xiem HoangVan

专题命中 其他LLM :language model(abstract)

AI总结 本研究提出了一种基于视频演示的机器人模仿学习方法,通过模块化框架结合时间位移模块和深度强化学习,实现机器人从无结构视频中学习基本操作技能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18689 2026-02-24 cs.SE cs.CR 50%

Automatic, Expressive, and Scalable Fuzzing with Stitching

通过拼接实现自动、表达性强且可扩展的模糊测试

Harrison Green, Fraser Brown, Claire Le Goues

专题命中 其他LLM :LLM(abstract)

AI总结 STITCH通过拼接技术实现自动、表达性强且可扩展的模糊测试,发现更多真实bug并提高精度。

详情

展开后加载摘要…

URL PDF HTML 收藏