arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-21 至 2025-11-21 共收录 10 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 10 篇

2511.16483 2025-11-21 cs.LG cs.AI cs.MA 90%

Large Language Model-Based Reward Design for Deep Reinforcement Learning-Driven Autonomous Cyber Defense

基于大型语言模型的深度强化学习驱动自主网络防御的奖励设计

Sayak Mukherjee, Samrat Chatterjee, Emilie Purvine, Ted Fujimoto, Tegan Emerson

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本文提出利用大型语言模型设计奖励,以提升深度强化学习在自主网络防御中的效果。

Comments Accepted in the AAAI-26 Workshop on Artificial Intelligence for Cyber Security (AICS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01082 2025-11-21 cs.CL 85%

Turning Up the Heat: Min-p Sampling for Creative and Coherent LLM Outputs

提升采样温度:用于生成创意且连贯LLM输出的min-p采样

Minh Nhat Nguyen, Andrew Baker, Clement Neo, Allen Roush, Andreas Kirsch, Ravid Shwartz-Ziv

机构 * Apart Research Independent(独立研究者) New York University(纽约大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 min-p采样通过动态截断方法提升LLM生成文本的质量和多样性,尤其在高温度下表现优异,已获多个开源框架采用。

Comments Oral presentation at ICLR 2025. Camera-ready version available at https://iclr.cc/virtual/2025/poster/30358

Journal ref In Proceedings of the 2025 International Conference on Learning Representations (ICLR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15733 2025-11-21 cs.SE cs.AI 77%

Technique to Baseline QE Artefact Generation Aligned to Quality Metrics

一种基于质量指标的QE人工生成技术

Eitan Farchi, Kiran Nayak, Papia Ghosh Majumdar, Saritha Route

机构 * IBM Research(IBM研究院) IBM Consulting(IBM咨询)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种基于质量指标的QE人工生成技术,结合LLM生成、反向生成和评分技术,通过实验验证其在不同输入质量下的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16018 2025-11-21 cs.AI cs.CL 76%

SpellForger: Prompting Custom Spell Properties In-Game using BERT supervised-trained model

SpellForger: 使用BERT监督训练模型在游戏中自定义咒语属性

Emanuel C. Silva, Emily S. M. Salum, Gabriel M. Arantes, Matheus P. Pereira, Vinicius F. Oliveira, Alessandro L. Bicho

专题命中 其他LLM :prompting(title);分类 cs.CL、cs.AI

AI总结 SpellForger利用BERT模型让玩家通过自然语言提示自定义游戏中的咒语属性,通过实时生成咒语验证AI作为游戏机制的可行性。

Comments Published in Anais Estendidos do XXIV Simpósio Brasileiro de Jogos e Entretenimento Digital (SBGames 2025)

Journal ref Anais Estendidos do XXIV Simpósio Brasileiro de Jogos e Entretenimento Digital (SBGames 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09751 2025-11-21 cs.CL cs.AI cs.HC cs.IR cs.LG 75%

OmniThink: Expanding Knowledge Boundaries in Machine Writing through Thinking

OmniThink: 通过思考扩展机器写作的知识边界

Zekun Xi, Wenbiao Yin, Jizhan Fang, Jialong Wu, Runnan Fang, Yong Jiang, Pengjun Xie, Fei Huang, Huajun Chen, Ningyu Zhang

机构 * Zhejiang University(浙江大学) Tongyi Lab, Alibaba Group(阿里云实验室,阿里巴巴集团) Zhejiang Key Laboratory of Big Data Intelligent Computing(浙江大数据智能计算重点实验室)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 OmniThink通过模拟人类思考过程,提升机器写作的知识密度和原创性,解决传统方法在生成长文时的不足。

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16092 2025-11-21 cs.SE 71%

The Future of Development Environments with AI Foundation Models: NII Shonan Meeting 222 Report

人工智能基础模型对未来开发环境的未来影响:NII Shonan会议222报告

Xing Hu, Raula Gaikovina Kula, Christoph Treude

专题命中 其他LLM :foundation model(title)

AI总结 人工智能基础模型对开发环境的影响及未来发展方向,通过专家讨论探讨其在软件工程和人机交互中的挑战与机遇。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17534 2025-11-21 cs.CV cs.CL cs.MM 70%

Co-Reinforcement Learning for Unified Multimodal Understanding and Generation

协同强化学习用于统一多模态理解和生成

Jingjing Jiang, Chongjie Si, Jun Luo, Hanwang Zhang, Chao Ma

机构 * Shanghai Jiao Tong University(上海交通大学) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出CoRL框架,通过协同强化学习提升多模态大语言模型在生成与理解任务上的性能。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10893 2025-11-21 cs.LG 57%

Multi-View Polymer Representations for the Open Polymer Prediction

多视图聚合表示用于开放聚合预测

Wonjin Jung, Yongseok Choi

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 本文提出多视图聚合方法,结合多种表示方式提升聚合物预测性能,在NeurIPS 2025挑战中取得第9名,公共MAE为0.057

Comments The authors have decided to withdraw this manuscript due to internal approval and authorship issues. A revised version may be posted in the future

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16113 2025-11-21 q-bio.QM 50%

ProtT-Affinity: Sequence-Based Protein-Protein Binding Affinity Prediction Using ProtT5 Embeddings

基于序列的蛋白质-蛋白质结合亲和力预测:使用ProtT5嵌入的ProtT-Affinity

Hongfu Lou

专题命中 其他LLM :language model(abstract)

AI总结 ProtT-Affinity通过结合ProtT5嵌入和轻量级Transformer架构,实现了基于序列的蛋白质-蛋白质结合亲和力预测,在缺乏结构信息时提供实用的替代方案。

Comments 9 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13863 2025-11-21 cs.CV cs.SD eess.AS 50%

Segmenting Collision Sound Sources in Egocentric Videos

从第一人称视频中分割碰撞声音源

Kranti Kumar Parida, Omar Emara, Hazel Doughty, Dima Damen

机构 * Samsung R&D Institute India – Bangalore(三星印度研发中心-班加罗尔) University of Bristol(布里斯托大学) Leiden University(莱顿大学)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文提出碰撞声音源分割任务,利用基础模型和第一人称线索,通过音频条件分割在两个基准上取得显著效果。

Comments Webpage: https://krantiparida.github.io/projects/cs3.html

详情

展开后加载摘要…

URL PDF HTML 收藏