arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-03 至 2026-02-03 共收录 607 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 46 篇

2507.12094 2026-02-03 cs.LG cs.GT 57%

Is This Predictor More Informative than Another? A Decision-Theoretical Comparison

这个预测器比另一个更有信息量吗?一种决策理论的比较

Yiding Feng, Liuhan Qian, Wei Tang

机构 * Hong Kong University of Science and Technology(香港科学与技术大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 其他LLM :LLM(abstract);分类 cs.LG

AI总结 本文提出信息量差距的概念,用于比较预测器的决策相关性,提供了一种评估预测模型在不同决策任务中表现的理论框架和实验验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09103 2026-02-03 cs.AI 57%

DANIEL: A fast Document Attention Network for Information Extraction and Labelling of handwritten documents

DANIEL:一种快速的文档注意力网络用于手写文档的信息提取与标注

Thomas Constum, Pierrick Tranouez, Thierry Paquet

机构 * LITIS, University of Rouen Normandie(LITIS,鲁昂诺曼大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 DANIEL是一种端到端架构,结合语言模型实现手写文档的信息提取与标注,同时支持多语言和多任务学习,性能优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01930 2026-02-03 cs.RO cs.CV 50%

LIEREx: Language-Image Embeddings for Robotic Exploration

LIEREx: 语言-图像嵌入用于机器人探索

Felix Igelbrink, Lennart Niecksch, Marian Renz, Martin Günther, Martin Atzmueller

专题命中 其他LLM :foundation model(abstract)

AI总结 LIEREx通过整合视觉-语言基础模型与3D语义场景图,实现机器人在部分未知环境中的目标导向探索。

Comments This preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this article is published in KI - Künstliche Intelligenz, and is available online at https://doi.org/10.1007/s13218-026-00902-6

Journal ref Künstliche Intelligenz (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.07114 2026-02-03 cs.CV 50%

Semantically Guided Dynamic Visual Prototype Refinement for Compositional Zero-Shot Learning

语义引导的动态视觉原型精炼用于组合零样本学习

Zhong Peng, Yishi Xu, Gerong Wang, Wenchao Chen, Bo Chen, Jing Zhang, Hongwei Liu

机构 * National Key Laboratory of Radar Signal Processing, Xidian University(雷达信号处理国家级重点实验室,西安电子科技大学) Research Institute of Systems Engineering, Academy of Military Science(系统工程研究所,军事科学院)

专题命中 其他LLM :language model(abstract)

AI总结 Duplex通过动态视觉原型精炼和双原型学习,提升组合零样本学习中语义判别性和泛化能力。

Comments Accepted for publication in Neurocomputing

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01305 2026-02-03 cs.CV 50%

StoryState: Agent-Based State Control for Consistent and Editable Storybooks

StoryState: 基于代理的状态控制用于一致且可编辑的故事书

Ayushman Sarkar, Zhenyu Yu, Wei Tang, Chu Chen, Kangning Cui, Mohd Yamani Idna Idris

机构 * Birbhum Institute of Engineering and Technology(比尔布尔工程科技学院) Universiti Malaya(马来亚大学) City University of Hong Kong(香港城市大学) Wake Forest University(威克森林大学)

专题命中 其他LLM :LLM(abstract)

AI总结 StoryState通过基于代理的状态控制实现多页故事书的一致性与编辑优化

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00531 2026-02-03 cs.CV 50%

Enhancing Open-Vocabulary Object Detection through Multi-Level Fine-Grained Visual-Language Alignment

通过多级细粒度视觉-语言对齐增强开放词汇物体检测

Tianyi Zhang, Antoine Simoulin, Kai Li, Sana Lakdawala, Shiqing Yu, Arpit Mittal, Hongyu Fu, Yu Lin

机构 * University of Minnesota(明尼苏达大学) Meta

专题命中 其他LLM :language model(abstract)

AI总结 VLDet通过多级细粒度视觉-语言对齐提升开放词汇物体检测性能,引入VL-PUB模块和SigRPN块,实现新类别检测的显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00409 2026-02-03 cs.SE 50%

Are Coding Agents Generating Over-Mocked Tests? An Empirical Study

编码代理是否生成过度的模拟测试?一项实证研究

Andre Hora, Romain Robbes

专题命中 其他LLM :LLM(abstract)

AI总结 本文研究了编码代理生成测试中模拟的使用情况,发现代理更倾向于修改测试并添加模拟,同时指出模拟测试可能更容易自动生成但验证效果较弱。

Comments Accepted for publication at MSR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏