arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-09 至 2026-01-09 共收录 20 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 20 篇

2601.05187 2026-01-09 cs.AI 87%

SimuAgent: An LLM-Based Simulink Modeling Assistant Enhanced with Reinforcement Learning

SimuAgent: 一种基于LLM的Simulink建模助手,结合强化学习增强

Yanchang Liang, Xiaowei Zhao

机构 * School of Engineering, University of Warwick(战争学院) Intelligent Control and Smart Energy (ICSE) Research Group(智能控制与智能能源研究组)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 SimuAgent结合强化学习和LLM,提升Simulink建模效率与准确性,实现快速收敛和高鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04801 2026-01-09 cs.AR cs.LG 87%

MPM-LLM4DSE: Reaching the Pareto Frontier in HLS with Multimodal Learning and LLM-Driven Exploration

MPM-LLM4DSE: 在HLS中通过多模态学习和LLM驱动探索达到帕累托前沿

Lei Xu, Shanshan Wang, Chenglong Xiao

机构 * Department of Computer Science, Shantou University, China(计算机科学系,汕头大学,中国)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文提出 MPM-LLM4DSE 框架,结合多模态学习与大语言模型驱动探索,提升 HLS 设计空间探索的性能和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05547 2026-01-09 cs.CV cs.AI 85%

Automated Invoice Data Extraction: Using LLM and OCR

自动化发票数据提取:利用大语言模型和OCR

Khushi Khanchandani, Advait Thakur, Akshita Shetty, Chaitravi Reddy, Ritisa Behera

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一个结合OCR、深度学习、LLMs和图分析的AI平台,以提高发票数据提取的质量和一致性。

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04209 2026-01-09 cs.CL 83%

Leveraging Language Models and RAG for Efficient Knowledge Discovery in Clinical Environments

利用语言模型和RAG实现临床环境中的高效知识发现

Seokhwan Ko, Donghyeon Lee, Jaewoo Chun, Hyungsoo Han, Junghwan Cho

机构 * Clinical Omics Institute, Kyungpook National University(临床组学研究所,庆北国立大学) Department of Biomedical Science, School of Medicine Kyungpook National University(生物医学科学系,庆北国立大学医学院) Department of Physiology, School of Medicine Kyungpook National University(生理学系,庆北国立大学医学院)

专题命中 领域大模型 :language model(title,abstract);large language model(abstract);分类 cs.CL

AI总结 本研究提出一种基于RAG的系统,利用本地部署的LLaMA3和PubMedBERT,在临床环境中实现高效生物医学知识发现。

Comments 11pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04195 2026-01-09 cs.CL cs.AI 79%

MedPI: Evaluating AI Systems in Medical Patient-facing Interactions

MedPI:评估医疗患者交互中的人工智能系统

Diego Fajardo V., Oleksii Proniakin, Victoria-Elisabeth Gruber, Razvan Marinescu

机构 * Lumos

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 MedPI是一个评估医疗患者交互中AI系统表现的高维基准,通过105个维度评估医疗对话质量,发现主流LLM在鉴别诊断上表现不佳。

Comments 24 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15993 2026-01-09 eess.SP cs.IT math.IT 78%

Filter-and-Attend: Wireless Channel Foundation Model with Noise-Plus-Interference Suppression Structure

滤波与关注:具有噪声加干扰抑制结构的无线信道基础模型

Yuwei Wang, Li Sun, Tingting Yang, Yuxuan Shi, Maged Elkashlan, Xiao Tang

专题命中 领域大模型 :foundation model(title,abstract)

AI总结 本文提出Filter-and-Attend范式,通过滤波抑制噪声加干扰和关注机制补全CSI,提升无线信道基础模型在多种下游任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04940 2026-01-09 cs.CR cs.AI 77%

CurricuLLM: Designing Personalized and Workforce-Aligned Cybersecurity Curricula Using Fine-Tuned LLMs

CurricuLLM: 利用微调大语言模型设计个性化且与劳动力市场对齐的网络安全课程

Arthur Nijdam, Harri Kähkönen, Valtteri Niemi, Paul Stankovski Wagner, Sara Ramezanian

机构 * Lund University, Department of Electrical University of Helsinki, Department of Computer Science, Helsinki, Finland Helsinki Institute for Information Technology, Helsinki, Finland Karlstad University, Department of Mathematics \& Computer Science, Karlstad, Sweden

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 CurricuLLM通过微调大语言模型设计个性化且与劳动力市场对齐的网络安全课程,解决课程与行业需求不匹配的问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04690 2026-01-09 cs.LG 77%

Do LLMs Benefit from User and Item Embeddings in Recommendation Tasks?

在推荐任务中,大语言模型是否受益于用户和物品嵌入?

Mir Rayat Imtiaz Hossain, Leo Feng, Leonid Sigal, Mohamed Osama Ahmed

机构 * University of British Columbia(不列颠哥伦比亚大学) RBC Borealis

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出通过投影用户和物品嵌入到LLM token空间,提升推荐性能,实现传统推荐系统与LLM的结合。

Comments Presented in Multimodal Algorithmic Reasoning Workshop at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04531 2026-01-09 cs.IR cs.AI 77%

Self-MedRAG: a Self-Reflective Hybrid Retrieval-Augmented Generation Framework for Reliable Medical Question Answering

Self-MedRAG:一种用于可靠医学问答的自反思混合检索增强生成框架

Jessica Ryan, Alexander I. Gumilang, Robert Wiliam, Derwin Suhartono

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 Self-MedRAG通过混合检索与自反思机制提升医学问答的准确性和可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04217 2026-01-09 cs.CY cs.AI 77%

Attachment Styles and AI Chatbot Interactions Among College Students

依附风格与大学生与AI聊天机器人互动

Ziqi Lin, Taiyu Hou

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究探讨了大学生依附风格如何影响其与AI聊天机器人的互动模式,发现依附风格决定了学生对AI作为情感空间、支持工具及亲密关系的认知与使用方式。

Comments 15 pages, 1 table, 2 appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04878 2026-01-09 cs.AI cond-mat.mtrl-sci cs.CL cs.LG 75%

Higher-Order Knowledge Representations for Agentic Scientific Reasoning

高阶知识表示用于代理科学推理

Isabella A. Stewart, Markus J. Buehler

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出基于超图的知识表示方法,用于代理科学推理,通过高阶路径生成机理假设,提升科学发现效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10230 2026-01-09 eess.IV cs.CV 75%

Leveraging Clinical Text and Class Conditioning for 3D Prostate MRI Generation

利用临床文本和类别条件化生成3D前列腺MRI

Emerson P. Grabke, Babak Taati, Masoom A. Haider

专题命中 领域大模型 :large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 本文提出CCELLA方法,通过结合临床文本和类别条件化,提升3D前列腺MRI生成质量及下游分类器性能。

Comments Accepted for publication in IEEE Transactions on Biomedical Engineering, 2025. This is the accepted author version. The final published version is available at https://doi.org/10.1109/TBME.2025.3648426

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04199 2026-01-09 cs.LG cs.AI cs.CL 75%

The Forgotten Shield: Safety Grafting in Parameter-Space for Medical MLLMs

被遗忘的盾牌:参数空间中的医疗大语言模型安全性移植

Jiale Zhao, Xing Mou, Jinlin Wu, Hongyuan Yu, Mingrui Sun, Yang Shi, Xuanwu Yin, Zhen Chen, Zhen Lei, Yaohua Wang

机构 * National University of Defense Technology(国防科技大学) Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统(MAIS)) Multimedia Department, Xiaomi Inc(小米公司多媒体部门) Centre for Artificial Intelligence and Robotics, Hong Kong Institute of Science and Innovation, Chinese Academy of Sciences, Hong Kong(香港科学院人工智能与机器人中心) School of Artificial Intelligence, University of Chinese Academy of Sciences, UCAS(中国科学院大学人工智能学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出参数空间干预方法,通过提取原始模型的安全知识并注入目标模型,提升医疗大语言模型的安全性,同时保持医疗性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05051 2026-01-09 cs.AI cs.CL cs.DL cs.IT math.IT 73%

Publishing FAIR and Machine-actionable Reviews in Materials Science: The Case for Symbolic Knowledge in Neuro-symbolic Artificial Intelligence

发布符合FAIR标准且可被机器执行的材料科学评论:神经符号人工智能中符号知识的案例

Jennifer D'Souza, Soren Auer, Eleni Poupaki, Alex Watkins, Anjana Devi, Riikka L. Puurunen, Bora Karasulu, Adrie Mackus, Erwin Kessels

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过FAIR标准和ORKG发布可机器执行的材料科学评论,强调符号层在神经符号AI中的核心作用。

Comments 35 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18526 2026-01-09 cs.CL cs.AI cs.DL 73%

SciClaims: An End-to-End Generative System for Biomedical Claim Analysis

SciClaims: 一种用于生物医学声明分析的端到端生成系统

Raúl Ortega, José Manuel Gómez-Pérez

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 SciClaims是一种基于大语言模型的端到端生物医学声明分析系统,可自动提取声明、检索证据并验证真实性,无需额外微调。

Comments In Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing: System Demonstrations

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04758 2026-01-09 cs.CL cs.AI 73%

PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks

PILOT-Bench:一个以专利领域法律推理为核心的基准,包含与IRAC对齐的分类任务

Yehoon Jang, Chaewon Lee, Hyun-seok Min, Sungchul Choi

机构 * Pukyong National University(浦项国立大学) Tomocube Inc.(Tomocube公司)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 PILOT-Bench通过IRAC对齐的分类任务评估专利领域法律推理能力,揭示闭源与开源模型在推理性能上的显著差异。

Comments Accepted at the NLLP Workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04597 2026-01-09 cs.CL 70%

THaLLE-ThaiLLM: Domain-Specialized Small LLMs for Finance and Thai -- Technical Report

THaLLE-ThaiLLM:面向金融和泰语的领域专用小型LLM——技术报告

KBTG Labs, :, Anuruth Lertpiya, Danupat Khamnuansin, Kantapong Sucharitpongpan, Pornchanan Balee, Tawunrat Chalothorn, Thadpong Pongthawornkamol, Monchai Lertsutthiwong

机构 * NLP-Voice Research Lab(自然语言处理-语音研究实验室) KBTG Labs(KBTG实验室) KASIKORN Business—Technology Group(Kasikorn 业务-技术集团)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 THaLLE-ThaiLLM通过模型合并技术,开发了面向金融和泰语的专用小型LLM,提升了泰语通用能力和金融领域性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04219 2026-01-09 cs.CY cs.AI cs.MA 70%

AgentTutor: Empowering Personalized Learning with Multi-Turn Interactive Teaching in Intelligent Education Systems

AgentTutor: 通过多轮互动教学赋能个性化学习在智能教育系统中

Yuxin Liu, Zeqing Song, Jiong Lou, Chentao Wu, Jie Li

专题命中 领域大模型 :LLM(abstract);language model(abstract);分类 cs.AI

AI总结 AgentTutor通过多轮互动教学系统提升个性化学习效果,结合多代理系统和学习者档案实现动态优化的教学策略。

Comments AAAI2026 Workshop AI4EDU

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04206 2026-01-09 cs.CL cs.CY cs.HC 57%

Enhancing Admission Inquiry Responses with Fine-Tuned Models and Retrieval-Augmented Generation

通过微调模型和检索增强生成提升录取咨询回复

Aram Virabyan

专题命中 领域大模型 :language model(abstract);分类 cs.CL

AI总结 本文提出通过微调模型和检索增强生成技术,提升大学招生咨询回复的准确性和效率,以满足招生沟通的特殊需求。

Comments 9 pages, 1 figure, 1 table. Proceedings of the 19th International Scientific Conference "Parallel Computing Technologies" (PCT'2025), Moscow, Russia

Journal ref Proc. 19th International Scientific Conference "Parallel Computing Technologies" (PCT'2025), South Ural State University, 2025, pp. 99-106

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06282 2026-01-09 cs.CV 50%

From Dataset to Real-world: General 3D Object Detection via Generalized Cross-domain Few-shot Learning

从数据集到现实世界:通过通用跨领域少样本学习实现通用3D目标检测

Shuangzhi Li, Junlong Shen, Lei Ma, Xingyu Li

专题命中 领域大模型 :language model(abstract)

AI总结 本文提出通用跨领域少样本学习方法,通过融合2D语义与3D空间推理,实现对现实世界中常见和新类目标的高效检测。

Comments The latest version refines the few-shot setting on common classes, enforcing a stricter object-level definition

详情

展开后加载摘要…

URL PDF HTML 收藏