arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-26 至 2025-11-26 共收录 22 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 22 篇

2508.09621 2025-11-26 cs.RO cs.AI cs.LG 88%

Interpretable Robot Control via Structured Behavior Trees and Large Language Models

通过结构化行为树和大语言模型实现可解释的机器人控制

Ingrid Maéva Chekam, Ines Pastor-Martinez, Ali Tourani, Jose Andres Millan-Romera, Laura Ribeiro, Pedro Miguel Bastos Soares, Holger Voos, Jose Luis Sanchez-Lopez

机构 * Automation and Robotics Research Group (ARG), Interdisciplinary Centre for Security, Reliability, and Trust (SnT), University of Luxembourg(自动化与机器人研究组(ARG)、安全、可靠性与信任跨学科中心(SnT)、卢森堡大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出结合大语言模型与结构化行为树的方法,实现机器人对自然语言指令的可解释执行,提升人机交互的实用性与适应性。

Comments 15 pages, 5 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20177 2025-11-26 cs.IR 86%

Enhancing Sequential Recommendation with World Knowledge from Large Language Models

利用大语言模型的世界知识增强序列推荐

Tianjie Dai, Xu Chen, Yunmeng Shu, Jinsong Lan, Xiaoyong Zhu, Jiangchao Yao, Bo Zheng

专题命中 领域大模型 :large language model(title);language model(title);LLM(abstract)

AI总结 GRASP通过整合生成增强检索和整体注意力机制,利用大语言模型的世界知识提升序列推荐性能,有效缓解幻觉噪声影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19940 2025-11-26 cs.HC 85%

Editing with AI: How Doctors Refine LLM-Generated Answers to Patient Queries

利用AI编辑:医生如何精炼LLM生成的患者问题回答

Rahul Sharma, Pragnya Ramjee, Kaushik Murali, Mohit Jain

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文研究医生如何利用AI编辑LLM生成的回答,发现情境化是主要编辑方式,间接编辑虽省力但易出错,直接编辑虽精确但工作量大。

Comments 9 pages, 2 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20100 2025-11-26 cs.DC cs.CL 79%

QiMeng-Kernel: Macro-Thinking Micro-Coding Paradigm for LLM-Based High-Performance GPU Kernel Generation

QiMeng-Kernel:基于宏思维微编码的LLM高性能GPU内核生成范式

Xinguo Zhu, Shaohui Peng, Jiaming Guo, Yunji Chen, Qi Guo, Yuanbo Wen, Hang Qin, Ruizhi Chen, Qirui Zhou, Ke Gao, Yanjun Wu, Chen Zhao, Ling Li

专题命中 领域大模型 :LLM(title,abstract);分类 cs.CL

AI总结 QiMeng-Kernel通过宏思维微编码范式,结合强化学习和通用LLM,实现高性能GPU内核生成,准确率和效率均优于现有方法。

Comments 9 pages, 2 figures, accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01265 2025-11-26 cs.CL 77%

AraFinNews: Arabic Financial Summarisation with Domain-Adapted LLMs

AraFinNews: 基于领域适应大语言模型的阿拉伯语金融摘要

Mo El-Haj, Paul Rayson

机构 * School of Computing(计算学院) Lancaster University(兰卡斯特大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);pretraining(abstract);分类 cs.CL

AI总结 AraFinNews通过领域适应大语言模型提升阿拉伯语金融文本摘要的连贯性和准确性。

Comments 9 pages

Journal ref IEEE BigData 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19546 2025-11-26 cs.AI cs.CL cs.HC 76%

CNS-Obsidian: A Neurosurgical Vision-Language Model Built From Scientific Publications

CNS-Obsidian:基于科学出版物构建的神经外科视觉-语言模型

Anton Alyakin, Jaden Stryker, Daniel Alexander Alber, Jin Vivian Lee, Karl L. Sangwon, Brandon Duderstadt, Akshay Save, David Kurland, Spencer Frome, Shrutika Singh, Jeff Zhang, Eunice Yang, Ki Yun Park, Cordelia Orillac, Aly A. Valliani, Sean Neifert, Albert Liu, Aneek Patel, Christopher Livia, Darryl Lau, Ilya Laufer, Peter A. Rozman, Eveline Teresa Hidalgo, Howard Riina, Rui Feng, Todd Hollon, Yindalon Aphinyanaphongs, John G. Golfinos, Laura Snyder, Eric Leuthardt, Douglas Kondziolka, Eric Karl Oermann

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

AI总结 CNS-Obsidian是一款基于科学文献训练的神经外科视觉-语言模型,通过对比实验展示了其在诊断准确性与用户评分上的表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20773 2025-11-26 cs.IR 75%

Adaptive Candidate Retrieval with Dynamic Knowledge Graph Construction for Cold-Start Recommendation

基于动态知识图谱构建的自适应候选检索用于冷启动推荐

Wooseong Yang, Weizhi Zhang, Yuqing Liu, Yuwei Han, Yu Wang, Junhyun Lee, Philip S. Yu

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 ColdRAG通过动态构建知识图谱和LLM引导的多跳推理,提升冷启动推荐性能。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19987 2025-11-26 cs.CL cs.IR 74%

$\text{R}^2\text{R}$: A Route-to-Rerank Post-Training Framework for Multi-Domain Decoder-Only Rerankers

R²R:多领域解码器-only重排序框架的路由到重排序后训练方法

Xinyu Wang, Hanwei Wu, Qingchen Hu, Zhenghan Tai, Jingrui Tian, Lei Ding, Jijun Chi, Hailin He, Tung Sum Thomas Kwok, Yufei Cui, Sicheng Lyu, Muzhi Li, Mingze Li, Xinyue Yu, Ling Zhou, Peng Lu

机构 * McGill University(麦吉尔大学) University of Toronto(多伦多大学) University of Manitoba(曼尼托巴大学) Mila CUHK(中国科技大学) Université de Montréal(蒙特利尔大学) CG Matrix(CG矩阵)

专题命中 领域大模型 :post-training(title);分类 cs.CL

AI总结 R²R通过动态专家路由和两阶段训练策略,提升多领域解码器重排序器的领域适应能力与跨领域鲁棒性。

Comments 13 pages, including 3 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17681 2025-11-26 cs.LG cs.AI 73%

Unlearning as Ablation: Toward a Falsifiable Benchmark for Generative Scientific Discovery

反例作为消去:面向生成科学发现的可证伪基准

Robert Yang

机构 * S6 Research(S6研究)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出'反例作为消去'作为可证伪基准,旨在检验AI在科学发现中的生成能力,通过系统移除目标结果并评估模型能否重新推导,推动AI-科学的基准发展。

Comments 6 pages + appendix. Accepted to NeurIPS 2025 AI4Science Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14191 2025-11-26 cs.CY cs.AI 70%

Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education

面向本体的RAG用于提升网络安全教育中的问答系统

Chengshuai Zhao, Garima Agrawal, Fan Zhang, Tharindu Kumarage, Zhen Tan, Yuli Deng, Ying-Chih Chen, Huan Liu

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出CyberRAG,一种面向本体的检索增强生成方法,用于提升网络安全教育中问答系统的可靠性与准确性。

Comments Accepted by the 2025 IEEE International Conference on Big Data (IEEE BigData 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19914 2025-11-26 cs.RO 67%

CoC-VLA: Delving into Adversarial Domain Transfer for Explainable Autonomous Driving via Chain-of-Causality Visual-Language-Action Model

CoC-VLA: 深入研究对抗域转移以实现可解释的自动驾驶 via 链式因果视觉语言动作模型

Dapeng Zhang, Fei Shen, Rui Zhao, Yinda Chen, Peng Zhi, Chenyang Li, Rui Zhou, Qingguo Zhou

机构 * Lanzhou University(兰州大学) National University of Singapore(新加坡国立大学) University of Science and Technology of China(中国科学技术大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 CoC-VLA通过链式因果视觉语言动作模型实现对抗域转移,提升自动驾驶的可解释性和长尾场景处理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19739 2025-11-26 cs.CL cs.LG 62%

Comparative Analysis of LoRA-Adapted Embedding Models for Clinical Cardiology Text Representation

LoRA适应嵌入模型在临床心脏病文本表示中的比较分析

Richard J. Young, Alice M. Matthews

机构 * University of Nevada Las Vegas(内华达大学拉斯维加斯分校) Concorde Career Colleges(康科德职业学院)

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.LG

AI总结 本研究通过LoRA微调比较了多种嵌入模型在心脏病文本表示中的性能,发现仅编码器架构在领域特定表现上优于更大解码器模型,并提供了临床NLP开发的实用指导。

Comments 25 pages, 13 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03068 2025-11-26 cs.LG cs.AI 62%

Domain Fusion Controllable Generalization for Cross-Domain Time Series Forecasting from Multi-Domain Integrated Distribution

多域集成分布下的领域融合可控泛化用于跨领域时间序列预测

Xiangkai Ma, Xiaobin Hong, Mingkai Lin, Han Zhang, Wenzhong Li, Sanglu Lu

机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI、cs.LG

AI总结 本文提出TimeControl模型,通过领域融合范式整合多领域时间序列信息,利用扩散模型实现跨领域时间序列预测的可控泛化。

Comments We have updated the abstract, introduction and related work. Additionally, we have incorporated the latest competitive baseline models

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20650 2025-11-26 cs.CV cs.AI 57%

MedROV: Towards Real-Time Open-Vocabulary Detection Across Diverse Medical Imaging Modalities

MedROV:面向跨多种医学影像模态的实时开放词汇检测

Tooba Tehreem Sheikh, Jean Lahoud, Rao Muhammad Anwer, Fahad Shahbaz Khan, Salman Khan, Hisham Cholakkal

机构 * Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 MedROV是首个实时开放词汇医学影像检测模型,通过大规模数据集和伪标签策略提升检测性能,实现40 mAP50的提升并达到70 FPS的实时处理速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20109 2025-11-26 cs.LG 57%

CLIMATEAGENT: Multi-Agent Orchestration for Complex Climate Data Science Workflows

CLIMATEAGENT: 多智能体编排用于复杂气候数据科学工作流

Hyeonjae Kim, Chenyue Li, Wen Deng, Mengxi Jin, Wen Huang, Mengqian Lu, Binhang Yuan

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

AI总结 ClimateAgent通过多智能体编排和动态API意识,实现端到端自动化气候数据分析,取得优于基线方法的高任务完成率和报告质量。

Comments 30 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00721 2025-11-26 eess.IV cs.CV cs.LG eess.SP 57%

FMPlug: Plug-In Foundation Flow-Matching Priors for Inverse Problems

FMPlug: 用于逆问题的插件基础流匹配先验

Yuxiang Wan, Ryan Devera, Wenjie Zhang, Ju Sun

机构 * Department of Computer Science and Engineering, University of Minnesota(计算机科学与工程系,明尼苏达大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

AI总结 FMPlug通过引入时间自适应预热策略和高斯性正则化,提升基础流匹配先验在逆问题中的性能,实现图像超分辨率和去模糊任务的显著改进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12051 2025-11-26 cs.CL 57%

MedS$^3$: Towards Medical Slow Thinking with Self-Evolved Soft Dual-sided Process Supervision

MedS$^3$: 向医疗慢思考迈进的自演化软双过程监督

Shuyang Jiang, Yusheng Liao, Zhe Chen, Ya Zhang, Yanfeng Wang, Yu Wang

机构 * footnotemark: 1 Corresponding Authors(通讯作者)

专题命中 领域大模型 :language model(abstract);分类 cs.CL

AI总结 MedS3通过自演化框架和软双过程奖励模型,提升医疗领域小模型的推理能力,实验显示其在准确率上优于现有最佳模型。

Comments 20 pages;Accepted as a Main paper at AAAI26

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19471 2025-11-26 eess.IV cs.AI cs.CV 57%

Not Quite Anything: Overcoming SAMs Limitations for 3D Medical Imaging

并非一切:克服SAMs在3D医学影像中的局限性

Keith Moore

机构 * Deptartment of Biomedical Data Science Stanford University(生物医学数据科学系 斯坦福大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 本文提出了一种无需微调基础模型的组合性替代方案,通过将基础模型输出作为额外输入通道来提高3D医学影像分割的准确性与鲁棒性。

Comments Preprint; Paper accepted at AIAS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20245 2025-11-26 cs.CV physics.optics 50%

HistoSpeckle-Net: Mutual Information-Guided Deep Learning for high-fidelity reconstruction of complex OrganAMNIST images via perturbed Multimode Fibers

HistoSpeckle-Net:基于互信息的深度学习用于通过扰动多模光纤重建复杂OrganAMNIST图像

Jawaria Maqbool, M. Imran Cheema

机构 * Department of Electrical Engineering, Syed Babar Ali School of Science and Engineering, Lahore University of Management Sciences(电气工程系,Syed Babar Ali科学与工程学院,拉合尔管理科学大学)

专题命中 领域大模型 :SLM(abstract)

AI总结 HistoSpeckle-Net通过互信息损失和三尺度特征细化模块,提升多模光纤在复杂医学图像重建中的性能和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19979 2025-11-26 cs.IR 50%

The 2nd Workshop on Human-Centered Recommender Systems

人类中心推荐系统研讨会第二届

Kaike Zhang, Jiakai Tang, Du Su, Shuchang Liu, Julian McAuley, Lina Yao, Qi Cao, Yue Feng, Fei Sun

专题命中 领域大模型 :LLM(abstract)

AI总结 该研讨会旨在推动推荐系统从优化参与度向设计真正理解、参与和惠及人类的系统转变,探讨如何整合人类价值观以提升推荐系统的社会责任感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19846 2025-11-26 cs.CV 50%

Face, Whole-Person, and Object Classification in a Unified Space Via The Interleaved Multi-Domain Identity Curriculum

通过交错多域身份课程实现面部、全身和物体分类的统一空间

Thomas M Metz, Matthew Q Hill, Alice J O'Toole

机构 * School of Behavioral and Brain Sciences, The University of Texas at Dallas(行为与脑科学学院,德克萨斯大学达拉斯分校)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本文提出了一种通过交错多域身份课程实现面部、全身和物体分类的统一空间模型,有效解决灾难性遗忘问题,并在多任务处理中优于人类。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16181 2025-11-26 cs.CV 50%

CLIP-IT: CLIP-based Pairing for Histology Images Classification

CLIP-IT: 基于CLIP的病理图像分类配对

Banafsheh Karimian, Giulia Avanzato, Soufian Belharbi, Alexis Guichemerre, Luke McCaffrey, Mohammadhadi Shateri, Eric Granger

机构 * LIVIA, ILLS, Dept. of Systems Engineering, ETS Montreal, Canada(LIVIA、ILLs、系统工程系、蒙特利尔大学ETSMontreal加拿大) Dept. of Computer Engineering, University of Cagliari, Italy(计算机工程系、卡利亚里大学意大利) Goodman Cancer Research, Centre, Dept. of Oncology, McGill University, Canada(Goodman癌症研究中心、肿瘤学系、麦吉尔大学加拿大)

专题命中 领域大模型 :language model(abstract)

AI总结 CLIP-IT通过利用未配对的病理报告提升病理图像分类性能,无需配对数据或复杂推理。

详情

展开后加载摘要…

URL PDF HTML 收藏