arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-26 至 2025-11-26 共收录 191 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 22 篇

2511.01265 2025-11-26 cs.CL 77%

AraFinNews: Arabic Financial Summarisation with Domain-Adapted LLMs

AraFinNews: 基于领域适应大语言模型的阿拉伯语金融摘要

Mo El-Haj, Paul Rayson

机构 * School of Computing(计算学院) Lancaster University(兰卡斯特大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);pretraining(abstract);分类 cs.CL

AI总结 AraFinNews通过领域适应大语言模型提升阿拉伯语金融文本摘要的连贯性和准确性。

Comments 9 pages

Journal ref IEEE BigData 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19546 2025-11-26 cs.AI cs.CL cs.HC 76%

CNS-Obsidian: A Neurosurgical Vision-Language Model Built From Scientific Publications

CNS-Obsidian:基于科学出版物构建的神经外科视觉-语言模型

Anton Alyakin, Jaden Stryker, Daniel Alexander Alber, Jin Vivian Lee, Karl L. Sangwon, Brandon Duderstadt, Akshay Save, David Kurland, Spencer Frome, Shrutika Singh, Jeff Zhang, Eunice Yang, Ki Yun Park, Cordelia Orillac, Aly A. Valliani, Sean Neifert, Albert Liu, Aneek Patel, Christopher Livia, Darryl Lau, Ilya Laufer, Peter A. Rozman, Eveline Teresa Hidalgo, Howard Riina, Rui Feng, Todd Hollon, Yindalon Aphinyanaphongs, John G. Golfinos, Laura Snyder, Eric Leuthardt, Douglas Kondziolka, Eric Karl Oermann

专题命中 领域大模型 :language model(title);分类 cs.CL、cs.AI

AI总结 CNS-Obsidian是一款基于科学文献训练的神经外科视觉-语言模型,通过对比实验展示了其在诊断准确性与用户评分上的表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20773 2025-11-26 cs.IR 75%

Adaptive Candidate Retrieval with Dynamic Knowledge Graph Construction for Cold-Start Recommendation

基于动态知识图谱构建的自适应候选检索用于冷启动推荐

Wooseong Yang, Weizhi Zhang, Yuqing Liu, Yuwei Han, Yu Wang, Junhyun Lee, Philip S. Yu

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 ColdRAG通过动态构建知识图谱和LLM引导的多跳推理,提升冷启动推荐性能。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19987 2025-11-26 cs.CL cs.IR 74%

$\text{R}^2\text{R}$: A Route-to-Rerank Post-Training Framework for Multi-Domain Decoder-Only Rerankers

R²R:多领域解码器-only重排序框架的路由到重排序后训练方法

Xinyu Wang, Hanwei Wu, Qingchen Hu, Zhenghan Tai, Jingrui Tian, Lei Ding, Jijun Chi, Hailin He, Tung Sum Thomas Kwok, Yufei Cui, Sicheng Lyu, Muzhi Li, Mingze Li, Xinyue Yu, Ling Zhou, Peng Lu

机构 * McGill University(麦吉尔大学) University of Toronto(多伦多大学) University of Manitoba(曼尼托巴大学) Mila CUHK(中国科技大学) Université de Montréal(蒙特利尔大学) CG Matrix(CG矩阵)

专题命中 领域大模型 :post-training(title);分类 cs.CL

AI总结 R²R通过动态专家路由和两阶段训练策略,提升多领域解码器重排序器的领域适应能力与跨领域鲁棒性。

Comments 13 pages, including 3 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17681 2025-11-26 cs.LG cs.AI 73%

Unlearning as Ablation: Toward a Falsifiable Benchmark for Generative Scientific Discovery

反例作为消去:面向生成科学发现的可证伪基准

Robert Yang

机构 * S6 Research(S6研究)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出'反例作为消去'作为可证伪基准,旨在检验AI在科学发现中的生成能力,通过系统移除目标结果并评估模型能否重新推导,推动AI-科学的基准发展。

Comments 6 pages + appendix. Accepted to NeurIPS 2025 AI4Science Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14191 2025-11-26 cs.CY cs.AI 70%

Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education

面向本体的RAG用于提升网络安全教育中的问答系统

Chengshuai Zhao, Garima Agrawal, Fan Zhang, Tharindu Kumarage, Zhen Tan, Yuli Deng, Ying-Chih Chen, Huan Liu

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出CyberRAG,一种面向本体的检索增强生成方法,用于提升网络安全教育中问答系统的可靠性与准确性。

Comments Accepted by the 2025 IEEE International Conference on Big Data (IEEE BigData 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19914 2025-11-26 cs.RO 67%

CoC-VLA: Delving into Adversarial Domain Transfer for Explainable Autonomous Driving via Chain-of-Causality Visual-Language-Action Model

CoC-VLA: 深入研究对抗域转移以实现可解释的自动驾驶 via 链式因果视觉语言动作模型

Dapeng Zhang, Fei Shen, Rui Zhao, Yinda Chen, Peng Zhi, Chenyang Li, Rui Zhou, Qingguo Zhou

机构 * Lanzhou University(兰州大学) National University of Singapore(新加坡国立大学) University of Science and Technology of China(中国科学技术大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 CoC-VLA通过链式因果视觉语言动作模型实现对抗域转移,提升自动驾驶的可解释性和长尾场景处理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19739 2025-11-26 cs.CL cs.LG 62%

Comparative Analysis of LoRA-Adapted Embedding Models for Clinical Cardiology Text Representation

LoRA适应嵌入模型在临床心脏病文本表示中的比较分析

Richard J. Young, Alice M. Matthews

机构 * University of Nevada Las Vegas(内华达大学拉斯维加斯分校) Concorde Career Colleges(康科德职业学院)

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.LG

AI总结 本研究通过LoRA微调比较了多种嵌入模型在心脏病文本表示中的性能,发现仅编码器架构在领域特定表现上优于更大解码器模型,并提供了临床NLP开发的实用指导。

Comments 25 pages, 13 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03068 2025-11-26 cs.LG cs.AI 62%

Domain Fusion Controllable Generalization for Cross-Domain Time Series Forecasting from Multi-Domain Integrated Distribution

多域集成分布下的领域融合可控泛化用于跨领域时间序列预测

Xiangkai Ma, Xiaobin Hong, Mingkai Lin, Han Zhang, Wenzhong Li, Sanglu Lu

机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI、cs.LG

AI总结 本文提出TimeControl模型,通过领域融合范式整合多领域时间序列信息,利用扩散模型实现跨领域时间序列预测的可控泛化。

Comments We have updated the abstract, introduction and related work. Additionally, we have incorporated the latest competitive baseline models

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20650 2025-11-26 cs.CV cs.AI 57%

MedROV: Towards Real-Time Open-Vocabulary Detection Across Diverse Medical Imaging Modalities

MedROV:面向跨多种医学影像模态的实时开放词汇检测

Tooba Tehreem Sheikh, Jean Lahoud, Rao Muhammad Anwer, Fahad Shahbaz Khan, Salman Khan, Hisham Cholakkal

机构 * Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 MedROV是首个实时开放词汇医学影像检测模型,通过大规模数据集和伪标签策略提升检测性能,实现40 mAP50的提升并达到70 FPS的实时处理速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20109 2025-11-26 cs.LG 57%

CLIMATEAGENT: Multi-Agent Orchestration for Complex Climate Data Science Workflows

CLIMATEAGENT: 多智能体编排用于复杂气候数据科学工作流

Hyeonjae Kim, Chenyue Li, Wen Deng, Mengxi Jin, Wen Huang, Mengqian Lu, Binhang Yuan

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

AI总结 ClimateAgent通过多智能体编排和动态API意识,实现端到端自动化气候数据分析,取得优于基线方法的高任务完成率和报告质量。

Comments 30 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00721 2025-11-26 eess.IV cs.CV cs.LG eess.SP 57%

FMPlug: Plug-In Foundation Flow-Matching Priors for Inverse Problems

FMPlug: 用于逆问题的插件基础流匹配先验

Yuxiang Wan, Ryan Devera, Wenjie Zhang, Ju Sun

机构 * Department of Computer Science and Engineering, University of Minnesota(计算机科学与工程系,明尼苏达大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

AI总结 FMPlug通过引入时间自适应预热策略和高斯性正则化,提升基础流匹配先验在逆问题中的性能,实现图像超分辨率和去模糊任务的显著改进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12051 2025-11-26 cs.CL 57%

MedS$^3$: Towards Medical Slow Thinking with Self-Evolved Soft Dual-sided Process Supervision

MedS$^3$: 向医疗慢思考迈进的自演化软双过程监督

Shuyang Jiang, Yusheng Liao, Zhe Chen, Ya Zhang, Yanfeng Wang, Yu Wang

机构 * footnotemark: 1 Corresponding Authors(通讯作者)

专题命中 领域大模型 :language model(abstract);分类 cs.CL

AI总结 MedS3通过自演化框架和软双过程奖励模型,提升医疗领域小模型的推理能力,实验显示其在准确率上优于现有最佳模型。

Comments 20 pages;Accepted as a Main paper at AAAI26

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19471 2025-11-26 eess.IV cs.AI cs.CV 57%

Not Quite Anything: Overcoming SAMs Limitations for 3D Medical Imaging

并非一切:克服SAMs在3D医学影像中的局限性

Keith Moore

机构 * Deptartment of Biomedical Data Science Stanford University(生物医学数据科学系 斯坦福大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 本文提出了一种无需微调基础模型的组合性替代方案,通过将基础模型输出作为额外输入通道来提高3D医学影像分割的准确性与鲁棒性。

Comments Preprint; Paper accepted at AIAS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20245 2025-11-26 cs.CV physics.optics 50%

HistoSpeckle-Net: Mutual Information-Guided Deep Learning for high-fidelity reconstruction of complex OrganAMNIST images via perturbed Multimode Fibers

HistoSpeckle-Net:基于互信息的深度学习用于通过扰动多模光纤重建复杂OrganAMNIST图像

Jawaria Maqbool, M. Imran Cheema

机构 * Department of Electrical Engineering, Syed Babar Ali School of Science and Engineering, Lahore University of Management Sciences(电气工程系,Syed Babar Ali科学与工程学院,拉合尔管理科学大学)

专题命中 领域大模型 :SLM(abstract)

AI总结 HistoSpeckle-Net通过互信息损失和三尺度特征细化模块,提升多模光纤在复杂医学图像重建中的性能和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19979 2025-11-26 cs.IR 50%

The 2nd Workshop on Human-Centered Recommender Systems

人类中心推荐系统研讨会第二届

Kaike Zhang, Jiakai Tang, Du Su, Shuchang Liu, Julian McAuley, Lina Yao, Qi Cao, Yue Feng, Fei Sun

专题命中 领域大模型 :LLM(abstract)

AI总结 该研讨会旨在推动推荐系统从优化参与度向设计真正理解、参与和惠及人类的系统转变,探讨如何整合人类价值观以提升推荐系统的社会责任感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19846 2025-11-26 cs.CV 50%

Face, Whole-Person, and Object Classification in a Unified Space Via The Interleaved Multi-Domain Identity Curriculum

通过交错多域身份课程实现面部、全身和物体分类的统一空间

Thomas M Metz, Matthew Q Hill, Alice J O'Toole

机构 * School of Behavioral and Brain Sciences, The University of Texas at Dallas(行为与脑科学学院,德克萨斯大学达拉斯分校)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本文提出了一种通过交错多域身份课程实现面部、全身和物体分类的统一空间模型,有效解决灾难性遗忘问题,并在多任务处理中优于人类。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16181 2025-11-26 cs.CV 50%

CLIP-IT: CLIP-based Pairing for Histology Images Classification

CLIP-IT: 基于CLIP的病理图像分类配对

Banafsheh Karimian, Giulia Avanzato, Soufian Belharbi, Alexis Guichemerre, Luke McCaffrey, Mohammadhadi Shateri, Eric Granger

机构 * LIVIA, ILLS, Dept. of Systems Engineering, ETS Montreal, Canada(LIVIA、ILLs、系统工程系、蒙特利尔大学ETSMontreal加拿大) Dept. of Computer Engineering, University of Cagliari, Italy(计算机工程系、卡利亚里大学意大利) Goodman Cancer Research, Centre, Dept. of Oncology, McGill University, Canada(Goodman癌症研究中心、肿瘤学系、麦吉尔大学加拿大)

专题命中 领域大模型 :language model(abstract)

AI总结 CLIP-IT通过利用未配对的病理报告提升病理图像分类性能,无需配对数据或复杂推理。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 10 篇

2510.15691 2025-11-26 q-fin.CP cs.AI cs.CL cs.LG 90%

Exploring the Synergy of Quantitative Factors and Newsflow Representations from Large Language Models for Stock Return Prediction

探索来自大语言模型的定量因素和新闻流表示的协同效应以预测股票收益率

Tian Guo, Emmanuel Hauptmann

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出融合学习框架和混合模型,利用大语言模型生成的定量因素和新闻流表示,提升股票收益率预测和选择的准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19569 2025-11-26 cs.LG 83%

An Invariant Latent Space Perspective on Language Model Inversion

语言模型反向的不变潜在空间视角

Wentao Ye, Jiaqi Hu, Haobo Wang, Xinpeng Ti, Zhiqing Xiao, Hao Chen, Liyao Li, Lei Feng, Sai Wu, Junbo Zhao

专题命中 知识编辑与模型理解 :language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出Inv^2A方法,通过不变潜在空间假设提升语言模型反向性能,有效对抗隐私和安全威胁。

Comments The Fortieth AAAI Conference on Artificial Intelligence (AAAI-26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19641 2025-11-26 cs.CV cs.AI 79%

On the Utility of Foundation Models for Fast MRI: Vision-Language-Guided Image Reconstruction

在快速MRI中基础模型的效用:基于视觉-语言的图像重建

Ruimin Feng, Xingxin He, Ronald Mercer, Zachary Stewart, Fang Liu

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

AI总结 本文提出利用视觉-语言基础模型通过语义空间优化提升欠采样MRI重建效果,实验表明其在保留解剖结构和提升感知质量方面优于传统方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12964 2025-11-26 cs.CL 70%

MA-COIR: Leveraging Semantic Search Index and Generative Models for Ontology-Driven Biomedical Concept Recognition

MA-COIR: 利用语义搜索索引和生成模型进行面向本体的生物医学概念识别

Shanshan Liu, Noriki Nishida, Rumana Ferdous Munne, Narumi Tokunaga, Yuki Yamagata, Kouji Kozaki, Yuji Matsumoto

机构 * RIKEN AIP University of Tsukuba(茨川大学) RIKEN R-IH RIKEN BRC Osaka Electro-Communication University(大阪电讯大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 MA-COIR通过结合语义搜索索引和生成模型,提升生物医学领域本体驱动的概念识别能力。

Comments preprint

Journal ref https://aclanthology.org/2025.acl-srw.39/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20273 2025-11-26 cs.LG cs.AI cs.CL 67%

Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits

超越组件:基于奇异向量的Transformer电路可解释性

Areeb Ahmad, Abhinav Joshi, Ashutosh Modi

机构 * Indian Institute of Technology Kanpur (IIT Kanpur)(印度理工学院坎浦尔(IIT坎浦尔))

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出基于奇异向量的Transformer电路可解释性方法,揭示了模型内部更分布、结构化和组合化的计算特性。

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17645 2025-11-26 cs.LG cs.AI cs.CL 67%

BlockCert: Certified Blockwise Extraction of Transformer Mechanisms

BlockCert: Transformer机制的认证分块提取

Sandro Andric

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 BlockCert通过认证分块提取Transformer机制,提供可验证的误差界和覆盖率指标,实验证明其在多个模型上有效且具有实际应用价值。

Comments 16 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15214 2025-11-26 q-fin.GN cs.CY 67%

Corporate Earnings Calls and Analyst Beliefs

企业盈利电话会议与分析师信念

Giuseppe Matera

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 研究通过分析企业盈利电话会议中的叙述,发现分析师对乐观情绪过度反应,对风险和不确定性的叙述反应不足,揭示了预期形成的机制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19506 2025-11-26 cs.LG cs.LO 57%

Profile Generators: A Link between the Narrative and the Binary Matrix Representation

生成器:叙事与二进制矩阵表示之间的桥梁

Raoul H. Kutil, Georg Zimmermann, Barbara Strasser-Kirchweger, Christian Borgelt

机构 * Department of AIHI, University of Salzburg(AIHI系,萨尔茨堡大学) Research Programme Biomedical Data Science, Paracelsus Medical University(生物医学数据科学研究计划,帕拉塞尔斯医学大学) Department of Psychology, University of Salzburg(心理学系,萨尔茨堡大学)

专题命中 知识编辑与模型理解 :prompting(abstract);分类 cs.LG

AI总结 本研究提出了一种将DSM-5叙述形式与二进制矩阵表示联系起来的生成器方法,用于高效生成复杂障碍的症状组合。

Comments 31 pages, 8 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20167 2025-11-26 cs.MM 50%

FINE: Factorized multimodal sentiment analysis via mutual INformation Estimation

FINE: 通过互信息估计进行因子化多模态情感分析

Yadong Liu, Shangfei Wang

专题命中 知识编辑与模型理解 :prompting(abstract)

AI总结 本文提出了一种基于互信息估计的因子化多模态情感分析框架,通过分解模态为共享和独特表示,抑制噪声并提升情感表示质量,从而在多个数据集上优于现有方法。

Comments 15 pages, 9 figures, conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17964 2025-11-26 cs.CV 50%

X-ReID: Multi-granularity Information Interaction for Video-Based Visible-Infrared Person Re-Identification

X-ReID:基于视频的可见-红外人重识别中的多粒度信息交互

Chenyang Yu, Xuehu Liu, Pingping Zhang, Huchuan Lu

专题命中 知识编辑与模型理解 :language model(abstract)

AI总结 X-ReID通过多粒度信息交互和跨模态特征学习,提升视频中可见-红外人重识别的性能。

Comments Accepted by AAAI2026. More modifications may be performed

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 13 篇

2511.19550 2025-11-26 cs.IT cs.AI math.IT 85%

The Semiotic Channel Principle: Measuring the Capacity for Meaning in LLM Communication

语义通道原理:测量LLM通信的意义容量

Davide Picca

机构 * University of Lausanne(洛桑大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于语义通道原理的框架,用于测量LLM通信的意义容量,通过优化lambda参数实现表达丰富性与可解码性之间的权衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20080 2025-11-26 cs.HC 85%

Adaptive LLM Agents: Toward Personalized Empathetic Care

自适应大语言模型代理:迈向个性化共情护理

Priyanka Singh, Sebastian Von Mammen

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出基于大语言模型的自适应代理框架,通过AIS量表量化用户心理状态,实现个性化治疗互动,探讨其在心理健康支持中的应用与设计。

Comments Accepted at workshop Future Wellbeing: Using Design Fiction to Explore Human-Agent Interaction and Mental Health at The 13th International Conference on Human-Agent Interaction (HAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏