arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-14 至 2026-01-14 共收录 17 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 17 篇

2512.20002 2026-01-14 cs.LG 92%

LoFT-LLM: Low-Frequency Time-Series Forecasting with Large Language Models

LoFT-LLM: 利用大语言模型进行低频时间序列预测

Jiacheng You, Jingcheng Yang, Yuhang Xie, Zhongxuan Wu, Xiucheng Li, Feng Li, Pengjie Wang, Jian Xu, Bo Zheng, Xinyang Chen

机构 * School of Computer Science and Technology, Harbin Institute of Technology (Shenzhen)(计算机科学与技术学院,哈尔滨工业大学(深圳)) The Chinese University of Hong Kong(香港中文大学)

专题命中 领域大模型 :LLM(title,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.LG

AI总结 LoFT-LLM通过整合大语言模型与低频学习,提升时间序列预测的准确性、鲁棒性和可解释性。

Comments This submission is withdrawn due to internal review and compliance considerations

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07954 2026-01-14 cs.CL 90%

A Human-Centric Pipeline for Aligning Large Language Models with Chinese Medical Ethics

面向中文医疗伦理的以人为本的大型语言模型对齐流水线

Haoan Jin, Han Ying, Jiacheng Ji, Hanhui Xu, Mengyue Wu

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);preference optimization(abstract)

AI总结 本文提出MedES基准和 guardian-in-the-loop 框架,通过监督微调和偏好优化,实现中文医疗伦理场景下的LLM对齐,提升伦理任务表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08148 2026-01-14 cs.IR cs.AI cs.LG 90%

Enriching Semantic Profiles into Knowledge Graph for Recommender Systems Using Large Language Models

利用大语言模型丰富知识图谱中的语义轮廓以提升推荐系统

Seokho Ahn, Sungbok Shin, Young-Duk Seo

机构 * Inha University(成均馆大学) Sogang University(成均馆大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本文提出SPiKE模型,利用大语言模型生成知识图谱中的语义轮廓,通过轮廓聚合和偏好匹配提升推荐系统性能。

Comments Accepted at KDD 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06780 2026-01-14 cs.CL cs.AI 86%

Foundations of LLM Knowledge Materialization: Termination, Reproducibility, Robustness

LLM知识材料化的基础:终止性、可重现性与鲁棒性

Luca Giordano, Simon Razniewski

机构 * ScaDS.AI Dresden/Leipzig & TU Dresden, Germany(ScaDS.AI 德累斯顿/莱比锡及德累斯顿技术大学,德国)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文研究了LLM知识材料化的终止性、可重现性和鲁棒性,通过实验揭示了不同因素对知识提取效果的影响。

Comments Accepted and published in Findings of EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07935 2026-01-14 cs.LG cs.AI cs.CL 85%

Towards Specialized Generalists: A Multi-Task MoE-LoRA Framework for Domain-Specific LLM Adaptation

迈向专业化的通用者:一种多任务MoE-LoRA框架用于领域特定LLM适应

Yuxin Yang, Aoxiong Zeng, Xiangquan Yang

机构 * Shanghai University(上海大学) East China Normal University(东华大学)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出Med-MoE-LoRA框架,通过结合MoE与LoRA实现高效多任务领域适应,尤其适用于医学场景,有效解决领域知识获取与通用能力保留的难题。

Comments Work in Progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16041 2026-01-14 cs.LG cs.AI 81%

Explainable Molecular Property Prediction: Aligning Chemical Concepts with Predictions via Language Models

可解释性分子性质预测:通过语言模型对齐化学概念与预测

Zhenzhong Wang, Zehui Lin, Wanyu Lin, Ming Yang, Minggang Zeng, Kay Chen Tan

机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学) Department of Data Science and Artificial Intelligence, The Hong Kong Polytechnic University(数据科学与人工智能系,香港理工大学) Department of Computing and Department of Data Science and Artificial Intelligence, The Hong Kong Polytechnic University(计算系和数据科学与人工智能系,香港理工大学) Department of Applied Physics, The Hong Kong Polytechnic University(应用物理系,香港理工大学) Institute of High Performance Computing, Agency for Science, Technology and Research (A*STAR)(高性能计算研究所,科技研究局(A*STAR))

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI、cs.LG

AI总结 Lamole通过结合自注意力权重和梯度,提供与化学概念对齐的解释,提升分子性质预测的可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20993 2026-01-14 cs.CL cs.AI cs.HC 79%

SAC: A Framework for Measuring and Inducing Personality Traits in LLMs with Dynamic Intensity Control

SAC:一种用于衡量和诱导LLM人格特质的框架,具有动态强度控制

Adithya Chittem, Aishna Shrivastava, Sai Tarun Pendela, Jagat Sesh Challa, Dhruv Kumar

机构 * BITS Pilani, India(印度比特斯理工学院)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出SAC框架,通过动态强度控制实现LLM人格特质的衡量与诱导,提升人机交互的可控性和细腻度。

Comments Accepted into 18th Edition of International Conference on Agents and Artificial Intelligence (ICAART)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01967 2026-01-14 cs.IT math.IT 78%

MUSE-FM: Multi-task Environment-aware Foundation Model for Wireless Communications

MUSE-FM:多任务环境感知基础模型用于无线通信

Tianyue Zheng, Jiajia Guo, Linglong Dai, Shi Jin, Jun Zhang

专题命中 领域大模型 :foundation model(title,abstract)

AI总结 MUSE-FM通过统一架构和环境感知机制,提升无线通信多任务处理的统一性和适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08441 2026-01-14 cs.AI 77%

YaPO: Learnable Sparse Activation Steering Vectors for Domain Adaptation

YaPO: 用于领域适应的可学习稀疏激活引导向量

Abdelaziz Bounhar, Rania Hossam Elmohamady Elbadry, Hadi Abdine, Preslav Nakov, Michalis Vazirgiannis, Guokan Shang

机构 * MBZUAI Ecole Polytechnique(巴黎政治学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.AI

AI总结 YaPO通过学习稀疏激活引导向量实现LLM的高效、稳定和细粒度对齐,适用于领域适应和文化对齐等任务。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19804 2026-01-14 cs.CL 70%

Compliance-to-Code: Enhancing Financial Compliance Checking via Code Generation

合规至代码:通过代码生成增强财务合规检查

Siyuan Li, Jian Chen, Rui Yao, Xuming Hu, Peilin Zhou, Weihua Qiu, Simin Zhang, Chucheng Dong, Zhiyao Li, Qipeng Xie, Zixuan Yuan

机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Sun Yat-Sen University(孙中山大学) University of California, Riverside(加州大学河滨分校)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 Compliance-to-Code通过构建大规模中文金融监管数据集,提升金融合规检查的自动化水平,采用代码生成技术实现法规结构化和合规逻辑验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15727 2026-01-14 cs.CL 70%

VocalBench: Benchmarking the Vocal Conversational Abilities for Speech Interaction Models

VocalBench: 语音交互模型的语音对话能力基准测试

Heyang Liu, Yuhao Wang, Ziyang Cheng, Hongcheng Liu, Yiqi Li, Yixuan Hou, Ronghua Wu, Qunshan Gu, Yanfeng Wang, Yu Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Ant Group(蚂蚁集团)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 VocalBench通过24000个精心挑选的语音实例,评估语音交互模型在语义质量、语音性能、对话能力和鲁棒性方面的表现,揭示当前模型的共同挑战并推动下一代语音交互系统的发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08040 2026-01-14 cs.CV 67%

Rescind: Countering Image Misconduct in Biomedical Publications with Vision-Language and State-Space Modeling

Rescind: 通过视觉-语言和状态空间建模对抗生物医学出版物中的图像不当行为

Soumyaroop Nandi, Prem Natarajan

专题命中 领域大模型 :language model(abstract);prompting(abstract)

AI总结 Rescind通过视觉-语言和状态空间建模方法,提出首个生物医学图像伪造生成与检测框架,实现高精度伪造定位与验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07850 2026-01-14 cs.MM cs.CV 67%

MLLM-VADStory: Domain Knowledge-Driven Multimodal LLMs for Video Ad Storyline Insights

MLLM-VADStory: 基于领域知识的多模态大语言模型用于视频广告剧情洞察

Jasmine Yang, Poppy Zhang, Shawndra Hill

机构 * Meta

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 MLLM-VADStory通过领域知识引导多模态大语言模型,系统量化和生成视频广告剧情洞察,提升视频广告创意效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07871 2026-01-14 q-bio.QM cs.AI cs.CV cs.LG 62%

Imaging-anchored Multiomics in Cardiovascular Disease: Integrating Cardiac Imaging, Bulk, Single-cell, and Spatial Transcriptomics

心血管疾病中的成像锚定多组学:整合心脏成像、批量、单细胞和空间转录组学

Minh H. N. Le, Tuan Vinh, Thanh-Huy Nguyen, Tao Li, Bao Quang Gia Le, Han H. Huynh, Monika Raj, Carl Yang, Min Xu, Nguyen Quoc Khanh Le

机构 * International Ph.D. Program in Medicine, College of Medicine, Taipei Medical University, Taipei, Taiwan AIBioMed Research Group, Taipei Medical University, Taipei, Taiwan Medical Sciences Division, University of Oxford, Oxford, United Kingdom Computational Biology Department, School of Computer Science, Carnegie Mellon University, Pittsburgh, PA, USA Department of Computer Science, Emory University, Atlanta, GA, USA Department of Chemistry, Emory University, Atlanta, GA, USA International Master Program for Translational Science, College of Medical Science Technology, Taipei Medical University, Taipei 110, Taiwan In-Service Master Program in Artificial Intelligence in Medicine, College of Medicine, Taipei Medical University, Taipei, Taiwan Translational Imaging Research Center, Taipei Medical University Hospital, Taipei, Taiwan

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出通过整合心脏成像与多组学数据,推动心血管疾病研究的多模态融合方法,提升疾病诊断和治疗的精准性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08531 2026-01-14 cs.AI 57%

Sketch-Based Facade Renovation With Generative AI: A Streamlined Framework for Bypassing As-Built Modelling in Industrial Adaptive Reuse

基于草图的立面修复与生成式AI:一种简化框架,用于在工业适应性再利用中绕过实际建成建模

Warissara Booranamaitree, Xusheng Du, Yushu Cai, Zhengyang Wang, Ye Zhang, Haoran Xie

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 本文提出了一种结合生成式AI和视觉-语言模型的三阶段框架,通过处理粗糙草图和文本描述,快速生成保留原结构且提高细节质量的立面修复提案,从而绕过传统详细建模流程。

Comments 10 pages, 9 figures, Proceedings of CAADRIA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08078 2026-01-14 cs.CV cs.CE cs.CL 57%

Exploiting DINOv3-Based Self-Supervised Features for Robust Few-Shot Medical Image Segmentation

利用基于DINOv3的自监督特征进行鲁棒的小样本医学图像分割

Guoping Xu, Jayaram K. Udupa, Weiguo Lu, You Zhang

专题命中 领域大模型 :foundation model(abstract);分类 cs.CL

AI总结 本文提出DINO-AugSeg框架,利用DINOv3特征结合小波域增强和上下文融合,提升小样本医学图像分割的鲁棒性。

Comments 36 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.19107 2026-01-14 cs.LG 57%

Brain network science modelling of sparse neural networks enables Transformers and LLMs to perform as fully connected

稀疏神经网络的脑网络科学建模使Transformer和LLMs能够像全连接网络一样运行

Yingtao Zhang, Diego Cerretti, Jialin Zhao, Wenjing Wu, Ziheng Liao, Umberto Michieli, Carlo Vittorio Cannistraci

机构 * Center for Complex Network Intelligence (CCNI)(复杂网络智能研究中心) Dept. of Computer Science & Technology(计算机科学与技术系) School of Biomedical Engineering(生物医学工程学院) University of Padova(帕多瓦大学) Canva Research(Canva研究)

专题命中 领域大模型 :language model(abstract);分类 cs.LG

AI总结 本文提出双分接收场模型,通过脑启发方法优化稀疏神经网络连接性,提升Transformer和LLMs性能。

详情

展开后加载摘要…

URL PDF HTML 收藏