arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12706 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12706 篇

2512.24754 2026-01-01 astro-ph.IM cs.AI 74%

AstroReview: An LLM-driven Multi-Agent Framework for Telescope Proposal Peer Review and Refinement

AstroReview: 一种基于大语言模型的多智能体框架用于望远镜提案同行评审与优化

Yutong Wang, Yunxiang Xiao, Yonglin Tian, Junyong Li, Jing Wang, Yisheng Lv

机构 * The Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统重点实验室,自动化研究所,中国科学院) The School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)

专题命中 领域大模型 :LLM(title);分类 cs.AI

AI总结 AstroReview通过多智能体框架实现望远镜提案的自动化评审与优化,显著提升评审效率和质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20080 2025-12-24 cs.NI cs.AI physics.optics 74%

CBA: Communication-Bound-Aware Cross-Domain Resource Assignment for Pipeline-Parallel Distributed LLM Training in Dynamic Multi-DC Optical Networks

CBA:面向动态多数据中心光网络的流水线并行分布式大语言模型训练的通信感知跨域资源分配

Dianxuan Fu, Xiaomin Liu, Yihao Zhang, Shikui Shen, Weisheng Hu, Qunbi Zhuge

专题命中 领域大模型 :LLM(title);分类 cs.AI

AI总结 本文提出CBA框架,通过通信感知的跨域资源分配,提升多数据中心光网络中流水线并行分布式大语言模型训练的效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14933 2025-12-12 cs.CE cs.AI cs.CR 74%

Explain First, Trust Later: LLM-Augmented Explanations for Graph-Based Crypto Anomaly Detection

先解释,后信任:基于图的加密异常检测中的LLM增强解释

Adriana Watson, Grant Richards, Daniel Schiff

机构 * School of Engineering Technology Purdue University(工程科技学院 Purdue 大学) College of Liberal Arts Purdue University(文理学院 Purdue 大学)

专题命中 领域大模型 :LLM(title);分类 cs.AI

AI总结 本文提出利用LLM增强的图基解释方法,用于加密货币异常检测,以提高自动化检测工具的有效性和准确性。

Comments 11 pages, 5 figures. Code available at: https://github.com/awatson246/crypto-anomaly-detection-policy

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06105 2025-12-09 cs.CV cs.AI 74%

Explainable Melanoma Diagnosis with Contrastive Learning and LLM-based Report Generation

基于对比学习和大语言模型的可解释性黑色素瘤诊断

Junwen Zheng, Xinran Xu, Li Rong Wang, Chang Cai, Lucinda Siyun Tan, Dingyuan Wang, Hong Liang Tey, Xiuyi Fan

专题命中 领域大模型 :LLM(title);分类 cs.AI

AI总结 本文提出基于对比学习和大语言模型的可解释性黑色素瘤诊断框架,通过将临床标准映射到视觉Transformer空间,实现图像与临床解释的透明连接,提升模型可解释性与临床信任度。

Comments AAAI-26-AIA

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05858 2025-12-08 cs.CL 74%

Prompting Science Report 4: Playing Pretend: Expert Personas Don't Improve Factual Accuracy

提示科学报告4:扮演角色:专家角色不提高事实准确性

Savir Basil, Ina Shapiro, Dan Shapiro, Ethan Mollick, Lilach Mollick, Lennart Meincke

专题命中 领域大模型 :prompting(title);分类 cs.CL

AI总结 该研究探讨了专家角色和低知识角色对AI模型在客观多项选择题上的准确性影响,发现角色提示通常未提升表现,专家角色在多数情况下无益,低知识角色反而降低准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04716 2025-12-05 physics.flu-dyn cs.AI 74%

Towards an AI Fluid Scientist: LLM-Powered Scientific Discovery in Experimental Fluid Mechanics

迈向人工智能流体科学家:基于LLM的实验流体力学中的科学发现

Haodong Feng, Lugang Ye, Dixia Fan

机构 * Westlake University(西湖大学)

专题命中 领域大模型 :LLM(title);分类 cs.AI

AI总结 本文提出基于LLM的AI流体科学家框架,实现自主实验流程,提升流体力学研究效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08135 2025-11-27 cs.SE cs.AI cs.DC cs.PF 74%

Leveraging AI for Productive and Trustworthy HPC Software: Challenges and Research Directions

利用AI实现高效且可信的HPC软件:挑战与研究方向

Keita Teranishi, Harshitha Menon, William F. Godoy, Prasanna Balaprakash, David Bau, Tal Ben-Nun, Abhinav Bhatele, Franz Franchetti, Michael Franusich, Todd Gamblin, Giorgis Georgakoudis, Tom Goldstein, Arjun Guha, Steven Hahn, Costin Iancu, Zheming Jin, Terry Jones, Tze Meng Low, Het Mankad, Narasinga Rao Miniskar, Mohammad Alaul Haque Monil, Daniel Nichols, Konstantinos Parasyris, Swaroop Pophale, Pedro Valero-Lara, Jeffrey S. Vetter, Samuel Williams, Aaron Young

机构 * Oak Ridge National Laboratory(奥克荷厄斯国家实验室) Lawrence Livermore National Laboratory(劳伦斯利弗莫尔国家实验室) Lawrence Berkeley National Laboratory(劳伦斯伯克利国家实验室) Carnegie Mellon University(卡内基梅隆大学) Northeastern University(东北大学) University of Maryland(马里兰大学) SpiralGen Inc.(SpiralGen公司)

专题命中 领域大模型 :large language model(abstract,comments);language model(abstract,comments);分类 cs.AI

AI总结 本文探讨了利用AI改进HPC软件的挑战与研究方向,提出通过Ellora和Durban项目推动AI在HPC软件发展中的应用。

Comments 12 pages, 1 Figure, Accepted at "The 1st International Workshop on Foundational Large Language Models Advances for HPC" LLM4HPC to be held in conjunction with ISC High Performance 2025

Journal ref In: Neuwirth, S., Paul, A.K., Weinzierl, T., Carson, E.C. (eds) High Performance Computing. ISC High Performance 2025. Lecture Notes in Computer Science, vol 16091. Springer, Cham

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19987 2025-11-26 cs.CL cs.IR 74%

$\text{R}^2\text{R}$: A Route-to-Rerank Post-Training Framework for Multi-Domain Decoder-Only Rerankers

R²R:多领域解码器-only重排序框架的路由到重排序后训练方法

Xinyu Wang, Hanwei Wu, Qingchen Hu, Zhenghan Tai, Jingrui Tian, Lei Ding, Jijun Chi, Hailin He, Tung Sum Thomas Kwok, Yufei Cui, Sicheng Lyu, Muzhi Li, Mingze Li, Xinyue Yu, Ling Zhou, Peng Lu

机构 * McGill University(麦吉尔大学) University of Toronto(多伦多大学) University of Manitoba(曼尼托巴大学) Mila CUHK(中国科技大学) Université de Montréal(蒙特利尔大学) CG Matrix(CG矩阵)

专题命中 领域大模型 :post-training(title);分类 cs.CL

AI总结 R²R通过动态专家路由和两阶段训练策略,提升多领域解码器重排序器的领域适应能力与跨领域鲁棒性。

Comments 13 pages, including 3 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02615 2025-11-10 cs.LG 74%

ExGra-Med: Extended Context Graph Alignment for Medical Vision-Language Models

Duy M. H. Nguyen, Nghiem T. Diep, Trung Q. Nguyen, Hoang-Bao Le, Tai Nguyen, Tien Nguyen, TrungTin Nguyen, Nhat Ho, Pengtao Xie, Roger Wattenhofer, James Zou, Daniel Sonntag, Mathias Niepert

机构 * German Research Centre for Artificial Intelligence (DFKI)(德国人工智能研究中心) Max Planck Research School for Intelligent Systems (IMPRS-IS)(马克斯·普朗克智能系统研究学校) University of Stuttgart(斯图加特大学) University Medical Center Gottingen(哥廷根大学医学中心) Max Planck Institute for Multidisciplinary Sciences(马克斯·普朗克多学科科学研究所) ARC Centre of Excellence for the Mathematical Analysis of Cellular Systems(细胞系统数学分析卓越中心) School of Mathematical Sciences, Queensland University of Technology(昆士兰科技大学数学科学学院) University of Oldenburg(奥尔登堡大学) University of Texas at Austin(德克萨斯大学奥斯汀分校) University of California San Diego(加州大学圣地亚哥分校) MBZUAI(马克斯·普朗克人工智能研究所) ETH Zurich(苏黎世联邦理工学院) Stanford University(斯坦福大学)

专题命中 领域大模型 :language model(title);分类 cs.LG

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27552 2025-11-03 cs.CL 74%

Multilingual BERT language model for medical tasks: Evaluation on domain-specific adaptation and cross-linguality

Yinghao Luo, Lang Zhou, Amrish Jhingoer, Klaske Vliegenthart Jongbloed, Carlijn Jordans, Ben Werkhoven, Tom Seinen, Erik van Mulligen, Casper Rokx, Yunlei Li

专题命中 领域大模型 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12958 2025-10-16 astro-ph.IM astro-ph.HE astro-ph.SR cs.LG 74%

Simulation-Based Pretraining and Domain Adaptation for Astronomical Time Series with Minimal Labeled Data

Rithwik Gupta, Daniel Muthukrishna, Jeroen Audenaert

机构 * Massachusetts Institute of Technology, Cambridge, MA 02139, USA Irvington High School, Fremont, CA 94538, USA Center for Astrophysics, Harvard \& Smithsonian, Cambridge, MA 02138, USA

专题命中 领域大模型 :pretraining(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12490 2025-10-15 cs.AI 74%

Using Medical Algorithms for Task-Oriented Dialogue in LLM-Based Medical Interviews

Rui Reis, Pedro Rangel Henriques, João Ferreira-Coimbra, Eva Oliveira, Nuno F. Rodrigues

机构 * University of Minho(米尼霍大学) University Hospital Center of São João(圣约翰大学医院中心) Ai - School of Technology, IPCA(2Ai-技术学校,IPCA)

专题命中 领域大模型 :LLM(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14738 2025-10-02 cs.AI 74%

R&D-Agent: An LLM-Agent Framework Towards Autonomous Data Science

Xu Yang, Xiao Yang, Shikai Fang, Yifei Zhang, Jian Wang, Bowen Xian, Qizheng Li, Jingyuan Li, Minrui Xu, Yuante Li, Haoran Pan, Yuge Zhang, Weiqing Liu, Yelong Shen, Weizhu Chen, Jiang Bian

专题命中 领域大模型 :LLM(title);分类 cs.AI

Comments 33 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18520 2025-09-24 cs.CR cs.AI 74%

Coherence-driven inference for cybersecurity

Steve Huntsman

专题命中 领域大模型 :large language model(abstract,comments);language model(abstract,comments);分类 cs.AI

Comments LLM4Sec - Workshop on the use of Large Language Models for Cybersecurity (https://llm4sec-workshop.github.io/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13846 2025-09-18 cs.CV cs.LG 74%

Consistent View Alignment Improves Foundation Models for 3D Medical Image Segmentation

Puru Vaish, Felix Meister, Tobias Heimann, Christoph Brune, Jelmer M. Wolterink

机构 * Department of Applied Mathematics, Technical Medical Centre, University of Twente(代尔夫特理工大学应用数学系) Digital Technology and Innovation, Siemens Healthineers, Erlangen, Germany(西门子医疗创新部,埃尔朗根,德国)

专题命中 领域大模型 :foundation model(title);分类 cs.LG

Comments MICCAI 2025: 1st Place in Transformer track and 2nd Place in Convolution track of SSL3D-OpenMind challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11944 2025-09-16 cs.AI 74%

Agentic Temporal Graph of Reasoning with Multimodal Language Models: A Potential AI Aid to Healthcare

Susanta Mitra

专题命中 领域大模型 :language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21354 2025-07-30 cs.AI cs.MA 74%

Games Agents Play: Towards Transactional Analysis in LLM-based Multi-Agent Systems

Monika Zamojska, Jarosław A. Chudziak

机构 * Faculty of Electronics and Information Technology(电子与信息技术学院)

专题命中 领域大模型 :LLM(title);分类 cs.AI

Comments Proceedings of the Annual Meeting of the Cognitive Science Society (CogSci 2025), https://escholarship.org/uc/item/7gg6j165

Journal ref Proceedings of the Annual Meeting of the Cognitive Science Society 47 (2025) 1598-1605

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03493 2025-07-08 cs.CL 74%

AI-VaxGuide: An Agentic RAG-Based LLM for Vaccination Decisions

Abdellah Zeggai, Ilyes Traikia, Abdelhak Lakehal, Abdennour Boulesnane

机构 * Department of Fundamental Computing and its Applications, Faculty of NTIC(基础计算及其应用系) BIOSTIM Laboratory, Medicine Faculty Salah Boubnider Constantine 03 University(BIOSTIM实验室,医学系Salah Boubnider Constantine 03大学)

专题命中 领域大模型 :LLM(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17601 2025-06-24 cs.RO cs.AI 74%

Risk-Guided Diffusion: Toward Deploying Robot Foundation Models in Space, Where Failure Is Not An Option

Rohan Thakker, Adarsh Patnaik, Vince Kurtz, Jonas Frey, Jonathan Becktor, Sangwoo Moon, Rob Royce, Marcel Kaufmann, Georgios Georgakis, Pascal Roth, Joel Burdick, Marco Hutter, Shehryar Khattak

机构 * NASA-JPL, Caltech(美国国家航空航天局喷气推进实验室,加州理工学院) CME, Caltech(加州理工学院计算机工程系) Robotic Systems Lab, ETH Zurich(苏黎世联邦理工学院机器人系统实验室)

专题命中 领域大模型 :foundation model(title);分类 cs.AI

Journal ref Robotics Science and Systems 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12516 2025-06-17 cond-mat.mtrl-sci cs.LG 74%

Information fusion strategy integrating pre-trained language model and contrastive learning for materials knowledge mining

Yongqian Peng, Zhouran Zhang, Longhui Zhang, Fengyuan Zhao, Yahao Li, Yicong Ye, Shuxin Bai

专题命中 领域大模型 :language model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07591 2025-06-10 cs.AI q-bio.QM 74%

Automating Exploratory Multiomics Research via Language Models

Shang Qu, Ning Ding, Linhai Xie, Yifei Li, Zaoqu Liu, Kaiyan Zhang, Yibai Xiong, Yuxin Zuo, Zhangren Chen, Ermo Hua, Xingtai Lv, Youbang Sun, Yang Li, Dong Li, Fuchu He, Bowen Zhou

专题命中 领域大模型 :language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02149 2025-06-04 eess.IV cs.LG eess.SP 74%

Tomographic Foundation Model -- FORCE: Flow-Oriented Reconstruction Conditioning Engine

Wenjun Xia, Chuang Niu, Ge Wang

机构 * Department of Biomedical Engineering, School of Engineering, Rensselaer Polytechnic Institute(生物医学工程系,工程学院,伦斯勒理工学院)

专题命中 领域大模型 :foundation model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24004 2025-06-02 cs.HC cs.CL cs.CY 74%

Redefining Research Crowdsourcing: Incorporating Human Feedback with LLM-Powered Digital Twins

Amanda Chan, Catherine Di, Joseph Rupertus, Gary Smith, Varun Nagaraj Rao, Manoel Horta Ribeiro, Andrés Monroy-Hernández

机构 * Princeton University(普林斯顿大学)

专题命中 领域大模型 :LLM(title);分类 cs.CL

Comments Accepted as a CHI Late Breaking Work (2025), cite appropriately

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23146 2025-05-30 cs.CL 74%

Cross-Domain Bilingual Lexicon Induction via Pretrained Language Models

Qiuyu Ding, Zhiqiang Cao, Hailong Cao, Tiejun Zhao

专题命中 领域大模型 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15422 2025-05-22 cs.CL 74%

Trends and Challenges in Authorship Analysis: A Review of ML, DL, and LLM Approaches

Nudrat Habib, Tosin Adewumi, Marcus Liwicki, Elisa Barney

专题命中 领域大模型 :LLM(title);分类 cs.CL

Comments 25 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05575 2025-04-09 cs.CV cs.LG 74%

A Lightweight Large Vision-language Model for Multimodal Medical Images

Belal Alsinglawi, Chris McCarthy, Sara Webb, Christopher Fluke, Navid Toosy Saidy

机构 * Swinburne University of Technology(斯威本科技大学) PropelHealthAI

专题命中 领域大模型 :language model(title);分类 cs.LG

Comments 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10671 2025-03-17 cs.CL 74%

Identifying Non-Replicable Social Science Studies with Language Models

Denitsa Saynova, Kajsa Hansson, Bastiaan Bruinsma, Annika Fredén, Moa Johansson

机构 * Chalmers University of Technology(查尔姆斯理工大学) University of Gothenburg(哥德堡大学) Lund University(隆德大学)

专题命中 领域大模型 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06828 2025-03-11 eess.IV cs.AI cs.CV 74%

Towards a Multimodal MRI-Based Foundation Model for Multi-Level Feature Exploration in Segmentation, Molecular Subtyping, and Grading of Glioma

Somayeh Farahani, Marjaneh Hejazi, Antonio Di Ieva, Emad Fatemizadeh, Sidong Liu

专题命中 领域大模型 :foundation model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05497 2025-02-14 cs.CL 74%

DeepThink: Aligning Language Models with Domain-Specific User Intents

Yang Li, Mingxuan Luo, Yeyun Gong, Chen Lin, Jian Jiao, Yi Liu, Kaili Huang

机构 * Xiamen University(厦门大学) Microsoft(微软)

专题命中 领域大模型 :language model(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21304 2025-02-07 cs.CV cs.LG 74%

VideoSAM: A Large Vision Foundation Model for High-Speed Video Segmentation

Chika Maduabuchi, Ericmoore Jossou, Matteo Bucci

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 领域大模型 :foundation model(title);分类 cs.LG

Comments Accepted at IEEE SSD 2025 (CSP Track)

详情

展开后加载摘要…

URL PDF HTML 收藏