arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12659 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12659 篇

2512.10370 2025-12-12 cs.AI 79%

LLM-Empowered Representation Learning for Emerging Item Recommendation

基于大语言模型的表示学习用于新兴物品推荐

Ziying Zhang, Quanming Yao, Yaqing Wang

机构 * Department of Electronic Engineering, Tsinghua University(清华大学电子工程系) Beijing Institute of Mathematical Sciences and Applications(北京数学科学研究院)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

AI总结 本文提出EmerFlow框架,利用大语言模型生成独特嵌入,通过丰富特征、对齐空间和元学习优化,提升新兴物品推荐性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09483 2025-12-11 cs.CL cs.CY 79%

Source Coverage and Citation Bias in LLM-based vs. Traditional Search Engines

基于大语言模型的搜索引擎与传统搜索引擎的来源覆盖与引用偏见

Peixian Zhang, Qiming Ye, Zifan Peng, Kiran Garimella, Gareth Tyson

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Rutgers University(罗格斯大学) Rutgers University New Brunswick United States(罗格斯大学新 Brunswick美国)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.CL

AI总结 本文研究了基于大语言模型的搜索引擎与传统搜索引擎在来源覆盖和引用偏见方面的差异,发现LLM-SEs在资源多样性上优于传统搜索引擎,但其可信度和中立性仍需进一步提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05993 2025-12-09 cs.CV cs.AI 79%

Domain-Specific Foundation Model Improves AI-Based Analysis of Neuropathology

领域特定基础模型提升基于人工智能的神经病理科分析

Ruchika Verma, Shrishtee Kandoi, Robina Afzal, Shengjia Chen, Jannes Jegminat, Michael W. Karlovich, Melissa Umphlett, Timothy E. Richardson, Kevin Clare, Quazi Hossain, Jorge Samanamud, Phyllis L. Faust, Elan D. Louis, Ann C. McKee, Thor D. Stein, Jonathan D. Cherry, Jesse Mez, Anya C. McGoldrick, Dalilah D. Quintana Mora, Melissa J. Nirenberg, Ruth H. Walker, Yolfrankcis Mendez, Susan Morgello, Dennis W. Dickson, Melissa E. Murray, Carlos Cordon-Cardo, Nadejda M. Tsankova, Jamie M. Walker, Diana K. Dangoor, Stephanie McQuillan, Emma L. Thorn, Claudia De Sanctis, Shuying Li, Thomas J. Fuchs, Kurt Farrell, John F. Crary, Gabriele Campanella

机构 * Windreich Department of AI and Human Health(AI与人类健康部门) Icahn School of Medicine at Mount Sinai(辛克医学院(梅奥医院)) Hasso Plattner Institute for Digital Health at Mount Sinai(梅奥医院数字健康研究所) Department of Neurology(神经病学部门) Department of Pathology(病理学部门) Columbia University(哥伦比亚大学) Memorial Sloan-Kettering Cancer Center(纪念斯隆凯特林癌症中心) Department of Neuroscience(神经科学部门) Mayo Clinic College of Medicine(梅奥诊所医学院) University of Texas Southwestern Medical Center(德克萨斯西南医学中心) Peter O’Donnell Jr. Brain Institute(彼得·奥多尼尔 Jr. 脑研究所) VA Boston Healthcare System(波士顿退伍军人医疗系统)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

AI总结 NeuroFM通过专门训练于脑组织的领域特定基础模型,提升神经病理学AI分析的准确性与可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03750 2025-12-04 cs.LG cond-mat.mtrl-sci 79%

Universally Converging Representations of Matter Across Scientific Foundation Models

跨科学基础模型中物质的普遍收敛表示

Sathya Edamadaka, Soojung Yang, Ju Li, Rafael Gómez-Bombarelli

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

AI总结 研究揭示科学基础模型在不同模态和数据集上对物质的普遍表示收敛性,表明模型学习了共同的物理现实表示,但受限于训练数据和归纳偏置。

Comments Oral spotlight at NeurIPS 2025 UniReps Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03068 2025-12-04 cs.CY cs.AI 79%

Echoes of AI Harms: A Human-LLM Synergistic Framework for Bias-Driven Harm Anticipation

AI危害的回响:一种人机协同框架用于偏见驱动的危害预见

Nicoleta Tantalaki, Sophia Vei, Athena Vakali

机构 * Aristotle University(亚里士多德大学)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

AI总结 本文提出ECHO框架,通过系统映射人工智能偏见类型到危害结果,实现主动危害预见,用于高风险领域中的偏见驱动危害检测与治理。

Comments 38 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01434 2025-12-02 cs.AI 79%

A Flexible Multi-Agent LLM-Human Framework for Fast Human Validated Tool Building

一种灵活的多智能体LLM-人类框架,用于快速的人类验证工具构建

Daull Xavier, Patrice Bellot, Emmanuel Bruno, Vincent Martin, Elisabeth Murisasco

机构 * Toulon Univ(图卢兹大学) Aix Marseille Univ(阿维尼翁-马赛大学) CNRS(国家科学研究中心) LIS(信息系统实验室)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

AI总结 该研究提出了一种灵活的多智能体框架,通过人类反馈和强化学习,实现快速的人类验证工具构建,适用于复杂迭代任务。

Journal ref 2025 IEEE/WIC International Conference on Web Intelligence and Intelligent Agent Technology (WI-IAT), Nov 2025, Londres, United Kingdom

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09019 2025-12-01 cs.CL 79%

MedMobile: A mobile-sized language model with clinical capabilities

MedMobile: 一种具有临床能力的移动尺寸语言模型

Krithik Vishwanath, Jaden Stryker, Anton Alyakin, Daniel Alexander Alber, Eric Karl Oermann

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL

AI总结 MedMobile是一种移动尺寸的语言模型,通过高效的方法在医疗领域实现高性能表现,成为参数最少且性能最佳的临床应用模型。

Comments 24 pages, 7 figures (4 main, 3 supplementary)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19220 2025-12-01 cs.CV cs.AI 79%

Are Large Vision Language Models Truly Grounded in Medical Images? Evidence from Italian Clinical Visual Question Answering

大视觉语言模型真的在医学图像上具有基础性吗?来自意大利临床视觉问答的证据

Federico Felizzi, Olivia Riccomi, Michele Ferramola, Francesco Andrea Causio, Manuel Del Medico, Vittorio De Vita, Lorenzo De Mori, Alessandra Piscitelli, Pietro Eric Risuleo, Bianca Destro Castaniti, Antonio Cristiano, Alessia Longo, Luigi De Angelis, Mariapia Vassalli, Marcello Di Pumpo

机构 * SIIAM NSBProject Dept. of Life Sciences & Public Health, UCSC(生命科学与公共卫生系,UCSC) ASL RM 4 UCSC Univ. Paris Cité(巴黎Cité大学) Univ. of Pisa(比萨大学)

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI

AI总结 研究通过测试四种先进模型在意大利医学问题上的表现,揭示了大视觉语言模型在视觉基础上的差异,强调了临床部署前的严格评估需求。

Comments Accepted at the Workshop on Multimodal Representation Learning for Healthcare (MMRL4H), EurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20100 2025-11-26 cs.DC cs.CL 79%

QiMeng-Kernel: Macro-Thinking Micro-Coding Paradigm for LLM-Based High-Performance GPU Kernel Generation

QiMeng-Kernel:基于宏思维微编码的LLM高性能GPU内核生成范式

Xinguo Zhu, Shaohui Peng, Jiaming Guo, Yunji Chen, Qi Guo, Yuanbo Wen, Hang Qin, Ruizhi Chen, Qirui Zhou, Ke Gao, Yanjun Wu, Chen Zhao, Ling Li

专题命中 领域大模型 :LLM(title,abstract);分类 cs.CL

AI总结 QiMeng-Kernel通过宏思维微编码范式,结合强化学习和通用LLM,实现高性能GPU内核生成,准确率和效率均优于现有方法。

Comments 9 pages, 2 figures, accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19147 2025-11-25 cs.CV cs.LG 79%

Collaborative Learning with Multiple Foundation Models for Source-Free Domain Adaptation

基于多基础模型的协同学习用于无源域适应

Huisoo Lee, Jisu Han, Hyunsouk Cho, Wonjun Hwang

机构 * Ajou University(阿乔大学) Korea University(韩国大学)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

AI总结 本文提出CoMA框架,通过协同利用两种互补基础模型提升无源域适应性能,实验显示在多个基准测试中优于现有方法。

Comments 15 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18670 2025-11-25 cs.AI 79%

MoveGPT: Scaling Mobility Foundation Models with Spatially-Aware Mixture of Experts

MoveGPT: 通过空间感知专家混合模型扩展移动基础模型

Chonghua Han, Yuan Yuan, Jingtao Ding, Jie Feng, Fanjin Meng, Yong Li

机构 * Department of Electronic Engineering, Tsinghua University(电子工程系,清华大学)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

AI总结 MoveGPT通过空间感知专家混合模型实现大规模移动基础模型的扩展,提升多种下游任务性能,展现强泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00221 2025-11-25 cs.LG eess.AS 79%

Speech Foundation Models Generalize to Time Series Tasks from Wearable Sensor Data

语音基础模型能泛化到可穿戴传感器数据的时间序列任务

Jaya Narain, Zakaria Aldeneh, Shirley Ren

机构 * Apple(苹果公司)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

AI总结 语音基础模型通过泛化能力在可穿戴传感器的时间序列任务中取得突破性性能。

Comments Preprint, under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09095 2025-11-19 eess.IV cs.AI cs.CV 79%

Foundation Models in Medical Imaging: A Review and Outlook

Vivien van Veldhuizen, Vanessa Botha, Chunyao Lu, Melis Erdal Cesur, Kevin Groot Lipman, Edwin D. de Jong, Hugo Horlings, Clárisa I. Sanchez, Cees G. M. Snoek, Lodewyk Wessels, Ritse Mann, Eric Marcus, Jonas Teuwen

机构 * Netherlands Cancer Institute(荷兰癌症研究所) Radboud University Medical Center(拉德堡德大学医学中心) Delft University of Technology(代尔夫特理工大学) University of Amsterdam(阿姆斯特丹大学) Oncode Institute(Oncode研究所)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11752 2025-11-18 cs.AI cs.DL quant-ph 79%

Towards autonomous quantum physics research using LLM agents with access to intelligent tools

Sören Arlt, Xuemei Gu, Mario Krenn

机构 * Machine Learning in Science Cluster, Department of Computer Science, Faculty of Science, University of Tuebingen, Germany(图宾根大学科学与技术学院计算机科学系机器学习科学集群) Max Planck Institute for the Science of Light, Erlangen, Germany(马克斯·普朗克光科学研究所)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

Comments 24 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15980 2025-11-17 cs.HC cs.AI cs.DB 79%

Text-to-SQL Domain Adaptation via Human-LLM Collaborative Data Annotation

Yuan Tian, Daniel Lee, Fei Wu, Tung Mai, Kun Qian, Siddhartha Sahai, Tianyi Zhang, Yunyao Li

机构 * Purdue University(普渡大学) Adobe Inc.(Adobe公司)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

Comments Accepted by IUI'25 Code & Demo: https://github.com/magic-YuanTian/SQLsynth

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09742 2025-11-14 cs.CV cs.AI 79%

Feature Quality and Adaptability of Medical Foundation Models: A Comparative Evaluation for Radiographic Classification and Segmentation

Frank Li, Theo Dapamede, Mohammadreza Chavoshi, Young Seok Jeon, Bardia Khosravi, Abdulhameed Dere, Beatrice Brown-Mulry, Rohan Satya Isaac, Aawez Mansuri, Chiratidzo Sanyika, Janice Newsome, Saptarshi Purkayastha, Imon Banerjee, Hari Trivedi, Judy Gichoya

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

Comments 7 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08109 2025-11-12 cs.CL 79%

Estranged Predictions: Measuring Semantic Category Disruption with Masked Language Modelling

Yuxuan Liu, Haim Dubossarsky, Ruth Ahnert

机构 * School of Arts, Queen Mary University of London(艺术学院,伦敦女王学院) School of Electronic Engineering and Computer Science, Queen Mary University of London(电子工程与计算机科学学院,伦敦女王学院)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09667 2025-11-11 cs.CL 79%

BLADE: Benchmarking Language Model Agents for Data-Driven Science

Ken Gu, Ruoxi Shang, Ruien Jiang, Keying Kuang, Richard-John Lin, Donghe Lyu, Yue Mao, Youran Pan, Teng Wu, Jiaqian Yu, Yikun Zhang, Tianmai M. Zhang, Lanyi Zhu, Mike A. Merrill, Jeffrey Heer, Tim Althoff

机构 * University of Washington(华盛顿大学) UC Berkeley(伯克利大学) New York University(纽约大学) Stanford University(斯坦福大学) University of British Columbia(不列颠哥伦比亚大学) Microsoft(微软) George Washington University(乔治·华盛顿大学)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL

Comments EMNLP 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06222 2025-11-11 cs.CL cs.CY 79%

SPA: Achieving Consensus in LLM Alignment via Self-Priority Optimization

Yue Huang, Xiangqi Wang, Xiangliang Zhang

机构 * University of Notre Dame(诺特大学)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.CL

Comments Accepted by AAAI 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05630 2025-11-11 q-bio.NC cs.AI 79%

BrainCSD: A Hierarchical Consistency-Driven MoE Foundation Model for Unified Connectome Synthesis and Multitask Brain Trait Prediction

Xiongri Shen, Jiaqi Wang, Yi Zhong, Zhenxi Song, Leilei Zhao, Liling Li, Yichen Wei, Lingyan Liang, Shuqiang Wang, Baiying Lei, Demao Deng, Zhiguo Zhang

机构 * Department of Computer Science and Technology, Harbin Institute of Technology(哈尔滨工业大学计算机科学与技术系) School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology(哈尔滨工业大学智能科学与工程学院) School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical, Measurements and Ultrasound Imaging, Shenzhen University Medical School, Shenzhen University(深圳大学医学院生物医学工程学院) Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05494 2025-11-11 cs.IR cs.AI 79%

Customized Retrieval-Augmented Generation with LLM for Debiasing Recommendation Unlearning

Haichao Zhang, Chong Zhang, Peiyu Hu, Shi Qiu, Jia Wang

机构 * Xi'an Jiaotong-Liverpool University(西安交通大学利物浦大学)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

Comments 10 pages, 4 figures. Accepted ICDM 2025 (IEEE International Conference on Data Mining)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.16995 2025-11-11 q-bio.QM cs.AI 79%

tcrLM: a lightweight protein language model for predicting T cell receptor and epitope binding specificity

Xing Fang, Chenpeng Yu, Shiye Tian, Hui Liu

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02451 2025-11-05 cs.CL 79%

Merging Continual Pretraining Models for Domain-Specialized LLMs: A Case Study in Finance

Kentaro Ueda, François Portet, Hirohiko Suwa, Keiichi Yasumoto

机构 * NARA Institute of Science and Technology(日本科学技术研究院)

专题命中 领域大模型 :pretraining(title);SFT(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26715 2025-10-31 cs.LG 79%

LSM-MS2: A Foundation Model Bridging Spectral Identification and Biological Interpretation

Gabriel Asher, Devesh Shah, Amy A. Caudy, Luke Ferro, Lea Amar, Ana S. H. Costa, Thomas Patton, Niall O'Connor, Jennifer M. Campbell, Jack Geremia

机构 * Matterworks, Inc.(Matterworks公司)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26217 2025-10-31 q-fin.CP cs.AI math.OC 79%

Hybrid LLM and Higher-Order Quantum Approximate Optimization for CSA Collateral Management

Tao Jin, Stuart Florescu, Heyu, Jin

机构 * Pyligent AI Dept. of Computer & Mathematical Sciences, Caltech(计算机与数学科学系,加州理工学院) Department of Economics, UCLA(经济学系,加州大学洛杉矶分校)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.AI

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22481 2025-10-30 eess.IV cs.AI cs.CV cs.MM 79%

Towards Blind Bitstream-corrupted Video Recovery via a Visual Foundation Model-driven Framework

Tianyi Liu, Kejun Wu, Chen Cai, Yi Wang, Kim-Hui Yap, Lap-Pui Chau

机构 * School of EEE, Nanyang Technological University(南洋理工大学电子工程系) School of EIC, Huazhong University of Science and Technology(华中科技大学电子信息学院) Dept. of EEE, The Hong Kong Polytechnic University(香港理工大学电子工程系)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

Comments 10 pages, 5 figures, accepted by ACMMM 2025

Journal ref Proceedings of the 33rd ACM International Conference on Multimedia, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20967 2025-10-27 cs.CV cs.AI 79%

3DReasonKnee: Advancing Grounded Reasoning in Medical Vision Language Models

Sraavya Sambara, Sung Eun Kim, Xiaoman Zhang, Luyang Luo, Shreya Johri, Mohammed Baharoon, Du Hyun Ro, Pranav Rajpurkar

机构 * Department of Biomedical Informatics, Harvard Medical School(生物医学信息学系,哈佛医学院) Seoul National University Hospital(首尔国立大学医院)

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17684 2025-10-21 cs.CV cs.AI 79%

Intelligent Communication Mixture-of-Experts Boosted-Medical Image Segmentation Foundation Model

Xinwei Zhang, Hu Chen, Zhe Yuan, Sukun Tian, Peng Feng

机构 * College of Optoelectronics Engineering, Chongqing University, Chongqing, China Center of Digital Dentistry, Peking University School Hospital of Stomatology \& NHC Key Laboratory of Digital Stomatology, Beijing, China

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14807 2025-10-21 eess.IV cs.AI cs.CV 79%

FetalCLIP: A Visual-Language Foundation Model for Fetal Ultrasound Image Analysis

Fadillah Maani, Numan Saeed, Tausifa Saleem, Zaid Farooq, Hussain Alasmawi, Werner Diehl, Ameera Mohammad, Gareth Waring, Saudabi Valappi, Leanne Bricker, Mohammad Yaqub

机构 * Department of Computer Vision(计算机视觉系) Mohamed bin Zayed University of Artificial Intelligence(马尔代夫比兹人工智能大学) Department of Machine Learning(机器学习系) Corniche Hospital, Abu Dhabi Health Services Company (SEHA)(阿布扎赫尔医院,阿布扎赫健康服务公司(SEHA))

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08821 2025-10-21 cs.IR cs.AI 79%

EasyRec: Simple yet Effective Language Models for Recommendation

Xubin Ren, Chao Huang

机构 * The University of Hong Kong(香港大学)

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI

Comments Published as an EMNLP'25 main paper

详情

展开后加载摘要…

URL PDF HTML 收藏