arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12705 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12705 篇

2507.12595 2025-07-18 eess.AS 78%

Enhancing In-Domain and Out-Domain EmoFake Detection via Cooperative Multilingual Speech Foundation Models

Orchid Chetia Phukan, Mohd Mujtaba Akhtar, Girish, Arun Balaji Buduru

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12236 2025-07-17 cs.CV 78%

Generate to Ground: Multimodal Text Conditioning Boosts Phrase Grounding in Medical Vision-Language Models

Felix Nützel, Mischa Dombrowski, Bernhard Kainz

机构 * Friedrich-Alexander-Universität Erlangen-Nürnberg(弗赖堡-亚历山大大学埃尔兰根-纽伦堡) Imperial College London(伦敦帝国理工学院)

专题命中 领域大模型 :language model(title,abstract)

Comments 20 pages, 6 figures. To appear in Proc. MIDL 2025 (PMLR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09209 2025-07-15 cs.CV 78%

Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models

Xiao Liang, Di Wang, Zhicheng Jiao, Ronghan Li, Pengfei Yang, Quan Wang, Tat-Seng Chua

机构 * The Key Laboratory of Smart Human-Computer Interaction and Wearable Technology of Shaanxi Province, Xidian University, China(陕西省智能人机交互与可穿戴技术重点实验室,西安电子科技大学) Warren Alpert Medical School, Brown University, USA(布朗大学沃伦·阿尔珀特医学院) National University of Singapore, Singapore(新加坡国立大学)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09097 2025-07-15 cs.CV 78%

RadEyeVideo: Enhancing general-domain Large Vision Language Model for chest X-ray analysis with video representations of eye gaze

Yunsoo Kim, Jinge Wu, Honghan Wu

机构 * Institute of Health Informatics(健康信息学研究所)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08527 2025-07-15 cs.CV 78%

Leveraging Segment Anything Model for Source-Free Domain Adaptation via Dual Feature Guided Auto-Prompting

Zheang Huai, Hui Tang, Yi Li, Zhuangzhuang Chen, Xiaomeng Li

机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学)

专题命中 领域大模型 :prompting(title,abstract)

Comments Accepted in TMI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08460 2025-07-14 cs.CV 78%

F3-Net: Foundation Model for Full Abnormality Segmentation of Medical Images with Flexible Input Modality Requirement

Seyedeh Sahar Taheri Otaghsara, Reza Rahmanzadeh

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15784 2025-07-08 cs.CV 78%

RL4Med-DDPO: Reinforcement Learning for Controlled Guidance Towards Diverse Medical Image Generation using Vision-Language Foundation Models

Parham Saremi, Amar Kumar, Mohamed Mohamed, Zahra TehraniNasab, Tal Arbel

机构 * Center for Intelligent Machines, McGill University, Montreal, Canada(智能机器中心,麦吉尔大学,加拿大) Mila - Quebec AI institute, Montreal, Canada(魁北克人工智能研究所,加拿大)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23822 2025-07-01 cs.CV 78%

Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model

Shiming Chen, Bowen Duan, Salman Khan, Fahad Shahbaz Khan

机构 * Mohamed bin Zayed University of AI(莫扎伊德大学人工智能学院) Huazhong University of Science and Technology(华中科技大学) Australian National University(澳大利亚国立大学) Linköping University(林雪平大学)

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted to ICCV'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15318 2025-07-01 cs.CV 78%

OpenPath: Open-Set Active Learning for Pathology Image Classification via Pre-trained Vision-Language Models

Lanfeng Zhong, Xin Liao, Shichuan Zhang, Shaoting Zhang, Guotai Wang

机构 * University of Electronic Science and Technology of China(电子科技大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Department of Pathology, West China Second University Hospital, Sichuan University(四川大学病理科西昌第二大学医院部) Department of Radiation Oncology, Sichuan Cancer Hospital and Institute, University of Electronic Science and Technology of China(四川癌症医院与研究所放射肿瘤科,电子科技大学)

专题命中 领域大模型 :language model(title,abstract)

Comments MICCAI 2025 early accept

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21863 2025-06-30 cs.CV 78%

Remote Sensing Large Vision-Language Model: Semantic-augmented Multi-level Alignment and Semantic-aware Expert Modeling

Sungjune Park, Yeongyun Kim, Se Yeon Kim, Yong Man Ro

机构 * Integrated Vision and Language Lab., School of Electrical Engineering, Korea Advanced Institute of Science and Technology (KAIST)(整合视觉与语言实验室,电气工程学院,韩国科学技术院(KAIST))

专题命中 领域大模型 :language model(title,abstract)

Comments 13 pages including reference pages, 7 tables, and 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20299 2025-06-26 cs.CY 78%

Enhancing Programming Pair Workshops: The Case of Teacher Pre-Prompting

Johan Petersson

专题命中 领域大模型 :prompting(title,abstract)

Comments 10 pages, 2 figures. Author's preprint of article published in SIGED/ECISER 2024 via AIS Electronic Library. The published version is available at: https://aisel.aisnet.org/siged2024/15/

Journal ref In Proc. SIGED 2024, AISel. https://aisel.aisnet.org/siged2024/15/ (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05142 2025-06-23 eess.IV cs.CV 78%

Chest X-ray Foundation Model with Global and Local Representations Integration

Zefan Yang, Xuanang Xu, Jiajin Zhang, Ge Wang, Mannudeep K. Kalra, Pingkun Yan

机构 * Department of Biomedical Engineering and Center for Biotechnology and Interdisciplinary Studies, Rensselaer Polytechnic Institute(生物医学工程系和生物技术及跨学科研究所以及罗切斯特理工学院) Department of Radiology, Massachusetts General Hospital, Harvard Medical School(放射学系、麻省总医院、哈佛医学院)

专题命中 领域大模型 :foundation model(title,abstract)

Comments Accepted by IEEE Transactions on Medical Imaging (TMI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13306 2025-06-17 eess.IV cs.CV 78%

Brain Imaging Foundation Models, Are We There Yet? A Systematic Review of Foundation Models for Brain Imaging and Biomedical Research

Salah Ghamizi, Georgia Kanli, Yu Deng, Magali Perquin, Olivier Keunen

机构 * Luxembourg Institute of Health (LIH)(卢森堡健康研究所) King's College London, UK(伦敦国王学院)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12733 2025-06-17 cs.CV 78%

Learning to Fuse: Modality-Aware Adaptive Scheduling for Robust Multimodal Foundation Models

Liam Bennett, Mason Clark, Lucas Anderson, Hana Satou, Olivia Martinez

机构 * Alan Mitkiy, Michael Johnson, Sofia García, Hana Satou(未知机构)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11870 2025-06-16 cs.DB 78%

LLM-based Dynamic Differential Testing for Database Connectors with Reinforcement Learning-Guided Prompt Selection

Ce Lyu, Minghao Zhao, Yanhao Wang, Liang Jie

专题命中 领域大模型 :LLM(title,abstract)

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06270 2025-06-16 cs.IR 78%

RecGPT: A Foundation Model for Sequential Recommendation

Yangqin Jiang, Xubin Ren, Lianghao Xia, Da Luo, Kangyi Lin, Chao Huang

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08949 2025-06-11 cs.CV 78%

SSS: Semi-Supervised SAM-2 with Efficient Prompting for Medical Imaging Segmentation

Hongjie Zhu, Xiwei Liu, Rundong Xue, Zeyu Zhang, Yong Xu, Daji Ergu, Ying Cai, Yang Zhao

机构 * SWUN(西南大学) MBZUAI(慕尼黑工业大学人工智能研究所) XJTU(西安交通大学) ANU(澳大利亚国立大学) La Trobe(拉特罗布大学)

专题命中 领域大模型 :prompting(title);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07988 2025-06-11 cs.CV 78%

MedVersa: A Generalist Foundation Model for Medical Image Interpretation

Hong-Yu Zhou, Julián Nicolás Acosta, Subathra Adithan, Suvrankar Datta, Eric J. Topol, Pranav Rajpurkar

机构 * Harvard Medical School(哈佛医学院) Scripps Research Translational Institute(斯克里普斯研究翻译研究所)

专题命中 领域大模型 :foundation model(title,abstract)

Comments Technical study

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05184 2025-06-06 cs.CV 78%

Single GPU Task Adaptation of Pathology Foundation Models for Whole Slide Image Analysis

Neeraj Kumar, Swaraj Nanda, Siddharth Singi, Jamal Benhamida, David Kim, Jie-Fu Chen, Amir Momeni-Boroujeni, Gregory M. Goldgof, Gabriele Campanella, Chad Vanderbilt

机构 * Memorial Sloan Kettering Cancer Center(纪念斯隆凯特琳癌症中心) Icahn School of Medicine at Mount Sinai(辛格内医学中心) Windreich Department of AI and Human Health(人工智能与人类健康部门) Hasso Platner Institute at Mount Sinai(辛格内霍斯索普研究所)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02854 2025-06-04 cs.CV 78%

Hierarchical Self-Prompting SAM: A Prompt-Free Medical Image Segmentation Framework

Mengmeng Zhang, Xingyuan Dai, Yicheng Sun, Jing Wang, Yueyang Yao, Xiaoyan Gong, Fuze Cong, Feiyue Wang, Yisheng Lv

机构 * State Key Laboratory of Multimodal Artificial Intelligence System, Institute of Automation, Chinese Academy of Science, China(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院,中国) School of Artificial Intelligence, University of Chinese Academy of Science, China(人工智能学院,中国科学院大学,中国) Department of Radiology, Peking Union Medical College Hospital, Peking Union Medical College, Chinese Academy of Medical Sciences, Beijing(放射科,北医三院,北京医科大学,中国医学科学院,北京)

专题命中 领域大模型 :prompting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24173 2025-06-02 cs.CV 78%

DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis?

Tianhong Zhou, Yin Xu, Yingtao Zhu, Chuxi Xiao, Haiyang Bian, Lei Wei, Xuegong Zhang

机构 * Tsinghua University(清华大学)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15625 2025-06-02 cs.LG cs.AI cs.CL cs.DC 78%

Improving Parallel Program Performance with LLM Optimizers via Agent-System Interfaces

Anjiang Wei, Allen Nie, Thiago S. F. X. Teixeira, Rohan Yadav, Wonchan Lee, Ke Wang, Alex Aiken

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15425 2025-05-26 cs.CV 78%

On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?

Raza Imam, Rufael Marew, Mohammad Yaqub

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·穆萨大学人工智能学院)

专题命中 领域大模型 :language model(title,abstract)

Comments Dataset and Code is available at https://github.com/BioMedIA-MBZUAI/RobustMedCLIP Accepted at: Medical Image Understanding and Analysis (MIUA) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10579 2025-05-20 cs.CV 78%

Bias and Generalizability of Foundation Models across Datasets in Breast Mammography

Elodie Germani, Ilayda Selin Türk, Fatima Zeineddine, Charbel Mourad, Shadi Albarqouni

机构 * University Hospital Bonn(波恩大学医院) Technical University of Munich(慕尼黑技术大学) Lebanese Hospital Geitaoui(黎巴嫩Geitaoui医院) Helmholtz AI, Helmholtz Munich(海德堡人工智能,海德堡慕尼黑)

专题命中 领域大模型 :foundation model(title,abstract)

Comments Accepted at the International Conference on Medical Image Computing and Computer-Assisted Intervention (MICCAI) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09564 2025-05-15 cs.CV 78%

Using Foundation Models as Pseudo-Label Generators for Pre-Clinical 4D Cardiac CT Segmentation

Anne-Marie Rickmann, Stephanie L. Thorn, Shawn S. Ahn, Supum Lee, Selen Uman, Taras Lysyy, Rachel Burns, Nicole Guerrera, Francis G. Spinale, Jason A. Burdick, Albert J. Sinusas, James S. Duncan

机构 * Yale University(耶鲁大学) University of Pennsylvania(宾夕法尼亚大学) Washington University School of Medicine in St Louis(华盛顿大学圣路易斯医学学院) University of South Carolina School of Medicine(南卡罗来纳大学医学院) University of Colorado Boulder(科罗拉多大学波德分校)

专题命中 领域大模型 :foundation model(title,abstract)

Comments accepted at FIMH 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08063 2025-05-14 cs.HC 78%

Who's the Leader? Analyzing Novice Workflows in LLM-Assisted Debugging of Machine Learning Code

Jessica Y. Bo, Majeed Kazemitabaar, Emma Zhuang, Ashton Anderson

专题命中 领域大模型 :LLM(title,abstract)

Comments Tools for Thought Workshop at CHI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06217 2025-05-12 cs.CV 78%

Adapting a Segmentation Foundation Model for Medical Image Classification

Pengfei Gu, Haoteng Tang, Islam A. Ebeid, Jose A. Nunez, Fabian Vazquez, Diego Adame, Marcus Zhan, Huimin Li, Bin Fu, Danny Z. Chen

机构 * Department of Computer Science, University of Texas Rio Grande Valley(德克萨斯理工大学里奥格兰德分校计算机科学系) Department of Computer Science, Texas Woman’s University(德克萨斯女性大学计算机科学系) Sewickley Academy(塞维克利学院) Department of Mathematical Sciences, The University of Texas at Dallas(德克萨斯大学达拉斯分校数学科学系) Department of Computer Science and Engineering, University of Notre Dame(诺特丹大学计算机科学与工程系)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02971 2025-05-07 cs.CV 78%

Adversarial Robustness Analysis of Vision-Language Models in Medical Image Segmentation

Anjila Budathoki, Manish Dhakal

机构 * Department of Computer Science(计算机科学系) Georgia State University(佐治亚州立大学)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02699 2025-05-06 cs.HC 78%

Exploring LLM-Powered Role and Action-Switching Pedagogical Agents for History Education in Virtual Reality

Zihao Zhu, Ao Yu, Xin Tong, Pan Hui

专题命中 领域大模型 :LLM(title,abstract)

Comments 14 pages excluding reference and appendix. Accepted at ACM CHI 2025. https://dl.acm.org/doi/10.1145/3706598.3713109

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01695 2025-05-06 cs.IR 78%

SimAug: Enhancing Recommendation with Pretrained Language Models for Dense and Balanced Data Augmentation

Yuying Zhao, Xiaodong Yang, Huiyuan Chen, Xiran Fan, Yu Wang, Yiwei Cai, Tyler Derr

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏