arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12705 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12705 篇

2505.01558 2025-05-06 cs.CV 78%

A Sensor Agnostic Domain Generalization Framework for Leveraging Geospatial Foundation Models: Enhancing Semantic Segmentation viaSynergistic Pseudo-Labeling and Generative Learning

Anan Yaghmour, Melba M. Crawford, Saurabh Prasad

机构 * University of Houston(休斯敦大学) Purdue University(普渡大学)

专题命中 领域大模型 :foundation model(title,abstract)

Comments Accepted in the 2025 CVPR Workshop on Foundation and Large Vision Models in Remote Sensing, to appear in CVPR 2025 Workshop Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15922 2025-04-24 cs.SE 78%

Language Models to Support Multi-Label Classification of Industrial Data

Waleed Abdeen, Michael Unterkalmsteiner, Krzysztof Wnuk, Alessio Ferrari, Panagiota Chatzipetrou

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted at SANER Conference 2025. Awaiting publication by IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15132 2025-04-22 cs.HC cs.CY 78%

Investigating Youth's Technical and Ethical Understanding of Generative Language Models When Engaging in Construction and Deconstruction Activities

Luis Morales-Navarro

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10642 2025-04-16 cs.CV 78%

SilVar-Med: A Speech-Driven Visual Language Model for Explainable Abnormality Detection in Medical Imaging

Tan-Hanh Pham, Chris Ngo, Trong-Duong Bui, Minh Luu Quang, Tan-Huong Pham, Truong-Son Hy

机构 * Florida Institute of Technology(佛罗里达理工学院) Knovel Engineering Lab(诺维尔工程实验室) Vietnam Military Medical University(越南军事医科大学) Military Central Hospital(108陆军中央医院) Can Tho University of Medicine and Pharmacy(芹苴医药大学) University of Alabama at Birmingham(阿拉巴马大学伯明翰分校)

专题命中 领域大模型 :language model(title,abstract)

Comments CVPR Multimodal Algorithmic Reasoning Workshop 2025 - SilVarMed

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.00636 2025-04-02 cs.HC 78%

Exploring the Impact of an LLM-Powered Teachable Agent on Learning Gains and Cognitive Load in Music Education

Lingxi Jin, Baicheng Lin, Mengze Hong, Kun Zhang, Hyo-Jeong So

专题命中 领域大模型 :LLM(title,abstract)

Comments Accepted at CHI 2025 Workshop on Augmented Educators and AI: Shaping the Future of Human and AI Cooperation in Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08813 2025-04-02 cs.CV 78%

Retrieval-augmented Few-shot Medical Image Segmentation with Foundation Models

Lin Zhao, Xiao Chen, Eric Z. Chen, Yikang Liu, Terrence Chen, Shanhui Sun

机构 * United Imaging Intelligence(联影智能)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24345 2025-04-01 cs.CV 78%

PathOrchestra: A Comprehensive Foundation Model for Computational Pathology with Over 100 Diverse Clinical-Grade Tasks

Fang Yan, Jianfeng Wu, Jiawen Li, Wei Wang, Jiaxuan Lu, Wen Chen, Zizhao Gao, Jianan Li, Hong Yan, Jiabo Ma, Minda Chen, Yang Lu, Qing Chen, Yizhi Wang, Xitong Ling, Xuenian Wang, Zihan Wang, Qiang Huang, Shengyi Hua, Mianxin Liu, Lei Ma, Tian Shen, Xiaofan Zhang, Yonghong He, Hao Chen, Shaoting Zhang, Zhe Wang

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Fourth Military Medical University(第四军医大学) Tsinghua University(清华大学) University of Science and Technology of China(中国科学技术大学) SenseTime Research(商汤科技研究院) Hong Kong University of Science and Technology(香港科技大学) Peking University(北京大学) Shengqiang Technology Co. Ltd.(胜强科技有限公司) Institute of Materials Research, Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院材料研究院)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23618 2025-04-01 cs.CV 78%

Leveraging Vision-Language Foundation Models to Reveal Hidden Image-Attribute Relationships in Medical Imaging

Amar Kumar, Anita Kriz, Barak Pertzov, Tal Arbel

机构 * McGill University(麦吉尔大学) MILA-Quebec AI Institute(魁北克人工智能研究所(MILA)) McMaster University(麦克马斯特大学)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23374 2025-04-01 cs.IR 78%

RuleAgent: Discovering Rules for Recommendation Denoising with Autonomous Language Agents

Zongwei Wang, Min Gao, Junliang Yu, Yupeng Hou, Shazia Sadiq, Hongzhi Yin

专题命中 领域大模型 :language agent(title,abstract)

Comments 11 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14522 2025-03-28 cs.CV 78%

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI

Tianbin Li, Yanzhou Su, Wei Li, Bin Fu, Zhe Chen, Ziyan Huang, Guoan Wang, Chenglong Ma, Ying Chen, Ming Hu, Yanjun Li, Pengcheng Chen, Xiaowei Hu, Zhongying Deng, Yuanfeng Ji, Jin Ye, Yu Qiao, Junjun He

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Shanghai Jiao Tong University(上海交通大学) Shenzhen Institute of Advanced Technology (SIAT), Chinese Academy of Sciences(中国科学院深圳先进技术研究院) Nanjing University(南京大学) East China Normal University(华东师范大学) Fudan University(复旦大学) Xiamen University(厦门大学) Monash University(莫纳什大学) University of Washington(华盛顿大学) University of Cambridge(剑桥大学) Stanford University(斯坦福大学)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09430 2025-03-28 cs.CV 78%

Evaluating Pre-trained Convolutional Neural Networks and Foundation Models as Feature Extractors for Content-based Medical Image Retrieval

Amirreza Mahbod, Nematollah Saeidi, Sepideh Hatamikia, Ramona Woitek

机构 * Danube Private University(多瑙私立大学) University of Isfahan(伊斯法罕大学) Austrian Center for Medical Innovation and Technology(奥地利医学创新与技术中心)

专题命中 领域大模型 :foundation model(title,abstract)

Comments 37 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18325 2025-03-25 cs.CV 78%

Towards Training-free Anomaly Detection with Vision and Language Foundation Models

Jinjin Zhang, Guodong Wang, Yizhou Jin, Di Huang

机构 * Beihang University(北京航空航天大学) State Key Laboratory of Complex and Critical Software Environment, Beihang University(北京航空航天大学复杂与关键软件环境国家重点实验室) School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)

专题命中 领域大模型 :foundation model(title,abstract)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15892 2025-03-21 cs.CV 78%

UMIT: Unifying Medical Imaging Tasks via Vision-Language Models

Haiyang Yu, Siyang Yi, Ke Niu, Minghan Zhuo, Bin Li

机构 * Shanghai Key Laboratory of Intelligent Information Processing(上海智能信息处理重点实验室) School of Computer Science, Fudan University(复旦大学计算机科学学院)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14129 2025-03-19 cs.CV 78%

SketchFusion: Learning Universal Sketch Features through Fusing Foundation Models

Subhadeep Koley, Tapas Kumar Dutta, Aneeshan Sain, Pinaki Nath Chowdhury, Ayan Kumar Bhunia, Yi-Zhe Song

机构 * University of Surrey(萨里大学) iFlyTek-Surrey Joint Research Centre on Artificial Intelligence(讯飞-萨里人工智能联合研究中心)

专题命中 领域大模型 :foundation model(title,abstract)

Comments Accepted in CVPR 2025. Project page available at https://subhadeepkoley.github.io/SketchFusion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17386 2025-03-18 eess.IV cs.CV 78%

vesselFM: A Foundation Model for Universal 3D Blood Vessel Segmentation

Bastian Wittmann, Yannick Wattenberg, Tamaz Amiranashvili, Suprosanna Shit, Bjoern Menze

机构 * University of Zurich(苏黎世大学) ETH Zurich(苏黎世联邦理工学院) Technical University of Munich(慕尼黑工业大学)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11619 2025-03-17 cs.CR 78%

Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense

Shuyang Hao, Yiwei Wang, Bryan Hooi, Ming-Hsuan Yang, Jun Liu, Chengcheng Tang, Zi Huang, Yujun Cai

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11576 2025-03-17 cs.CV 78%

SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion

Ahmed Nassar, Andres Marafioti, Matteo Omenetti, Maksym Lysak, Nikolaos Livathinos, Christoph Auer, Lucas Morin, Rafael Teixeira de Lima, Yusik Kim, A. Said Gurbuz, Michele Dolfi, Miquel Farré, Peter W. J. Staar

机构 * IBM Research(IBM研究院) HuggingFace

专题命中 领域大模型 :language model(title,abstract)

Comments 24 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06190 2025-03-11 eess.IV cs.CV 78%

Attention on the Wires (AttWire): A Foundation Model for Detecting Devices and Catheters in X-ray Fluoroscopic Images

YingLiang Ma, Sandra Howell, Aldo Rinaldi, Tarv Dhanjal, Kawal S. Rhode

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06003 2025-03-11 cs.CV 78%

Integrating Frequency-Domain Representations with Low-Rank Adaptation in Vision-Language Models

Md Azim Khan, Aryya Gangopadhyay, Jianwu Wang, Robert F. Erbacher

机构 * University of Maryland Baltimore County (UMBC)(马里兰大学巴尔的摩县分校) Center for Real-time Distributed Sensing and Autonomy (CARDS)(实时分布式感知与自主中心) DEVCOM Army Research Laboratory(DEVCOM陆军研究实验室)

专题命中 领域大模型 :language model(title,abstract)

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04826 2025-03-10 eess.IV cs.CV 78%

Rethinking Few-Shot Medical Image Segmentation by SAM2: A Training-Free Framework with Augmentative Prompting and Dynamic Matching

Haiyue Zu, Jun Ge, Heting Xiao, Jile Xie, Zhangzhe Zhou, Yifan Meng, Jiayi Ni, Junjie Niu, Linlin Zhang, Li Ni, Huilin Yang

机构 * The First Affiliated Hospital of Soochow University(苏州大学附属第一医院) Soochow University(苏州大学)

专题命中 领域大模型 :prompting(title);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12915 2025-03-05 cs.CV 78%

VILA-M3: Enhancing Vision-Language Models with Medical Expert Knowledge

Vishwesh Nath, Wenqi Li, Dong Yang, Andriy Myronenko, Mingxin Zheng, Yao Lu, Zhijian Liu, Hongxu Yin, Yucheng Tang, Pengfei Guo, Can Zhao, Ziyue Xu, Yufan He, Greg Heinrich, Yee Man Law, Benjamin Simon, Stephanie Harmon, Stephen Aylward, Marc Edgar, Michael Zephyr, Song Han, Pavlo Molchanov, Baris Turkbey, Holger Roth, Daguang Xu

机构 * NVIDIA(英伟达) SingHealth(新加坡保健服务集团) NIH(美国国立卫生研究院)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01020 2025-03-04 cs.CV 78%

Delving into Out-of-Distribution Detection with Medical Vision-Language Models

Lie Ju, Sijin Zhou, Yukun Zhou, Huimin Lu, Zhuoting Zhu, Pearse A. Keane, Zongyuan Ge

机构 * Monash University(莫纳什大学) Moorfields Eye Hospital(摩尔菲尔德眼科医院) University College London(伦敦大学学院) Airdoc Technology Inc(鹰瞳科技) Southeast University(东南大学) Melbourne University(墨尔本大学)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00802 2025-03-04 cs.CV 78%

MFM-DA: Instance-Aware Adaptor and Hierarchical Alignment for Efficient Domain Adaptation in Medical Foundation Models

Jia-Xuan Jiang, Wenhui Lei, Yifeng Wu, Hongtao Wu, Furong Li, Yining Xie, Xiaofan Zhang, Zhong Wang

机构 * Lanzhou University(兰州大学) Shanghai Jiaotong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20749 2025-03-03 eess.IV cs.CV 78%

SemiSAM+: Rethinking Semi-Supervised Medical Image Segmentation in the Era of Foundation Models

Yichi Zhang, Bohao Lv, Le Xue, Wenbo Zhang, Yuchen Liu, Yu Fu, Yuan Cheng, Yuan Qi

机构 * Artificial Intelligence Innovation and Incubation Institute, Fudan University(复旦大学人工智能创新与孵化研究院) Shanghai Academy of Artificial Intelligence for Science(上海人工智能科学研究院) Huashan Hospital, Fudan University(复旦大学附属华山医院) Human Phenome Institute, Fudan University(复旦大学人类表型组研究院) School of Information Science and Engineering, Lanzhou University(兰州大学信息科学与工程学院)

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19133 2025-02-27 cs.HC 78%

DBox: Scaffolding Algorithmic Programming Learning through Learner-LLM Co-Decomposition

Shuai Ma, Junling Wang, Yuanhao Zhang, Xiaojuan Ma, April Yi Wang

专题命中 领域大模型 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09001 2025-02-27 eess.IV cs.CV 78%

Vision Foundation Models for Computed Tomography

Suraj Pai, Ibrahim Hadzic, Dennis Bontempi, Keno Bressem, Benjamin H. Kann, Andriy Fedorov, Raymond H. Mak, Hugo J. W. L. Aerts

专题命中 领域大模型 :foundation model(title,abstract)

Comments 6 figures, followed by 9 Extended Data Figures and a Supplementary Information document

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18735 2025-02-27 cs.RO cs.CV 78%

QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries

Nicolas Harvey Chapman, Feras Dayoub, Will Browne, Christopher Lehnert

机构 * School of Electrical Engineering and Robotics, Queensland University of Technology(昆士兰科技大学电气工程与机器人学院)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16832 2025-02-25 cs.CV 78%

FedBM: Stealing Knowledge from Pre-trained Language Models for Heterogeneous Federated Learning

Meilu Zhu, Qiushi Yang, Zhifan Gao, Yixuan Yuan, Jun Liu

机构 * City University of Hong Kong(香港城市大学) Sun Yat-sen University(中山大学) Chinese University of Hong Kong(香港中文大学) The University of Hong Kong(香港大学)

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted by MedIA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16223 2025-02-25 cs.CV 78%

Prompt as Knowledge Bank: Boost Vision-language model via Structural Representation for zero-shot medical detection

Yuguang Yang, Tongfei Chen, Haoyu Huang, Linlin Yang, Chunyu Xie, Dawei Leng, Xianbin Cao, Baochang Zhang

机构 * School of Electronic Information Engineering, Beihang University(北京航空航天大学电子信息工程学院) AI Research, Qihoo 360(奇虎360公司360人工智能研究院) School of Artificial Intelligence, Beihang University(北京航空航天大学人工智能学院) State Key Laboratory of Media Convergence and Communication, Communication University of China(中国传媒大学媒体融合与传播国家重点实验室) National Superior College for Engineers, Beihang University(北京航空航天大学国家优秀工程师学院) Artificial Intelligence Research Center, Lobachevsky State University(罗巴切夫斯基国立大学人工智能研究中心)

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted as ICLR 2025 conference paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14584 2025-02-24 eess.IV cs.CV 78%

Vision Foundation Models in Medical Image Analysis: Advances and Challenges

Pengchen Liang, Bin Pu, Haishan Huang, Yiwei Li, Hualiang Wang, Weibo Ma, Qing Chang

机构 * Shanghai University(上海大学) The Hong Kong University of Science and Technology(香港科技大学) Sun Yat-sen University(中山大学) Shanghai Jiao Tong University(上海交通大学) East China Normal University(华东师范大学) Shanghai Jiao Tong University School of Medicine(上海交通大学医学院)

专题命中 领域大模型 :foundation model(title,abstract)

Comments 17 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏