arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 11714 信号源:cs.CL, cs.AI, cs.LG

1. 指令微调 11714 篇

2509.07613 2025-10-16 cs.CV 78%

Data-Efficient Fine-Tuning of Vision-Language Models for Diagnosis of Alzheimer's Disease

Fangqi Cheng, Surajit Ray, Xiaochen Yang

机构 * School of Mathematics and Statistics, University of Glasgow, UK(数学与统计学学院,格拉斯哥大学)

专题命中 指令微调 :language model(title,abstract)

Comments Accepted at MICAD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12741 2025-10-15 cs.CV cs.DC 78%

Personalized Federated Fine-Tuning of Vision Foundation Models for Healthcare

Adam Tupper, Christian Gagné

机构 * Institut intelligence et données (IID)(智能与数据研究所) Université Laval(拉瓦尔大学) Mila - Quebec AI Institute(魁北克人工智能研究所)

专题命中 指令微调 :foundation model(title,abstract)

Comments Accepted to the Symposium on Model Accountability, Sustainability and Healthcare (SMASH) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10584 2025-10-14 cs.CV 78%

Equipping Vision Foundation Model with Mixture of Experts for Out-of-Distribution Detection

Shizhen Zhao, Jiahui Liu, Xin Wen, Haoru Tan, Xiaojuan Qi

机构 * The University of Hong Kong(香港大学)

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.00433 2025-10-14 cs.CR 78%

PrivTuner with Homomorphic Encryption and LoRA: A P3EFT Scheme for Privacy-Preserving Parameter-Efficient Fine-Tuning of AI Foundation Models

Yang Li, Wenhan Yu, Jun Zhao

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09867 2025-10-14 cs.CV 78%

Cluster-Aware Prompt Ensemble Learning for Few-Shot Vision-Language Model Adaptation

Zhi Chen, Xin Yu, Xiaohui Tao, Yan Li, Zi Huang

机构 * University of Southern Queensland(昆士兰南方大学) University of Queensland(昆士兰大学)

专题命中 指令微调 :language model(title,abstract)

Comments Accepted to the journal Pattern Recognition in 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07319 2025-10-09 cs.CV 78%

Temporal Prompting Matters: Rethinking Referring Video Object Segmentation

Ci-Siang Lin, Min-Hung Chen, I-Jieh Liu, Chien-Yi Wang, Sifei Liu, Yu-Chiang Frank Wang

机构 * Graduate Institute of Communication Engineering, National Taiwan University, Taiwan(台湾国立台湾大学通信工程研究所) NVIDIA

专题命中 指令微调 :prompting(title);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07277 2025-10-09 cs.CV 78%

Evaluating Fundus-Specific Foundation Models for Diabetic Macular Edema Detection

Franco Javier Arellano, José Ignacio Orlando

机构 * Yatiris Group(Yatiris集团) PLADEMA Institute(PLADEMA研究所) UNICEN(UNICEN大学) Tandil, Argentina(阿根廷坦迪尔)

专题命中 指令微调 :foundation model(title,abstract)

Comments Accepted for publication at SIPAIM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21273 2025-09-26 cs.CV 78%

A Sentinel-3 foundation model for ocean colour

Geoffrey Dawson, Remy Vandaele, Andrew Taylor, David Moffat, Helen Tamura-Wicks, Sarah Jackson, Rosie Lickorish, Paolo Fraccaro, Hywel Williams, Chunbo Luo, Anne Jones

机构 * IBM Research Europe(IBM欧洲研究院) University of Exeter(埃克塞特大学) STFC Hartree Centre(科学与技术设施委员会哈特里中心) Plymouth Marine Laboratory National Center for Earth Observation(普利茅斯海洋实验室地球观测国家中心)

专题命中 指令微调 :foundation model(title,abstract)

Comments 15 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18104 2025-09-22 cs.CV 78%

PromptMID: Modal Invariant Descriptors Based on Diffusion and Vision Foundation Models for Optical-SAR Image Matching

Han Nie, Bin Luo, Jun Liu, Zhitao Fu, Huan Zhou, Shuo Zhang, Weixing Liu

机构 * State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing, Wuhan University(信息工程测绘与遥感国家重点实验室,武汉大学)

专题命中 指令微调 :foundation model(title,abstract)

Comments 15 pages, 8 figures

Journal ref ISPRS2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14664 2025-09-19 cs.CV 78%

Attention Lattice Adapter: Visual Explanation Generation for Visual Foundation Model

Shinnosuke Hirano, Yuiga Wada, Tsumugi Iida, Komei Sugiura

机构 * Keio University(庆应大学)

专题命中 指令微调 :foundation model(title,abstract)

Comments Accepted for presentation at ICONIP2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14707 2025-09-16 cs.CV 78%

Seeing Further on the Shoulders of Giants: Knowledge Inheritance for Vision Foundation Models

Jiabo Huang, Chen Chen, Lingjuan Lyu

专题命中 指令微调 :foundation model(title,abstract)

Comments Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09572 2025-09-12 cs.CV 78%

PeftCD: Leveraging Vision Foundation Models with Parameter-Efficient Fine-Tuning for Remote Sensing Change Detection

Sijun Dong, Yuxuan Hu, LiBo Wang, Geng Chen, Xiaoliang Meng

机构 * School of Remote Sensing and Information Engineering, Wuhan University(武汉大学遥感与信息工程学院) School of Remote Sensing and Geomatics Engineering, Nanjing University of Information Science and Technology(南京信息工程大学遥感与地理信息工程学院) Guangxi Water & Power Design Institute CO., Ltd.(广西水电设计院有限公司)

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06992 2025-09-10 cs.CV 78%

FedAPT: Federated Adversarial Prompt Tuning for Vision-Language Models

Kun Zhai, Siheng Chen, Xingjun Ma, Yu-Gang Jiang

机构 * Fudan University, Shanghai, China(复旦大学) Shanghai Jiao Tong University, Shanghai, China(上海交通大学)

专题命中 指令微调 :language model(title,abstract)

Comments ACM MM25

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06096 2025-09-09 cs.CV 78%

MedSeqFT: Sequential Fine-tuning Foundation Models for 3D Medical Image Segmentation

Yiwen Ye, Yicheng Wu, Xiangde Luo, He Zhang, Ziyang Chen, Ting Dang, Yanning Zhang, Yong Xia

机构 * School of Computer Science and Engineering, Northwestern Polytechnical University(计算机科学与工程学院,西北工业大学) Monash University(墨尔本大学) Department of Radiation Oncology, Sichuan Cancer Hospital(肿瘤医院放射肿瘤科) RMIT(皇家墨尔本理工大学) University of Melbourne(墨尔本大学)

专题命中 指令微调 :foundation model(title,abstract)

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03895 2025-09-05 cs.CV 78%

Attn-Adapter: Attention Is All You Need for Online Few-shot Learner of Vision-Language Model

Phuoc-Nguyen Bui, Khanh-Binh Nguyen, Hyunseung Choo

机构 * Sungkyunkwan University(顺天大学) Deakin University(德金大学)

专题命中 指令微调 :language model(title,abstract)

Comments ICCV 2025 - LIMIT Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03635 2025-09-05 cs.CV 78%

Reg3D: Reconstructive Geometry Instruction Tuning for 3D Scene Understanding

Hongpei Zheng, Lintao Xiang, Qijun Yang, Qian Lin, Hujun Yin

机构 * Department of Electrical and Electronic Engineering, The University of Manchester(曼彻斯特大学电子与电气工程系)

专题命中 指令微调 :instruction tuning(title,abstract)

Comments 16 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20860 2025-09-03 cs.CV 78%

FedMVP: Federated Multimodal Visual Prompt Tuning for Vision-Language Models

Mainak Singha, Subhankar Roy, Sarthak Mehrotra, Ankit Jha, Moloud Abdar, Biplab Banerjee, Elisa Ricci

机构 * University of Trento(特伦托大学) University of Bergamo(贝拉姆奥大学) Indian Institute of Technology Bombay(印度班加罗尔理工学院) LNMIIT Jaipur(斋普尔LNMIIT) The University of Queensland(昆士兰大学) Fondazione Bruno Kessler(布鲁诺·凯塞勒基金会)

专题命中 指令微调 :language model(title,abstract)

Comments Accepted in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15144 2025-09-03 cs.CV 78%

Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models

Ankit Yadav, Lingqiao Liu, Yuankai Qi

机构 * The University of Adelaide(阿德莱德大学) Macquarie University(麦考瑞大学)

专题命中 指令微调 :language model(title,abstract)

Comments 25 Pages Accepted in DICTA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00374 2025-09-03 cs.CV 78%

Adaptive Point-Prompt Tuning: Fine-Tuning Heterogeneous Foundation Models for 3D Point Cloud Analysis

Mengke Li, Lihao Chen, Peng Zhang, Yiu-ming Cheung, Hui Huang

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济实验室) National Laboratory of Radar Signal Processing, Xidian University(雷达信号处理国家实验室) Department of Computer Science, Hong Kong Baptist University(香港 Baptist 大学计算机科学系)

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20830 2025-08-29 cs.CV 78%

Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation

Krit Duangprom, Tryphon Lambrou, Binod Bhattarai

机构 * University of Aberdeen(阿伯丁大学)

专题命中 指令微调 :language model(title,abstract)

Comments Accepted to MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19746 2025-08-28 cs.CV 78%

SPLF-SAM: Self-Prompting Segment Anything Model for Light Field Salient Object Detection

Qiyao Xu, Qiming Wu, Xiaowei Li

机构 * College of Electronic Engineering and Information(电子工程与信息学院) Sichuan University(四川大学) Chengdu, China(中国成都)

专题命中 指令微调 :prompting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14931 2025-08-22 eess.IV cs.GR 78%

Pixels Under Pressure: Exploring Fine-Tuning Paradigms for Foundation Models in High-Resolution Medical Imaging

Zahra TehraniNasab, Amar Kumar, Tal Arbel

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12877 2025-08-19 cs.CV 78%

Preserve and Sculpt: Manifold-Aligned Fine-tuning of Vision-Language Models for Few-Shot Learning

Dexia Chen, Qianjie Zhu, Weibing Li, Yue Yu, Tong Zhang, Ruixuan Wang

专题命中 指令微调 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09061 2025-08-13 cs.CV 78%

VLM-3D:End-to-End Vision-Language Models for Open-World 3D Perception

Fuhao Chang, Shuxin Li, Yabei Li, Lei He

机构 * School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动性学院) State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University(清华大学智能绿色车辆与移动性国家重点实验室) College of Information and Electrical Engineering, China Agricultural University(中国农业大学信息与电气工程学院) School of Statistics and Data Science, Southwestern University of Finance and Economics(西南财经大学统计与数据科学学院) Meituan, China(美团(中国))

专题命中 指令微调 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17963 2025-08-08 cs.CV 78%

M$^{2}$Chat: Empowering VLM for Multimodal LLM Interleaved Text-Image Generation

Xiaowei Chi, Junbo Qi, Rongyu Zhang, Shanghang Zhang, Qifeng Liu, Yike Guo

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Waseda University(早稻田大学) Peking University(北京大学)

专题命中 指令微调 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01582 2025-08-05 cs.CV 78%

Set Pivot Learning: Redefining Generalized Segmentation with Vision Foundation Models

Xinhui Li, Xinyu He, Qiming Hu, Xiaojie Guo

机构 * College of Intelligence and Computing, Tianjin University(智能与计算学院,天津大学)

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01361 2025-08-05 cs.RO 78%

VLH: Vision-Language-Haptics Foundation Model

Luis Francisco Moreno Fuentes, Muhammad Haris Khan, Miguel Altamirano Cabrera, Valerii Serpiva, Dmitri Iarchuk, Yara Mahmoud, Issatay Tokmurziyev, Dzmitry Tsetserukou

机构 * Intelligent Space Robotics Laboratory(智能空间机器人实验室) Skolkovo Institute of Science and Technology(斯克尔科沃科学与技术研究所)

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19847 2025-07-30 cs.CV 78%

Knowledge Regularized Negative Feature Tuning of Vision-Language Models for Out-of-Distribution Detection

Wenjie Zhu, Yabin Zhang, Xin Jin, Wenjun Zeng, Lei Zhang

机构 * Hong Kong Polytechnic University(香港理工大学) Eastern Institute of Technology(东部技术研究所) Stanford University(斯坦福大学) Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究所,东部技术研究所) Institute for Clarity in Documentation(文档清晰研究所) Inria Paris-Rocquencourt(巴黎-罗克琴堡研究所) Rajiv Gandhi University(拉贾·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕勒研究实验室)

专题命中 指令微调 :language model(title,abstract)

Comments accepted by ACMMM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15582 2025-07-30 cond-mat.mtrl-sci 78%

Fine-tuning foundation models of materials interatomic potentials with frozen transfer learning

Mariia Radova, Wojciech G. Stark, Connor S. Allen, Reinhard J. Maurer, Albert P. Bartók

专题命中 指令微调 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20972 2025-07-29 astro-ph.IM astro-ph.SR 78%

Finetuning Stellar Spectra Foundation Models with LoRA

Xiaosheng Zhao, Yuan-Sen Ting, Alexander S. Szalay, Yang Huang

专题命中 指令微调 :foundation model(title,abstract)

Comments 7 pages, 2 figures. Accepted to the Machine Learning for Astrophysics (ML4Astro) Colocated Workshop at ICML 2025. Presented as a spotlight talk

详情

展开后加载摘要…

URL PDF HTML 收藏