arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12705 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12705 篇

2410.13242 2024-10-21 cs.CV 78%

Fundus to Fluorescein Angiography Video Generation as a Retinal Generative Foundation Model

Weiyi Zhang, Jiancheng Yang, Ruoyu Chen, Siyu Huang, Pusheng Xu, Xiaolan Chen, Shanfu Lu, Hongyu Cao, Mingguang He, Danli Shi

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12645 2024-10-17 eess.AS eess.SP 78%

Beyond Speech and More: Investigating the Emergent Ability of Speech Foundation Models for Classifying Physiological Time-Series Signals

Orchid Chetia Phukan, Swarup Ranjan Behera, Girish, Mohd Mujtaba Akhtar, Arun Balaji Buduru, Rajesh Sharma

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10526 2024-10-15 cs.CR 78%

Generalized Adversarial Code-Suggestions: Exploiting Contexts of LLM-based Code-Completion

Karl Rubel, Maximilian Noppel, Christian Wressnegger

专题命中 领域大模型 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07460 2024-10-11 cs.CV 78%

Generalizing Segmentation Foundation Model Under Sim-to-real Domain-shift for Guidewire Segmentation in X-ray Fluoroscopy

Yuxuan Wen, Evgenia Roussinova, Olivier Brina, Paolo Machi, Mohamed Bouri

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04251 2024-10-08 cs.LG cs.AI cs.CL cs.SI quant-ph 78%

Enhancing Future Link Prediction in Quantum Computing Semantic Networks through LLM-Initiated Node Features

Gilchan Park, Paul Baity, Byung-Jun Yoon, Adolfy Hoisie

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01573 2024-10-03 cs.CV 78%

PASS:Test-Time Prompting to Adapt Styles and Semantic Shapes in Medical Image Segmentation

Chuyan Zhang, Hao Zheng, Xin You, Yefeng Zheng, Yun Gu

专题命中 领域大模型 :prompting(title,abstract)

Comments Submitted to IEEE TMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15574 2024-09-25 cs.CV 78%

Clinical-grade Multi-Organ Pathology Report Generation for Multi-scale Whole Slide Images via a Semantically Guided Medical Text Foundation Model

Jing Wei Tan, SeungKyu Kim, Eunsu Kim, Sung Hak Lee, Sangjeong Ahn, Won-Ki Jeong

专题命中 领域大模型 :foundation model(title);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12011 2024-09-19 cs.CV 78%

Mixture of Prompt Learning for Vision Language Models

Yu Du, Tong Niu, Rong Zhao

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07048 2024-09-12 cs.CV 78%

Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations

Keumgang Cha, Donggeun Yu, Junghoon Seo

专题命中 领域大模型 :language model(title);foundation model(abstract)

Comments This study was primarily conducted during the latter half of 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.11421 2024-09-05 cs.CV 78%

Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement

Weijian Huang, Cheng Li, Hao Yang, Jiarun Liu, Yong Liang, Hairong Zheng, Shanshan Wang

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.16769 2024-08-30 cs.CV cs.CR 78%

PromptSmooth: Certifying Robustness of Medical Vision-Language Models via Prompt Learning

Noor Hussein, Fahad Shamshad, Muzammal Naseer, Karthik Nandakumar

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted to MICCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12396 2024-08-23 cs.CV physics.geo-ph 78%

Cross-Domain Foundation Model Adaptation: Pioneering Computer Vision Models for Geophysical Data Analysis

Zhixiang Guo, Xinming Wu, Luming Liang, Hanlin Sheng, Nuo Chen, Zhengfa Bi

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.08796 2024-08-15 cs.IR 78%

The Elephant in the Room: Rethinking the Usage of Pre-trained Language Model in Sequential Recommendation

Zekai Qu, Ruobing Xie, Chaojun Xiao, Xingwu Sun, Zhanhui Kang

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted at RecSys 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21317 2024-08-07 cs.CV 78%

Pathology Foundation Models

Mieko Ochi, Daisuke Komura, Shumpei Ishikawa

专题命中 领域大模型 :foundation model(title,abstract)

Comments 19 pages, 1 figure, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21465 2024-08-01 cs.CV 78%

MarvelOVD: Marrying Object Recognition and Vision-Language Models for Robust Open-Vocabulary Object Detection

Kuo Wang, Lechao Cheng, Weikai Chen, Pingping Zhang, Liang Lin, Fan Zhou, Guanbin Li

专题命中 领域大模型 :language model(title,abstract)

Comments Codes are available at https://github.com/wkfdb/MarvelOVD

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15728 2024-07-25 eess.IV cs.CV 78%

SAM2CLIP2SAM: Vision Language Model for Segmentation of 3D CT Scans for Covid-19 Detection

Dimitrios Kollias, Anastasios Arsenos, James Wingate, Stefanos Kollias

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.09585 2024-07-23 cs.SD eess.AS 78%

Domain Adaptation for Contrastive Audio-Language Models

Soham Deshmukh, Rita Singh, Bhiksha Raj

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted at INTERSPEECH 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.12128 2024-07-16 cs.CV cs.AI cs.CL cs.LG 78%

DiagrammerGPT: Generating Open-Domain, Open-Platform Diagrams via LLM Planning

Abhay Zala, Han Lin, Jaemin Cho, Mohit Bansal

专题命中 领域大模型 :LLM(title);分类 cs.CL、cs.AI、cs.LG

Comments COLM 2024; Project page: https://diagrammerGPT.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06104 2024-07-16 cs.CV 78%

Debiased Noise Editing on Foundation Models for Fair Medical Image Classification

Ruinan Jin, Wenlong Deng, Minghui Chen, Xiaoxiao Li

专题命中 领域大模型 :foundation model(title,abstract)

Comments 13 pages, 3 figures. Accepted by MICCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06508 2024-07-12 eess.IV cs.CV 78%

A Clinical Benchmark of Public Self-Supervised Pathology Foundation Models

Gabriele Campanella, Shengjia Chen, Ruchika Verma, Jennifer Zeng, Aryeh Stock, Matt Croken, Brandon Veremis, Abdulkadir Elmas, Kuan-lin Huang, Ricky Kwan, Jane Houldsworth, Adam J. Schoenfeld, Chad Vanderbilt

专题命中 领域大模型 :foundation model(title,abstract)

Comments arXiv admin note: text overlap with arXiv:2310.07033

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.02235 2024-07-12 cs.CV 78%

Image Fusion via Vision-Language Model

Zixiang Zhao, Lilun Deng, Haowen Bai, Yukun Cui, Zhipeng Zhang, Yulun Zhang, Haotong Qin, Dongdong Chen, Jiangshe Zhang, Peng Wang, Luc Van Gool

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted by International Conference on Machine Learning (ICML) 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00621 2024-07-04 cs.IR cs.MM 78%

Multimodal Pretraining, Adaptation, and Generation for Recommendation: A Survey

Qijiong Liu, Jieming Zhu, Yanting Yang, Quanyu Dai, Zhaocheng Du, Xiao-Ming Wu, Zhou Zhao, Rui Zhang, Zhenhua Dong

专题命中 领域大模型 :pretraining(title,abstract)

Comments Accepted by KDD 2024. See our tutorial materials at https://mmrec.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18070 2024-07-02 cs.CV 78%

EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation

Baoqi Pei, Guo Chen, Jilan Xu, Yuping He, Yicheng Liu, Kanghua Pan, Yifei Huang, Yali Wang, Tong Lu, Limin Wang, Yu Qiao

专题命中 领域大模型 :foundation model(title,abstract)

Comments Champion solutions in the EgoVis CVPR 2024 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.11037 2024-06-18 cs.SD eess.AS 78%

NAST: Noise Aware Speech Tokenization for Speech Language Models

Shoval Messica, Yossi Adi

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted at Interspeech 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19675 2024-05-31 cs.CV 78%

Knowledge-grounded Adaptation Strategy for Vision-language Models: Building Unique Case-set for Screening Mammograms for Residents Training

Aisha Urooj Khan, John Garrett, Tyler Bradshaw, Lonie Salkowski, Jiwoong Jason Jeong, Amara Tariq, Imon Banerjee

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09446 2024-05-16 eess.IV 78%

M$^4$oE: A Foundation Model for Medical Multimodal Image Segmentation with Mixture of Experts

Yufeng Jiang, Yiqing Shen

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.06927 2024-05-14 cs.IR 78%

Multimodal Pretraining and Generation for Recommendation: A Tutorial

Jieming Zhu, Chuhan Wu, Rui Zhang, Zhenhua Dong

专题命中 领域大模型 :pretraining(title,abstract)

Comments Published in WWW 2024 Tutorial. Find the tutorial materials at https://mmrec.github.io/tutorial/www2024/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.04771 2024-05-09 cs.CV 78%

Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches

Qing Yu, Mikihiro Tanaka, Kent Fujiwara

专题命中 领域大模型 :language model(title,abstract)

Comments Accepted to CVPR 2024, Project website: https://yu1ut.com/MotionPatches-HP/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18720 2024-04-30 cs.RO 78%

Innovative Integration of Visual Foundation Model with a Robotic Arm on a Mobile Platform

Shimian Zhang, Qiuhong Lu

专题命中 领域大模型 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.17534 2024-04-29 cs.CV cs.MM 78%

Exploring the Distinctiveness and Fidelity of the Descriptions Generated by Large Vision-Language Models

Yuhang Huang, Zihan Wu, Chongyang Gao, Jiawei Peng, Xu Yang

专题命中 领域大模型 :language model(title,abstract)

Comments 11 pages, 9 figures, 6 tables. For associated code, see https://anonymous.4open.science/r/Explore_FGVDs-E277

详情

展开后加载摘要…

URL PDF HTML 收藏