arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-24 至 2025-09-24 共收录 10 信号源:cs.CL, cs.AI, cs.LG

1. 预训练与数据 10 篇

2409.04183 2025-09-24 cs.CL cs.AI 90%

GALLa: Graph Aligned Large Language Models for Improved Source Code Understanding

Ziyin Zhang, Hang Yu, Shijie Li, Peng Di, Jianguo Li, Rui Wang

机构 * Ant Group(蚂蚁集团) Shanghai Jiao Tong University(上海交通大学)

专题命中 预训练与数据 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18839 2025-09-24 cs.CV 88%

Benchmarking Vision-Language and Multimodal Large Language Models in Zero-shot and Few-shot Scenarios: A study on Christian Iconography

Gianmarco Spinaci, Lukas Klic, Giovanni Colavizza

机构 * Gianmarco Spinaci Department of Classical Philology and Italian Studies, University of Bologna, Italy Villa i Tatti, The Harvard University Center for Italian Renaissance Studies, Florence, Italy(Gianmarco Spinaci 文艺复兴研究系,博洛尼亚大学,意大利 塔蒂别墅,哈佛大学意大利文艺复兴研究中心,佛罗伦萨,意大利) Lukas Klic Villa i Tatti, The Harvard University Center for Italian Renaissance Studies, Florence, Italy(Lukas Klic 塔蒂别墅,哈佛大学意大利文艺复兴研究中心,佛罗伦萨,意大利) Giovanni Colavizza Department of Classical Philology and Italian Studies, University of Bologna, Italy Department of Communication, University of Copenhagen, Denmark(Giovanni Colavizza 文艺复兴研究系,博洛尼亚大学,意大利 传播系,哥本哈根大学,丹麦)

专题命中 预训练与数据 :large language model(title,abstract);language model(title,abstract)

Comments 11 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18569 2025-09-24 cs.SD cs.AI eess.AS 85%

Explore the Reinforcement Learning for the LLM based ASR and TTS system

Changfeng Gao, Yabin Li, Keyu An, Zhifu Gao, Zhihao Du, Han Zhao, Xiangang Li

机构 * Speech Team, Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)

专题命中 预训练与数据 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18152 2025-09-24 cs.LG cs.AI 84%

WLFM: A Well-Logs Foundation Model for Multi-Task and Cross-Well Geological Interpretation

Zhenyu Qi, Qing Yu, Jichen Wang, Yun-Bo Zhao, Zerui Li, Wenjun Lv

机构 * Institute of Advanced Technology, University of Science and Technology of China(科学技术大学先进技术研究所) Department of Automation, University of Science and Technology of China(科学技术大学自动化系) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究所)

专题命中 预训练与数据 :foundation model(title,abstract);pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18383 2025-09-24 cs.CL 77%

NileChat: Towards Linguistically Diverse and Culturally Aware LLMs for Local Communities

Abdellah El Mekki, Houdaifa Atou, Omer Nacar, Shady Shehata, Muhammad Abdul-Mageed

机构 * The University of British Columbia(不列颠哥伦比亚大学) Mohammed VI Polytechnic University(穆莱·阿卜杜勒阿齐兹国王理工学院) Tuwaiq Academy(图瓦伊克学院) Invertible AI(可逆人工智能)

专题命中 预训练与数据 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 (Main Conference). Camera-ready version. Data & models: https://github.com/UBC-NLP/nilechat

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18938 2025-09-24 cs.CV cs.AI cs.LG 73%

No Labels Needed: Zero-Shot Image Classification with Collaborative Self-Learning

Matheus Vinícius Todescato, Joel Luís Carbonera

机构 * Institute of Informatics(信息学院) UFRGS(乌拉圭国家研究学院)

专题命中 预训练与数据 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments This paper was accepted at International Conference on Tools with Artificial Intelligence (ICTAI) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18535 2025-09-24 cs.CL eess.SP 70%

Trace Is In Sentences: Unbiased Lightweight ChatGPT-Generated Text Detector

Mo Mu, Dianqiao Lei, Chang Li

专题命中 预训练与数据 :LLM(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18483 2025-09-24 cs.LG quant-ph 70%

Physics-informed time series analysis with Kolmogorov-Arnold Networks under Ehrenfest constraints

Abhijit Sen, Illya V. Lukin, Kurt Jacobs, Lev Kaplan, Andrii G. Sotnikov, Denys I. Bondar

机构 * Department of Physics and Engineering Physics, Tulane University, New Orleans, Louisiana 70118, USA(物理系和工程物理系, Tulane大学, 新奥尔良,路易斯安那州70118,美国) Karazin Kharkiv National University, Svobody Square 4, 61022 Kharkiv, Ukraine(卡扎林基赫夫国家大学,Svobody广场4号,61022基赫夫,乌克兰) Akhiezer Institute for Theoretical Physics, NSC KIPT, Akademichna 1, 61108 Kharkiv, Ukraine(阿赫伊泽尔理论物理研究所,NSC KIPT,Akademichna 1号,61108基赫夫,乌克兰) United States DEVCOM Army Research Laboratory, Adelphi, Maryland 20783, USA(美国DEVCOM陆军研究实验室,阿德菲,马里兰州20783,美国) Department of Physics, University of Massachusetts at Boston, Boston, Massachusetts 02125, USA(物理系,马萨诸塞大学波士顿分校,波士顿,马萨诸塞州02125,美国)

专题命中 预训练与数据 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18369 2025-09-24 cs.CV cs.AI 70%

Align Where the Words Look: Cross-Attention-Guided Patch Alignment with Contrastive and Transport Regularization for Bengali Captioning

Riad Ahmed Anonto, Sardar Md. Saffat Zabin, M. Saifur Rahman

机构 * Bangladesh University of Engineering and Technology (BUET)(孟加拉工程与技术大学)

专题命中 预训练与数据 :language model(abstract);pretraining(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18182 2025-09-24 cs.CV cs.LG eess.IV 57%

AI-Derived Structural Building Intelligence for Urban Resilience: An Application in Saint Vincent and the Grenadines

Isabelle Tingzon, Yoji Toriumi, Caroline Gevaert

机构 * The World Bank Group(世界银行集团)

专题命中 预训练与数据 :foundation model(abstract);分类 cs.LG

Comments Accepted at the 2nd Workshop on Computer Vision for Developing Countries (CV4DC) at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏