arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-08-20 至 2025-08-20 共收录 118 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 17 篇

2502.14275 2025-08-20 cs.CL cs.LG 88%

Fact or Guesswork? Evaluating Large Language Models' Medical Knowledge with Structured One-Hop Judgments

Jiaxi Li, Yiwei Wang, Kai Zhang, Yujun Cai, Bryan Hooi, Nanyun Peng, Kai-Wei Chang, Jin Lu

机构 * University of Georgia(佐治亚大学) University of California, Merced(加州大学默塞德分校) Lehigh University(莱特大学) The University of Queensland(昆士兰大学) National University of Singapore(新加坡国立大学) University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG

Comments 15 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13754 2025-08-20 cs.AI 85%

Expertise-aware Multi-LLM Recruitment and Collaboration for Medical Decision-Making

Liuxin Bao, Zhihao Peng, Xiaofei Zhou, Runmin Cong, Jiyong Zhang, Yixuan Yuan

机构 * School of Automation, Hangzhou Dianzi University(杭州电子大学自动化学院) School of Control Science and Engineering, Shandong University(山东大学控制科学与工程学院) Department of Electronic Engineering, Chinese University of Hong Kong(香港中文大学电子工程系)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13943 2025-08-20 cs.HC cs.MA 85%

LLM-Powered Virtual Patient Agents for Interactive Clinical Skills Training with Automated Feedback

Henrik Voigt, Yurina Sugamiya, Kai Lawonn, Sina Zarrieß, Atsuo Takanishi

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13240 2025-08-20 cs.CR cs.AI 83%

Quantifying Loss Aversion in Cyber Adversaries via LLM Analysis

Soham Hans, Nikolos Gurney, Stacy Marsella, Sofia Hirschmann

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13743 2025-08-20 cs.CL 81%

Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA

Kaiwei Zhang, Qi Jia, Zijian Chen, Wei Sun, Xiangyang Zhu, Chunyi Li, Dandan Zhu, Guangtao Zhai

专题命中 领域大模型 :large language model(abstract);language model(abstract);post-training(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08897 2025-08-20 cs.CL cs.AI 81%

PlantDeBERTa: An Open Source Language Model for Plant Science

Hiba Khey, Amine Lakhder, Salma Rouichi, Imane El Ghabi, Kamal Hejjaoui, Younes En-nahli, Fahd Kalloubi, Moez Amri

机构 * Mohammed VI Polytechnic University (UM6P)(摩洛哥Mohammed VI理工学院) Sidi Mohamed Ben Abdellah University (USMBA)(Sidi Mohamed Ben Abdellah大学) Faculty of Sciences Semlalia, Cadi Ayyad University(Semlalia科学学院,卡迪·艾亚德大学)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13787 2025-08-20 cs.MA cs.AI cs.NI 77%

BetaWeb: Towards a Blockchain-enabled Trustworthy Agentic Web

Zihan Guo, Yuanjian Zhou, Chenyi Wang, Linlin You, Minjie Bian, Weinan Zhang

机构 * Shanghai Innovation Institute(上海创新研究院) Sun Yat-sen University(中山大学) Zhejiang University(浙江大学) Shanghai Data Group Co., Ltd(上海数据集团有限公司) Shanghai Jiao Tong University(上海交通大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments A technical report with 21 pages, 3 figures, and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20309 2025-08-20 cs.CL 77%

Understanding the Impact of Confidence in Retrieval Augmented Generation: A Case Study in the Medical Domain

Shintaro Ozaki, Yuta Kato, Siyuan Feng, Masayo Tomita, Kazuki Hayashi, Wataru Hashimoto, Ryoma Obara, Masafumi Oyamada, Katsuhiko Hayashi, Hidetaka Kamigaito, Taro Watanabe

机构 * Nara Institute of Science and Technology(奈良科学技术研究所) The University of Tokyo(东京大学) NEC Corporation(日本电气株式会社)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted to BioNLP2025 (Workshop colocated with ACL2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12152 2025-08-20 cs.HC 75%

Contextualizing Recommendation Explanations with LLMs: A User Study

Yuanjun Feng, Stefan Feuerriegel, Yash Raj Shrestha

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

Comments Accepted to the International AAAI Conference on Web and Social Media (ICWSM 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13580 2025-08-20 cs.CL cs.AI 73%

A Comparative Study of Decoding Strategies in Medical Text Generation

Oriana Presacan, Alireza Nik, Vajira Thambawita, Bogdan Ionescu, Michael Riegler

机构 * AI Multimedia Lab(人工智能多媒体实验室) CAMPUS Research Institute(CAMPUS研究学院) National University of Science and Technology Politehnica Bucharest(巴尔的效率科学技术大学) Department of Holistic Systems(整体系统部门) SimulaMet Oslo Metropolitan University(奥斯陆 Metropolitan 大学) Cyber Security(网络安全) Simula Research Laboratory(Simula研究实验室)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00943 2025-08-20 cs.CL 70%

Universal Abstraction: Harnessing Frontier Models to Structure Real-World Data at Scale

Cliff Wong, Sam Preston, Qianchu Liu, Zelalem Gero, Jaspreet Bagga, Sheng Zhang, Shrey Jain, Theodore Zhao, Yu Gu, Yanbo Xu, Sid Kiblawi, Srinivasan Yegnasubramanian, Taxiarchis Botsis, Marvin Borja, Luis M. Ahumada, Joseph C. Murray, Guo Hui Gan, Roshanthi Weerasinghe, Kristina Young, Rom Leidner, Brian Piening, Carlo Bifulco, Tristan Naumann, Mu Wei, Hoifung Poon

机构 * Microsoft Research(微软研究院) Johns Hopkins University School of Medicine(约翰霍普金斯大学医学院) Providence Portland Medical Center(普罗维德波特医疗中心) Earle A. Chiles Research Institute(埃勒·A·奇尔斯研究所) Providence Research Network(普罗维德研究网络) Providence Genomics(普罗维德基因组学) The Oregon Clinic(俄勒冈诊所)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.07796 2025-08-20 cs.CV cs.AI cs.LG 62%

Fusing Echocardiography Images and Medical Records for Continuous Patient Stratification

Nathan Painchaud, Jérémie Stym-Popper, Pierre-Yves Courand, Nicolas Thome, Pierre-Marc Jodoin, Nicolas Duchateau, Olivier Bernard

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

Comments 13 pages + 2 pages of supplementary material, accepted for publication in IEEE TUFFC

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13518 2025-08-20 cs.CV cs.AI 57%

Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency

Yanbiao Ma, Wei Dai, Bowei Liu, Jiayi Chen, Wenke Huang, Guancheng Wan, Zhiwu Lu, Junchi Yan

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学耿丽人工智能学院) Tsinghua University(清华大学) Xidian University(西安电子科技大学) Wuhan University(武汉大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

Comments 15 pages, CVPR Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13791 2025-08-20 cs.CV 56%

Shape-from-Template with Generalised Camera

Agniva Sengupta, Stefan Zachow

机构 * Zuse Institute Berlin (ZIB)(柏林泽尼克研究所)

专题命中 领域大模型 :SFT(abstract,comments)

Comments Pre-print of the IMAVIS article: https://www.sciencedirect.com/science/article/abs/pii/S0262885625001672 Code and data in: https://git.zib.de/asengupta/sft-generalised

Journal ref Image and Vision Computing; Year: 2025; Vol.: 162; ISSN: 0262-8856; URL: https://www.sciencedirect.com/science/article/pii/S0262885625001672

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14024 2025-08-20 eess.IV cs.CV 50%

UNICON: UNIfied CONtinual Learning for Medical Foundational Models

Mohammad Areeb Qazi, Munachiso S Nwadike, Ibrahim Almakky, Mohammad Yaqub, Numan Saeed

机构 * Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 领域大模型 :foundation model(abstract)

Comments 10 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 5 篇

2411.17110 2025-08-20 cs.DB cs.LG 88%

TabulaX: Leveraging Large Language Models for Multi-Class Table Transformations

Arash Dargahi Nobari, Davood Rafiei

机构 * University of Alberta(阿尔伯塔大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19498 2025-08-20 cs.CV cs.AI 79%

Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models

Nanxing Hu, Xiaoyue Duan, Jinchao Zhang, Guoliang Kang

机构 * Beihang University(北航大学) Tencent WXG(腾讯 WXG)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13908 2025-08-20 physics.ed-ph 75%

Translating the Force Concept Inventory in the age of AI

Marina Babayeva, Justin Dunlap, Marie Snětinová, Ralf Widenhorn

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13729 2025-08-20 cs.CL cs.AI cs.LG 75%

Prediction is not Explanation: Revisiting the Explanatory Capacity of Mapping Embeddings

Hanna Herasimchyk, Alhassan Abdelhalim, Sören Laue, Michaela Regneri

机构 * Universität Hamburg(汉堡大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 10 pages, 6 Figures. Published at ECAI 2025 in a version without the Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18678 2025-08-20 cs.CV cs.RO 50%

MCN-SLAM: Multi-Agent Collaborative Neural SLAM with Hybrid Implicit Neural Scene Representation

Tianchen Deng, Guole Shen, Xun Chen, Shenghai Yuan, Hongming Shen, Guohao Peng, Zhenyu Wu, Jingchuan Wang, Lihua Xie, Danwei Wang, Hesheng Wang, Weidong Chen

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 8 篇

2503.13250 2025-08-20 cs.RO cs.HC 87%

MindEye-OmniAssist: A Gaze-Driven LLM-Enhanced Assistive Robot System for Implicit Intention Recognition and Task Execution

Zejia Zhang, Bo Yang, Xinxing Chen, Weizhuang Shi, Haoyuan Wang, Wei Luo, Jian Huang

机构 * Hubei Key Laboratory of Brain-Inspired Intelligent Systems, Huazhong University of Science and Technology(脑启发智能系统湖北省重点实验室,华中科技大学) Key Laboratory of the Ministry of Education for Image Processing and Intelligent Control, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(图像处理与智能控制教育部重点实验室,人工智能与自动化学院,华中科技大学) state key laboratory of intelligent vehicle safety technology, chongqing changan automobile co ltd(智能车辆安全技术 state key laboratory,重庆长安汽车有限公司) Science and technology innovation center, China ship development and design centre(科技创新中心,中国船舶工业设计研究中心)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13975 2025-08-20 cs.AI 85%

ChronoLLM: Customizing Language Models for Physics-Based Simulation Code Generation

Jingquan Wang, Andrew Negrut, Harry Zhang, Khailanii Slaton, Shu Wang, Radu Serban, Jinlong Wu, Dan Negrut

机构 * Department of Mechanical Engineering, University of Wisconsin-Madison(威斯康星大学麦迪逊分校机械工程系) Department of Computer Science, Rice University(里德大学计算机科学系)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13948 2025-08-20 cs.HC cs.AI cs.CL cs.PL 79%

Prompt Orchestration Markup Language

Yuge Zhang, Nan Chen, Jiahang Xu, Yuqing Yang

机构 * Microsoft Research(微软研究院)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments All findings in this paper are derived from a POML snapshot as of February 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04107 2025-08-20 cs.CV cs.AI 77%

Unlocking the Potential of MLLMs in Referring Expression Segmentation via a Light-weight Mask Decoder

Jingchao Wang, Zhijian Wu, Dingjiang Huang, Yefeng Zheng, Hong Wang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 9 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22884 2025-08-20 cs.CV 67%

AutoComPose: Automatic Generation of Pose Transition Descriptions for Composed Pose Retrieval Using Multimodal LLMs

Yi-Ting Shen, Sungmin Eum, Doheon Lee, Rohit Shete, Chiao-Yi Wang, Heesung Kwon, Shuvra S. Bhattacharyya

机构 * University of Maryland, College Park(马里兰大学学院市分校) DEVCOM Army Research Laboratory(陆军研究实验室)

专题命中 其他LLM :large language model(abstract);language model(abstract)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12837 2025-08-20 cs.LG 57%

Learning In-context n-grams with Transformers: Sub-n-grams Are Near-stationary Points

Aditya Varre, Gizem Yüce, Nicolas Flammarion

机构 * Theory of Machine Learning Lab, EPFL, Switzerland(机器学习理论实验室,瑞士联邦理工学院)

专题命中 其他LLM :language model(abstract);分类 cs.LG

Comments ICML2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02332 2025-08-20 cs.CR 50%

PII Jailbreaking in LLMs via Activation Steering Reveals Personal Information Leakage

Krishna Kanth Nakka, Xue Jiang, Dmitrii Usynin, Xuebing Zhou

专题命中 其他LLM :LLM(abstract)

Comments Preprint. V2 Updated with dataset filtering, benchmarking privacy evaluator and additional latent space visualizations

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11265 2025-08-20 physics.space-ph astro-ph.EP astro-ph.SR physics.plasm-ph 50%

Evidence of Nonlinear Signatures in Solar Wind Proton Density at the L1 Lagrange point

Dario Javier Zamora, Facundo Abaca, Bruno Zossi, Ana Georgina Elias

专题命中 其他LLM :prompting(abstract)

Journal ref A&A 700, A166 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏