arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-08-27 至 2025-08-27 共收录 164 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 22 篇

2406.12719 2025-08-27 cs.CL cs.AI 86%

Exploring the Robustness of Language Models for Tabular Question Answering via Attention Analysis

Kushal Raj Bhandari, Sixue Xing, Soham Dan, Jianxi Gao

机构 * Rensselaer Polytechnic Institute(拉特盖尔理工学院) Microsoft(微软)

专题命中 领域大模型 :language model(title,abstract);large language model(abstract);instruction tuning(abstract);分类 cs.CL、cs.AI

Comments Accepted TMLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18872 2025-08-27 cs.CL 85%

Empowering Computing Education Researchers Through LLM-Assisted Content Analysis

Laurie Gale, Sebastian Mateos Nicolajsen

机构 * Raspberry Pi Computing Education Research Centre University of Cambridge Cambridge UK(Raspberry Pi 计算教育研究中心 剑桥大学 剑桥 英国) Center for Computing Education Research IT University of Copenhagen Copenhagen Denmark(计算教育研究中心 丹麦技术大学 哥本哈根 丹麦) University of Cambridge(剑桥大学) IT University of Copenhagen(丹麦技术大学)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 7 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06687 2025-08-27 cs.LG cond-mat.mtrl-sci cs.AI physics.bio-ph physics.chem-ph 81%

UniGenX: a unified generative foundation model that couples sequence, structure and function to accelerate scientific design across proteins, molecules and materials

Gongbo Zhang, Yanting Li, Renqian Luo, Pipi Hu, Yang Yang, Zeru Zhao, Lingbo Li, Guoqing Liu, Zun Wang, Ran Bi, Kaiyuan Gao, Liya Guo, Yu Xie, Chang Liu, Jia Zhang, Tian Xie, Robert Pinsler, Claudio Zeni, Ziheng Lu, Hongxia Hao, Yingce Xia, Marwin Segler, Maik Riechert, Wei Yang, Hao Jiang, Wen-Bin Zhang, Zhijun Zeng, Yi Zhu, Li Dong, Xiuyuan Hu, Li Yuan, Lei Chen, Haiguang Liu, Tao Qin

机构 * Microsoft Research AI for Science(微软研究院人工智能科学部门) School of Electronic and Computer Engineering, Peking University(北京大学电子与计算机工程学院) School of AI4S, Shenzhen Graduate School, Peking University(北京大学深圳研究生院人工智能学院) Beijing Institute of Mathematical Sciences and Applications(北京数学科学研究院) School of Computer Science and Technology, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院) School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院) Department of Automation, Tsinghua University(清华大学自动化系) Yau Mathematical Sciences Center, Tsinghua University(清华大学数学科学中心)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18613 2025-08-27 eess.IV cs.LG 79%

ModAn-MulSupCon: Modality-and Anatomy-Aware Multi-Label Supervised Contrastive Pretraining for Medical Imaging

Eichi Takaya, Ryusei Inamori

机构 * AI Lab Tohoku University Hospital(东京东北大学医院人工智能实验室) Department of Diagnostic Imaging Tohoku University Graduate School of Medicine(东北大学医学研究科诊断影像科)

专题命中 领域大模型 :pretraining(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00910 2025-08-27 cs.CR cs.CL cs.LG 79%

Cyber-Zero: Training Cybersecurity Agents without Runtime

Terry Yue Zhuo, Dingmin Wang, Hantian Ding, Varun Kumar, Zijian Wang

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Public Link: https://github.com/amazon-science/cyber-zero

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18886 2025-08-27 cs.CV 78%

Toward Robust Medical Fairness: Debiased Dual-Modal Alignment via Text-Guided Attribute-Disentangled Prompt Learning for Vision-Language Models

Yuexuan Xia, Benteng Ma, Jiang He, Zhiyong Wang, Qi Dou, Yong Xia

机构 * National Engineering Laboratory for Integrated Aero-Space-Ground-Ocean Big Data Application Technology(集成空天地海大数据应用技术国家工程实验室) Northwestern Polytechnical University(西北工业大学) Huiying Medical Technology Company Ltd.(慧影医疗技术有限公司) The School of Computer Science(计算机学院) The University of Sydney(悉尼大学) Department of Computer Science and Engineering(计算机科学与工程系) The Chinese University of Hong Kong(香港中文大学)

专题命中 领域大模型 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19097 2025-08-27 cs.AI 77%

Reasoning LLMs in the Medical Domain: A Literature Survey

Armin Berger, Sarthak Khanna, David Berghaus, Rafet Sifa

机构 * Fraunhofer IAIS - Department of Media Engineering(弗劳恩霍夫研究所媒体工程部门) University of Bonn - Department of Computer Science(波恩大学计算机科学系) West-AI - Federal Ministry of Education and Research(西德人工智能 - 教育与研究部)

专题命中 领域大模型 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18406 2025-08-27 cs.MA cs.AI cs.HC 77%

Toward Generalized Autonomous Agents: A Neuro-Symbolic AI Framework for Integrating Social and Technical Support in Education

Ryan Hare, Ying Tang

机构 * Department of Electrical and Computer Engineering, Rowan University(电气与计算机工程系,罗文大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments Preprint. This work has been submitted to the IEEE for possible publication. In review for IEEE's Systems, Man, and Cybernetics Magazine. 8 pages, 3 figures. arxiv abstract has been shortened as the magazine format uses a long-form abstract

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18489 2025-08-27 cs.DC 75%

Experiences with Model Context Protocol Servers for Science and High Performance Computing

Haochen Pan, Ryan Chard, Reid Mello, Christopher Grams, Tanjin He, Alexander Brace, Owen Price Skelly, Will Engler, Hayden Holbrook, Song Young Oh, Maxime Gonthier, Michael Papka, Ben Blaiszik, Kyle Chard, Ian Foster

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

Comments 11 pages, including a 4-page appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18313 2025-08-27 cs.LG cs.AI 73%

ProtoEHR: Hierarchical Prototype Learning for EHR-based Healthcare Predictions

Zi Cai, Yu Liu, Zhiyao Luo, Tingting Zhu

机构 * University of Cambridge(剑桥大学) University of Oxford(牛津大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments CIKM 2025 Full Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19004 2025-08-27 cs.AI 70%

AI Models Exceed Individual Human Accuracy in Predicting Everyday Social Norms

Pontus Strimling, Simon Karlsson, Irina Vartanova, Kimmo Eriksson

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 18 pages + supplementy materials

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19136 2025-08-27 econ.TH 67%

Using Machine Learning to Generate, Clarify, and Improve Economic Models

Annie Liang

专题命中 领域大模型 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01240 2025-08-27 cs.CV 67%

Inspiring the Next Generation of Segment Anything Models: Comprehensively Evaluate SAM and SAM 2 with Diverse Prompts Towards Context-Dependent Concepts under Different Scenes

Xiaoqi Zhao, Youwei Pang, Shijie Chang, Yuan Zhao, Lihe Zhang, Chenyang Yu, Hanqi Liu, Jiaming Zuo, Jinsong Ouyang, Weisi Lin, Georges El Fakhri, Huchuan Lu, Xiaofeng Liu

机构 * Yale University, USA(耶鲁大学) Nanyang Technological University, Singapore(南洋理工大学) Dalian University of Technology, China(大连理工大学) X3000 Inspection Co., Ltd, China(X3000检测有限公司)

专题命中 领域大模型 :foundation model(abstract);prompting(abstract)

Comments Under submission to International Journal of Computer Vision (IJCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17661 2025-08-27 cs.AI cs.LG cs.NE 62%

Spacer: Towards Engineered Scientific Inspiration

Minhyeong Lee, Suyoung Hwang, Seunghyun Moon, Geonho Nah, Donghyun Koh, Youngjun Cho, Johyun Park, Hojin Yoo, Jiho Park, Haneul Choi, Sungbin Moon, Taehoon Hwang, Seungwon Kim, Jaeyeong Kim, Seongjun Kim, Juneau Jung

专题命中 领域大模型 :LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18687 2025-08-27 cs.CL 57%

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning

Songtao Jiang, Yuxi Chen, Sibo Song, Yan Zhang, Yeying Jin, Yang Feng, Jian Wu, Zuozhu Liu

机构 * Zhejiang University, Zhejiang, China(浙江大学) Alibaba Group, Zhejiang, China(阿里巴巴集团) Angelalign Technology Inc., Shanghai, China(Angelalign技术有限公司) ChohoTech Inc., Hangzhou, China(楚合科技有限公司) Zhejiang Key Laboratory of Medical Imaging Artificial Intelligence, Zhejiang, China(浙江省医学影像人工智能重点实验室)

专题命中 领域大模型 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18652 2025-08-27 cs.CR cs.CL 57%

UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation

Runpeng Geng, Yanting Wang, Ying Chen, Jinyuan Jia

机构 * Pennsylvania State University(宾夕法尼亚州立大学)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL

Comments 21 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11167 2025-08-27 cs.CV 50%

VFM-Guided Semi-Supervised Detection Transformer under Source-Free Constraints for Remote Sensing Object Detection

Jianhong Han, Yupei Wang, Liang Chen

机构 * School of Information and Electronics, Beijing Institute of Technology(信息与电子学院,北京理工大学) Beijing Institute of Technology Chongqing Innovation Center(北京理工大学重庆创新中心) National Key Laboratory for Space-Born Intelligent Information Processing(空间智能信息处理国家重点实验室)

专题命中 领域大模型 :foundation model(abstract)

Comments Manuscript submitted to IEEE TCSVT

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 7 篇

2508.19099 2025-08-27 cs.CL 70%

Beyond the Black Box: Integrating Lexical and Semantic Methods in Quantitative Discourse Analysis with BERTopic

Thomas Compton

机构 * University of York(约克大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 5 pages conference paper, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19096 2025-08-27 cs.AI 70%

Trustworthy Agents for Electronic Health Records through Confidence Estimation

Yongwoo Song, Minbyul Jeong, Mujeen Sung

机构 * Kyung Hee University(庆熙大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18953 2025-08-27 cs.AI 70%

Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method

I. I. Priezzhev, D. A. Danko, A. V. Shubin

机构 * National University of Oil and Gas «Gubkin University»(石油国家大学「古比金大学」) IPLab LLC(IPLab公司)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 18 pages, 6 figures. Novel hierarchical neural networks based on k-nearest neighbors method for addressing hallucination effects, training complexity, and catastrophic forgetting in modern AI systems. Includes mathematical formulations using Kohonen self-organizing maps and experimental validation on MNIST handwritten digit recognition and machine translation tasks

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19209 2025-08-27 cs.CV 67%

OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation

Jianwen Jiang, Weihong Zeng, Zerong Zheng, Jiaqi Yang, Chao Liang, Wang Liao, Han Liang, Yuan Zhang, Mingyuan Gao

机构 * Intelligent Creation Lab, ByteDance(字节跳动智能创作实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments Homepage: https://omnihuman-lab.github.io/v1_5/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19036 2025-08-27 cs.CY 67%

Of the People, By the Algorithm: How AI Transforms Democratic Representation

Yuval Rymon

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18297 2025-08-27 cs.CV cs.AI cs.CL 62%

Can VLMs Recall Factual Associations From Visual References?

Dhananjay Ashok, Ashutosh Chaubey, Hirona J. Arai, Jonathan May, Jesse Thomason

机构 * University of Southern California(南加州大学) Information Sciences Institute, University of Southern California(信息科学研究所,南加州大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

Comments To appear at EMNLP 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22675 2025-08-27 cs.CV 50%

MergeSAM: Unsupervised change detection of remote sensing images based on the Segment Anything Model

Meiqi Hu, Lingzhi Lu, Chengxi Han, Xiaoping Liu

专题命中 知识编辑与模型理解 :foundation model(abstract)

Comments 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 20 篇

2508.15648 2025-08-27 cs.CL 88%

SDGO: Self-Discrimination-Guided Optimization for Consistent Safety in Large Language Models

Peng Ding, Wen Sun, Dailin Li, Wei Zou, Jiaming Wang, Jiajun Chen, Shujian Huang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments Accepted by EMNLP 2025 (Main Conference), 15 pages, 4 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09787 2025-08-27 cs.SE cs.AI cs.HC 85%

TableTalk: Scaffolding Spreadsheet Development with a Language Agent

Jenny T. Liang, Aayush Kumar, Yasharth Bajpai, Sumit Gulwani, Vu Le, Chris Parnin, Arjun Radhakrishna, Ashish Tiwari, Emerson Murphy-Hill, Guastavo Soares

机构 * Carnegie Mellon University(卡内基梅隆大学) Microsoft(微软)

专题命中 其他LLM :language agent(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08249 2025-08-27 cs.NI cs.AI 85%

GeNet: A Multimodal LLM-Based Co-Pilot for Network Topology and Configuration

Beni Ifland, Elad Duani, Rubin Krief, Miro Ohana, Aviram Zilberman, Andres Murillo, Ofir Manor, Ortal Lavi, Hikichi Kenji, Asaf Shabtai, Yuval Elovici, Rami Puzis

机构 * 1 Ben Gurion University of the Negev, Software

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18918 2025-08-27 cs.HC cs.SD eess.AS 78%

DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality

Youngwon Choi, Donghyuk Jung, Hwayeon Kim

机构 * MAUM AI Inc.(MAUM人工智能公司) Korea Culture Technology Institute(韩国文化科技研究所)

专题命中 其他LLM :LLM(title,abstract)

Comments 2 pages, 2 figures. Accepted for presentation as a UIST 2025 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19111 2025-08-27 cs.CL 77%

Do LVLMs Know What They Know? A Systematic Study of Knowledge Boundary Perception in LVLMs

Zhikai Ding, Shiyu Ni, Keping Bi

机构 * State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院人工智能安全重点实验室) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments EMNLP2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18547 2025-08-27 cs.SE 75%

How do Humans and LLMs Process Confusing Code?

Youssef Abdelsalam, Norman Peitek, Anna-Maria Maurer, Mariya Toneva, Sven Apel

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏