arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-10-27 至 2025-10-27 共收录 214 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 22 篇

2508.05201 2025-10-27 cs.LG cs.AI cs.CL 80%

FAITH: A Framework for Assessing Intrinsic Tabular Hallucinations in Finance

Mengao Zhang, Jiayu Fu, Tanya Warrier, Yuwen Wang, Tianhui Tan, Ke-wei Huang

机构 * Asian Institute of Digital Finance, National University of Singapore(亚洲数字金融研究所,新加坡国立大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 9 pages, AMC ICAIF'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02992 2025-10-27 cs.AI cs.CL cs.LG 80%

Mitigating Manipulation and Enhancing Persuasion: A Reflective Multi-Agent Approach for Legal Argument Generation

Li Zhang, Kevin D. Ashley

机构 * Intelligent Systems Program University of Pittsburgh Pittsburgh Pennsylvania USA Intelligent Systems Program University of Pittsburgh

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 13 pages, 2 figures, 2nd ConventicLe on Artificial Intelligence Regulation and Safety Workshop at ICAIL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20967 2025-10-27 cs.CV cs.AI 79%

3DReasonKnee: Advancing Grounded Reasoning in Medical Vision Language Models

Sraavya Sambara, Sung Eun Kim, Xiaoman Zhang, Luyang Luo, Shreya Johri, Mohammed Baharoon, Du Hyun Ro, Pranav Rajpurkar

机构 * Department of Biomedical Informatics, Harvard Medical School(生物医学信息学系,哈佛医学院) Seoul National University Hospital(首尔国立大学医院)

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21228 2025-10-27 cs.CL cs.HC 77%

DispatchMAS: Fusing taxonomy and artificial intelligence agents for emergency medical services

Xiang Li, Huizi Yu, Wenkong Wang, Yiran Wu, Jiayan Zhou, Wenyue Hua, Xinxin Lin, Wenjia Tan, Lexuan Zhu, Bingyi Chen, Guang Chen, Ming-Li Chen, Yang Zhou, Zhao Li, Themistocles L. Assimes, Yongfeng Zhang, Qingyun Wu, Xin Ma, Lingyao Li, Lizhou Fan

机构 * Shandong University(山东大学) The Chinese University of Hong Kong(香港中文大学) Pennsylvania State University(宾夕法尼亚州立大学) Stanford University School of Medicine(斯坦福大学医学院) University of California(加州大学) University of Macau(澳门大学) New York University(纽约大学) Broad Institute of MIT(MITBroad研究所) The University of Hong Kong(香港大学) Chinese Academy of Medical Sciences(中国医学科学院) Peking Union Medical College(北京协和医学院) Rutgers University(罗格斯大学) University of South Florida(佛罗里达州立大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 27 pages, 7 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20279 2025-10-27 cs.LG 77%

ResearchGPT: Benchmarking and Training LLMs for End-to-End Computer Science Research Workflows

Penghao Wang, Yuhao Zhou, Mengxuan Wu, Ziheng Qin, Bangyuan Zhu, Shengbin Huang, Xuanlei Zhao, Panpan Zhang, Xiaojiang Peng, Yuzhang Shang, Jianfei Yang, Zheng Zhu, Tianlong Chen, Zhangyang Wang, Kai Wang

机构 * NUS(国立新加坡大学) NTU(国立科技大学) SZTU(深圳技术大学) UCF(佛罗里达大学) GigaAI(GigaAI研究所) UNC(北卡罗来纳大学教堂山分校) UT Austin(得克萨斯大学奥斯汀分校)

专题命中 领域大模型 :large language model(abstract);language model(abstract);pretraining(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08730 2025-10-27 cs.CL 77%

Magical: Medical Lay Language Generation via Semantic Invariance and Layperson-tailored Adaptation

Weibin Liao, Tianlong Wang, Yinghao Zhu, Yasha Wang, Junyi Gao, Liantao Ma

机构 * National Engineering Research Center For Software Engineering, Peking University(软件工程国家工程研究中心,北京大学) School of Computer Science, Peking University(北京大学计算机学院) School of Computing and Data Science, The University of Hong Kong(香港大学计算科学与数据科学学院) Centre for Medical Informatics, University of Edinburgh(爱丁堡大学医学信息中心) Health Data Research UK(英国健康数据研究)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21508 2025-10-27 cs.HC cs.CR 75%

Actionable Cybersecurity Notifications for Smart Homes: A User Study on the Role of Length and Complexity

Victor Jüttner, Charlotte S. Löffler, Erik Buchmann

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

Comments This version of the article has been accepted for publication, but is not the Version of Record and does not reflect post-acceptance improvements, or any corrections. The Version of Record is available online at: https://doi.org/10.1007/978-3-032-07989-3_19

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24841 2025-10-27 cs.CL cs.AI 73%

A Hierarchical Error Framework for Reliable Automated Coding in Communication Research: Applications to Health and Political Communication

Zhilong Zhao, Yindi Liu

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Version 2: Enhanced clarification of precision-matching task characteristics and framework applicability conditions. 20 pages, 4 figures, 4 tables. Replication package available at https://doi.org/10.7910/DVN/NDXVLZ

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20362 2025-10-27 physics.comp-ph cond-mat.mtrl-sci cs.LG 70%

ComProScanner: A multi-agent based framework for composition-property structured data extraction from scientific literature

Aritra Roy, Enrico Grisan, John Buckeridge, Chiara Gattinoni

机构 * Energy, Materials and Environment Research Centre, London South Bank University, London SE1 0AA, UK.(能源、材料与环境研究中心,伦敦南银行大学) Bioscience and Bioengineering Research Centre, London South Bank University, London SE1 0AA, UK.(生物科学与生物工程研究中心,伦敦南银行大学) Department of Physics, Kings College London, London WC2R 2LS, UK.(物理系,国王学院伦敦)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21370 2025-10-27 cs.MA cs.AI cs.CL cs.DL 62%

HIKMA: Human-Inspired Knowledge by Machine Agents through a Multi-Agent Framework for Semi-Autonomous Scientific Conferences

Zain Ul Abideen Tariq, Mahmood Al-Zubaidi, Uzair Shah, Marco Agus, Mowafa Househ

机构 * College of Science and Engineering, Hamad Bin Khalifa University, Qatar(科学与工程学院,哈马德·本·哈利法大学,卡塔尔)

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26017 2025-10-27 cs.LG cs.CL 62%

FITS: Towards an AI-Driven Fashion Information Tool for Sustainability

Daphne Theodorakopoulos, Elisabeth Eberling, Miriam Bodenheimer, Sabine Loos, Frederic Stahl

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.LG

Journal ref Frontiers in Artificial Intelligence and Applications Volume 413: ECAI 2025 (2025) 1205-1212

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21551 2025-10-27 cs.LG 57%

Interpretable Multimodal Zero-Shot ECG Diagnosis via Structured Clinical Knowledge Alignment

Jialu Tang, Hung Manh Pham, Ignace De Lathauwer, Henk S. Schipper, Yuan Lu, Dong Ma, Aaqib Saeed

机构 * Eindhoven University of Technology(埃因霍温理工大学) Singapore Management University(新加坡管理大学) Maxima Medical Center(马克斯医疗中心) Erasmus Medical Center(埃因霍温医学院)

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14193 2025-10-27 cs.GT cs.AI cs.CY 57%

Modeling the Economic Impacts of AI Openness Regulation

Tori Qiu, Benjamin Laufer, Jon Kleinberg, Hoda Heidari

机构 * Carnegie Mellon University(卡内基梅隆大学) Cornell Tech(康奈尔科技) Cornell University(康奈尔大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21209 2025-10-27 eess.AS cs.SD 50%

SpecTokenizer: A Lightweight Streaming Codec in the Compressed Spectrum Domain

Zixiang Wan, Guochang Zhang, Yifeng He, Jianqiang Wei

机构 * Audio Innovation Technology Department(音频创新技术部) School of Computer Science(计算机科学学院)

专题命中 领域大模型 :language model(abstract)

Comments Accepted by Interspeech 2025; 5 pages, 1 figure, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21000 2025-10-27 cs.CV 50%

BioDet: Boosting Industrial Object Detection with Image Preprocessing Strategies

Jiaqi Hu, Hongli Xu, Junwen Huang, Peter KT Yu, Slobodan Ilic, Benjamin Busam

机构 * Technical University of Munich(慕尼黑技术大学) XYZ Robotics(XYZ机器人)

专题命中 领域大模型 :foundation model(abstract)

Comments 8 pages, accepted by ICCV 2025 R6D

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 13 篇

2505.13763 2025-10-27 cs.AI cs.CL q-bio.NC 84%

Language Models Are Capable of Metacognitive Monitoring and Control of Their Internal Activations

Li Ji-An, Hua-Dong Xiong, Robert C. Wilson, Marcelo G. Mattar, Marcus K. Benna

机构 * Neurosciences Graduate Program University of California San Diego(加州大学圣地亚哥分校神经科学研究生项目) School of Psychology Georgia Tech(佐治亚理工学院心理学系) Department of Psychology New York University(纽约大学心理学系) Department of Neurobiology University of California San Diego(加州大学圣地亚哥分校神经生物学系)

专题命中 知识编辑与模型理解 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08966 2025-10-27 cs.CL cs.LG cs.NE 81%

Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers

Marek Kadlčík, Michal Štefánik, Timothee Mickus, Michal Spiegel, Josef Kuchař

机构 * Faculty of Informatics, Masaryk University(马萨里克大学信息学院) University of Helsinki(赫尔辛基大学) Kempelen Institute of Intelligent Technologies(凯姆佩尔智能技术研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18512 2025-10-27 cs.IR cs.AI cs.CL cs.LG 80%

AcuRank: Uncertainty-Aware Adaptive Computation for Listwise Reranking

Soyoung Yoon, Gyuwan Kim, Gyu-Hwung Cho, Seung-won Hwang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at NeurIPS 2025. The first two authors contributed equally. Author order is randomly determined via coin toss

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21664 2025-10-27 cs.CV q-bio.QM 78%

Foundation Models in Dermatopathology: Skin Tissue Classification

Riya Gupta, Yiwei Zong, Dennis H. Murphree

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19661 2025-10-27 cs.AI 77%

AgentSense: LLMs Empower Generalizable and Explainable Web-Based Participatory Urban Sensing

Xusen Guo, Mingxing Peng, Xixuan Hao, Xingchen Zou, Qiongyan Wang, Sijie Ruan, Yuxuan Liang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Beijing Institute of Technology(北京理工大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments 13 pages, 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17477 2025-10-27 cs.CL cs.AI cs.LG 75%

Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination

Jerry Huang, Prasanna Parthasarathi, Mehdi Rezagholizadeh, Boxing Chen, Sarath Chandar

机构 * Mila & Université de Montréal(Mila与蒙特利尔大学) Noah’s Ark Lab(Noah’s Ark实验室) Advanced Micro Devices Chandar Research Lab(Chandar研究实验室) Polytechnique Montréal(蒙特利尔理工学院) CIFAR AI Chair(CIFAR人工智能主席)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to Findings of The 63rd Annual Meeting of the Association for Computational Linguistics (ACL) 2025. Official proceedings version available at https://aclanthology.org/2025.findings-acl.60/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07036 2025-10-27 cs.LG cs.AI 73%

Methodological Insights into Structural Causal Modelling and Uncertainty-Aware Forecasting for Economic Indicators

Federico Cerutti

机构 * University of Brescia, Italy(意大利布雷西亚大学) Imperial College London, UK(伦敦帝国理工学院) Cardiff University, UK(卡迪夫大学) University of Southampton, UK(南安普顿大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Accepted at the 2nd edition of the Workshop in AI and Finance at ECAI-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21121 2025-10-27 cs.RO cs.AI 70%

Generalizable Hierarchical Skill Learning via Object-Centric Representation

Haibo Zhao, Yu Qi, Boce Hu, Yizhe Zhu, Ziyan Chen, Heng Tian, Xupeng Zhu, Owen Howell, Haojie Huang, Robin Walters, Dian Wang, Robert Platt

专题命中 知识编辑与模型理解 :language model(abstract);foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13737 2025-10-27 cs.AI 70%

Causal Head Gating: A Framework for Interpreting Roles of Attention Heads in Transformers

Andrew Nam, Henry Conklin, Yukang Yang, Thomas Griffiths, Jonathan Cohen, Sarah-Jane Leslie

机构 * Princeton Laboratory for AI Natural and Artificial Minds(普林斯顿人工智能实验室) Princeton University(普林斯顿大学) Department of Electrical and Computer Engineering(电气与计算机工程系) Department of Psychology(心理学系) Princeton Neuroscience Institute(普林斯顿神经科学研究所) Department of Philosophy(哲学系) Center for Statistics and Machine Learning(统计与机器学习中心)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI

Comments 10 pages, 5 figures, 2 tables. The Thirty-Ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21119 2025-10-27 stat.ME stat.ML 67%

Leveraging semantic similarity for experimentation with AI-generated treatments

Lei Shi, David Arbour, Raghavendra Addanki, Ritwik Sinha, Avi Feller

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 31 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13345 2025-10-27 cs.CY cs.AI cs.CL cs.HC cs.LG 67%

Beyond Accuracy: Rethinking Hallucination and Regulatory Response in Generative AI

Zihao Li, Weiwei Yi, Jiahong Chen

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21323 2025-10-27 cs.CV cs.LG 57%

VL-SAE: Interpreting and Enhancing Vision-Language Alignment with a Unified Concept Set

Shufan Shen, Junshu Sun, Qingming Huang, Shuhui Wang

机构 * Key Lab of Intell. Info. Process., Inst. of Comput. Tech., CAS(智能信息处理重点实验室,计算技术研究所,中国科学院) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02964 2025-10-27 cs.CV cs.LG 57%

FORLA: Federated Object-centric Representation Learning with Slot Attention

Guiqiu Liao, Matjaz Jogan, Eric Eaton, Daniel A. Hashimoto

机构 * PCASO Laboratory, Dept. of Surgery, University of Pennsylvania(宾夕法尼亚大学外科部PCASO实验室) Dept. of Computer and Information Science, University of Pennsylvania(宾夕法尼亚大学计算机与信息科学系)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

Comments Accepted by Neurips2025

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 6 篇

2510.21189 2025-10-27 cs.CR 88%

Adjacent Words, Divergent Intents: Jailbreaking Large Language Models via Task Concurrency

Yukun Jiang, Mingjie Li, Michael Backes, Yang Zhang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

Comments Accepted in NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21053 2025-10-27 cs.CR 85%

A Reinforcement Learning Framework for Robust and Secure LLM Watermarking

Li An, Yujian Liu, Yepeng Liu, Yuheng Bu, Yang Zhang, Shiyu Chang

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏