arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-11 至 2025-09-11 共收录 110 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12 篇

2509.08570 2025-09-11 cs.CV 71%

Vision-Language Semantic Aggregation Leveraging Foundation Model for Generalizable Medical Image Segmentation

Wenjun Yu, Yinchen Zhou, Jia-Xuan Jiang, Shubin Zeng, Yuee Li, Zhong Wang

机构 * organization= School of Information Science \& Engineering, Lanzhou University , addressline= , city= Lanzhou , postcode= 730000 , country= China

专题命中 领域大模型 :foundation model(title)

Comments 29 pages and 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08199 2025-09-11 cs.CY stat.AP 67%

Algorithmic Tradeoffs, Applied NLP, and the State-of-the-Art Fallacy

AJ Alvero, Ruohong Dong, Klint Kanopka, David Lang

专题命中 领域大模型 :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08116 2025-09-11 cs.LG cs.AI 62%

Domain Knowledge is Power: Leveraging Physiological Priors for Self Supervised Representation Learning in Electrocardiography

Nooshin Maghsoodi, Sarah Nassar, Paul F R Wilson, Minh Nguyen Nhat To, Sophia Mannina, Shamel Addas, Stephanie Sibley, David Maslove, Purang Abolmaesumi, Parvin Mousavi

机构 * School of Computing, Queen’s University(女王大学计算机学院) Vector Institute(向量研究所) Department of Electrica(电气系)

专题命中 领域大模型 :pretraining(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08820 2025-09-11 cs.RO 50%

RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation

Zongzheng Zhang, Chenghao Yue, Haobo Xu, Minwen Liao, Xianglin Qi, Huan-ang Gao, Ziwei Wang, Hao Zhao

专题命中 领域大模型 :language model(abstract)

Comments Accepted to CoRL 2025, Project Page: https://zzongzheng0918.github.io/RoboChemist.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 5 篇

2509.05230 2025-09-11 cs.CL cs.AI cs.LG 82%

CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models

Aysenur Kocak, Shuo Yang, Bardh Prenkaj, Gjergji Kasneci

机构 * Technical University of Munich(慕尼黑技术大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at the Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00300 2025-09-11 cs.HC cs.AI cs.LG 79%

MetaExplainer: A Framework to Generate Multi-Type User-Centered Explanations for AI Systems

Shruthi Chari, Oshani Seneviratne, Prithwish Chakraborty, Pablo Meyer, Deborah L. McGuinness

机构 * Rensselaer Polytechnic Institute(伦斯勒理工学院) Amazon Science(亚马逊科学) IBM Research(IBM研究院)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00074 2025-09-11 cs.CY cs.AI cs.DL cs.IR cs.SI physics.soc-ph 74%

Whose Name Comes Up? Auditing LLM-Based Scholar Recommendations

Daniele Barolo, Chiara Valentin, Fariba Karimi, Luis Galárraga, Gonzalo G. Méndez, Lisette Espín-Noboa

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.AI

Comments 40 pages: 10 main (incl. 9 figures), 3 references, and 27 appendix. Paper under-review

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08778 2025-09-11 cs.CL 57%

Do All Autoregressive Transformers Remember Facts the Same Way? A Cross-Architecture Analysis of Recall Mechanisms

Minyeong Choe, Haehyun Cho, Changho Seo, Hyunil Kim

机构 * Chosun University(全州大学) Soongsil University(顺天大学) Kongju National University(康州国立大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08640 2025-09-11 eess.IV cs.AI cs.CV 57%

RoentMod: A Synthetic Chest X-Ray Modification Model to Identify and Correct Image Interpretation Model Shortcuts

Lauren H. Cooke, Matthias Jung, Jan M. Brendel, Nora M. Kerkovits, Borek Foldyna, Michael T. Lu, Vineet K. Raghu

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI

Comments 25 + 8 pages, 4 + 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 11 篇

2505.15337 2025-09-11 cs.CL cs.AI 90%

Your Language Model Can Secretly Write Like Humans: Contrastive Paraphrase Attacks on LLM-Generated Text Detectors

Hao Fang, Jiawei Kong, Tianqu Zhuang, Yixiang Qiu, Kuofeng Gao, Bin Chen, Shu-Tao Xia, Yaowei Wang, Min Zhang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) Harbin Institute of Technology(哈尔滨工业大学) Pengcheng Laboratory(鹏城实验室)

专题命中 其他LLM :LLM(title,abstract);language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments Accepted by EMNLP-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08480 2025-09-11 cs.CL 88%

Acquiescence Bias in Large Language Models

Daniel Braun

机构 * Marburg University(马尔堡大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08218 2025-09-11 cs.CY 88%

PolicyStory: Leveraging Large Language Models to Generate Comprehensible Summaries of Policy-News in India

Aatif Nisar Dar, Aditya Raj Singh, Anirban Sen

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08395 2025-09-11 cs.CL 85%

IssueBench: Millions of Realistic Prompts for Measuring Issue Bias in LLM Writing Assistance

Paul Röttger, Musashi Hinck, Valentin Hofmann, Kobi Hackenburg, Valentina Pyatkin, Faeze Brahman, Dirk Hovy

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments accepted at TACL (pre-MIT Press publication version)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08203 2025-09-11 cs.HC cs.AI cs.SE 83%

Componentization: Decomposing Monolithic LLM Responses into Manipulable Semantic Units

Ryan Lingo, Rajeev Chhajer, Martin Arroyo, Luka Brkljacic, Ben Davis, Nithin Santhanam

机构 * Honda Research Institute, USA, Inc.(本田研究院(美国公司))

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

Comments 12 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06130 2025-09-11 cs.CV cs.CL 79%

Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models

Ce Zhang, Zifu Wan, Zhehan Kan, Martin Q. Ma, Simon Stepputtis, Deva Ramanan, Russ Salakhutdinov, Louis-Philippe Morency, Katia Sycara, Yaqi Xie

机构 * School of Computer Science, Carnegie Mellon University(计算机科学系,卡内基梅隆大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted by ICLR 2025. Project page: https://zhangce01.github.io/DeGF/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08266 2025-09-11 cs.CV 78%

Examining Vision Language Models through Multi-dimensional Experiments with Vision and Text Features

Saurav Sengupta, Nazanin Moradinasab, Jiebei Liu, Donald E. Brown

机构 * School of Data Science, University of Virginia(数据科学学院,弗吉尼亚大学)

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08400 2025-09-11 cs.NI 67%

Ubiquitous Intelligence Via Wireless Network-Driven LLMs Evolution

Xingkun Yin, Feiran You, Hongyang Du, Kaibin Huang

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05570 2025-09-11 cs.IR 67%

LESER: Learning to Expand via Search Engine-feedback Reinforcement in e-Commerce

Yipeng Zhang, Bowen Liu, Xiaoshuang Zhang, Aritra Mandal, Canran Xu, Zhe Wu

专题命中 其他LLM :LLM(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08090 2025-09-11 cs.SE 67%

ChatGPT for Code Refactoring: Analyzing Topics, Interaction, and Effective Prompts

Eman Abdullah AlOmar, Luo Xu, Sofia Martinez, Anthony Peruma, Mohamed Wiem Mkaouer, Christian D. Newman, Ali Ouni

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07998 2025-09-11 cs.CL cs.AI 62%

Bilingual Word Level Language Identification for Omotic Languages

Mesay Gemeda Yigezu, Girma Yohannis Bade, Atnafu Lambebo Tonja, Olga Kolesnikova, Grigori Sidorov, Alexander Gelbukh

机构 * Institutetext: Instituto Politécnico Nacional (IPN), Centro de Investigación en Computación (CIC), Mexico City, Mexico(墨西哥城国家理工学院(IPN)计算机研究中心)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏