arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-10 至 2025-11-10 共收录 148 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 15 篇

2508.20514 2025-11-10 cs.CL 83%

SciTopic: Enhancing Topic Discovery in Scientific Literature through Advanced LLM

Pengjiang Li, Zaitian Wang, Xinhao Zhang, Ran Zhang, Lu Jiang, Pengfei Wang, Yuanchun Zhou

机构 * Computer Network Information Center, Chinese Academy of Sciences, Beijing, China(中国科学院计算机网络信息中心) University of Chinese Academy of Sciences, Beijing, China(中国科学院大学) Department of Computer Science, Portland State University, Portland, US(波特兰州立大学计算机科学系) Information Science and Technology College, Dalian Maritime University, Dalian, China(大连海事大学信息科学与技术学院)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04270 2025-11-10 cs.CV cs.AI 83%

ZERO: Industry-ready Vision Foundation Model with Multi-modal Prompts

Sangbum Choi, Kyeongryeol Go, Taewoong Jang

机构 * Superb AI Seoul, South Korea(超霸AI首尔韩国)

专题命中 领域大模型 :foundation model(title,abstract);prompting(abstract);分类 cs.AI

Comments 9 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21700 2025-11-10 cs.CR cs.AI cs.LG 79%

XBreaking: Understanding how LLMs security alignment can be broken

Marco Arazzi, Vignesh Kumar Kembu, Antonino Nocera, Vinod P

机构 * Department of Electrical, Computer Biomedical Engineering, University of Pavia, Italy\ . Department of Computer Applications,\ University of Science \& Technology, India\ .

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05718 2025-11-10 cs.HC 78%

Do Vision-Language Models See Visualizations Like Humans? Alignment in Chart Categorization

Péter Ferenc Gyarmati, Manfred Klaffenböck, Laura Koesten, Torsten Möller

专题命中 领域大模型 :language model(title,abstract)

Comments 2 pages, 2 figures. Accepted submission to the poster track of IEEE VIS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16447 2025-11-10 cs.LG 77%

Boardwalk: Towards a Framework for Creating Board Games with LLMs

Álvaro Guglielmin Becker, Gabriel Bauer de Oliveira, Lana Bertoldo Rossato, Anderson Rocha Tavares

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Presented at SBGames 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02615 2025-11-10 cs.LG 74%

ExGra-Med: Extended Context Graph Alignment for Medical Vision-Language Models

Duy M. H. Nguyen, Nghiem T. Diep, Trung Q. Nguyen, Hoang-Bao Le, Tai Nguyen, Tien Nguyen, TrungTin Nguyen, Nhat Ho, Pengtao Xie, Roger Wattenhofer, James Zou, Daniel Sonntag, Mathias Niepert

机构 * German Research Centre for Artificial Intelligence (DFKI)(德国人工智能研究中心) Max Planck Research School for Intelligent Systems (IMPRS-IS)(马克斯·普朗克智能系统研究学校) University of Stuttgart(斯图加特大学) University Medical Center Gottingen(哥廷根大学医学中心) Max Planck Institute for Multidisciplinary Sciences(马克斯·普朗克多学科科学研究所) ARC Centre of Excellence for the Mathematical Analysis of Cellular Systems(细胞系统数学分析卓越中心) School of Mathematical Sciences, Queensland University of Technology(昆士兰科技大学数学科学学院) University of Oldenburg(奥尔登堡大学) University of Texas at Austin(德克萨斯大学奥斯汀分校) University of California San Diego(加州大学圣地亚哥分校) MBZUAI(马克斯·普朗克人工智能研究所) ETH Zurich(苏黎世联邦理工学院) Stanford University(斯坦福大学)

专题命中 领域大模型 :language model(title);分类 cs.LG

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15418 2025-11-10 cs.CL cs.AI 62%

Fine-Tuning MedGemma for Clinical Captioning to Enhance Multimodal RAG over Malaysia CPGs

Lee Qi Zun, Mohamad Zulhilmi Bin Abdul Halim, Goh Man Fye

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10751 2025-11-10 cs.LG cs.CL 62%

Neural at ArchEHR-QA 2025: Agentic Prompt Optimization for Evidence-Grounded Clinical Question Answering

Sai Prasanna Teja Reddy Bogireddy, Abrar Majeedi, Viswanatha Reddy Gajjala, Zhuoyan Xu, Siddhant Rai, Vaishnav Potlapalli

机构 * University of Chicago(芝加哥大学)

专题命中 领域大模型 :prompting(abstract);分类 cs.CL、cs.LG

Comments Accepted to Proceedings of the 24th Workshop on Biomedical Language Processing (https://aclanthology.org/2025.bionlp-share.13/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22900 2025-11-10 cs.CV cs.CL 57%

MOTOR: Multimodal Optimal Transport via Grounded Retrieval in Medical Visual Question Answering

Mai A. Shaaban, Tausifa Jan Saleem, Vijay Ram Papineni, Mohammad Yaqub

机构 * Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) Department of Mathematics and Computer Science, Faculty of Science, Alexandria University(亚历山大大学数学与计算机科学系) Sheikh Shakhbout Medical City(谢赫·沙赫布OUT医疗城)

专题命中 领域大模型 :language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08511 2025-11-10 cs.LG 57%

Distributionally robust self-supervised learning for tabular data

Shantanu Ghosh, Tiankang Xie, Mikhail Kuznetsov

机构 * Boston University(波士顿大学) Amazon(亚马逊)

专题命中 领域大模型 :language model(abstract);分类 cs.LG

Comments TRL Workshop@NeurIPS2024

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 9 篇

2507.03279 2025-11-10 cs.LG cs.AI stat.ML 90%

Conformal Information Pursuit for Interactively Guiding Large Language Models

Kwan Ho Ryan Chan, Yuyan Ge, Edgar Dobriban, Hamed Hassani, René Vidal

机构 * University of Pennsylvania(宾夕法尼亚大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07869 2025-11-10 cs.CL cs.HC 89%

Are Humans as Brittle as Large Language Models?

Jiahui Li, Sean Papay, Roman Klinger

机构 * University of Bamberg(巴姆伯格大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06265 2025-11-10 cs.LG cs.AI 88%

Large language models as uncertainty-calibrated optimizers for experimental discovery

Bojana Ranković, Ryan-Rhys Griffiths, Philippe Schwaller

机构 * École Polytechnique Fédérale de Lausanne (EPFL)(瑞士联邦理工学院(EPFL)) National Centre of Competence in Research (NCCR) Catalysis(研究型研究中心(NCCR)催化)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.20054 2025-11-10 cs.CL 74%

To Word Senses and Beyond: Inducing Concepts with Contextualized Language Models

Bastien Liétard, Pascal Denis, Mikaela Keller

机构 * University of Lille(里尔大学) Inria(法国国家信息与自动化技术研究院) CNRS(法国国家科学研究中心) Centrale Lille(里尔中央理工学院) UMR 9189 - CRIStAL

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL

Comments Published in EMNLP 2024 main conference proceedings

Journal ref In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, pages 2684-2696 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05583 2025-11-10 cs.LG cs.AI stat.ML 73%

Conformal Prediction Adaptive to Unknown Subpopulation Shifts

Nien-Shao Wang, Duygu Nur Yaldiz, Yavuz Faruk Bakman, Sai Praneeth Karimireddy

机构 * University of Southern California(南加州大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 21 pages, 7 figures, 5 tables, submitted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09360 2025-11-10 cs.CL 70%

MetaRAG: Metamorphic Testing for Hallucination Detection in RAG Systems

Channdeth Sok, David Luz, Yacine Haddam

机构 * Forvia Paris Tech Center, GIT, Immeuble Lumière, 40 avenue des Terroirs de France, 75012 Paris, France(巴黎Forvia技术中心,GIT,Lumière大厦,法国巴黎75012)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

Comments Identity-Aware AI workshop at 28th European Conference on Artificial Intelligence, October 25, 2025, Bologna, Italy

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04173 2025-11-10 cs.AI 57%

Open Agent Specification (Agent Spec): A Unified Representation for AI Agents

Soufiane Amini, Yassine Benajiba, Cesare Bernardis, Paul Cayet, Hassan Chafi, Abderrahim Fathan, Louis Faucon, Damien Hilloulin, Sungpack Hong, Ingo Kossyk, Tran Minh Son Le, Rhicheek Patra, Sujith Ravi, Jonas Schweizer, Jyotika Singh, Shailender Singh, Weiyi Sun, Kartik Talamadupula, Jerry Xu

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05432 2025-11-10 cs.CV 50%

Shared Latent Representation for Joint Text-to-Audio-Visual Synthesis

Dogucan Yaman, Seymanur Akti, Fevziye Irem Eyiokur, Alexander Waibel

机构 * Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院) Carnegie Mellon University(卡内基梅隆大学)

专题命中 知识编辑与模型理解 :pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11350 2025-11-10 cs.RO 50%

Search-TTA: A Multimodal Test-Time Adaptation Framework for Visual Search in the Wild

Derek Ming Siang Tan, Shailesh, Boyang Liu, Alok Raj, Qi Xuan Ang, Weiheng Dai, Tanishq Duhan, Jimmy Chiun, Yuhong Cao, Florian Shkurti, Guillaume Sartoretti

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Accepted for presentation at CORL 2025. Code, models, and data are available at https://search-tta.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 9 篇

2406.19544 2025-11-10 cs.SE 89%

Where Is Self-admitted Code Generated by Large Language Models on GitHub?

Xiao Yu, Lei Liu, Xing Hu, Jin Liu, Xin Xia

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06980 2025-11-10 cs.CL 88%

Exploring Multimodal Perception in Large Language Models Through Perceptual Strength Ratings

Jonghyun Lee, Dojun Park, Jiwoo Lee, Hoekeon Choi, Sung-Eun Lee

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments Published in IEEE Access

Journal ref IEEE Access, vol. 13, pp. 176751-176769, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04697 2025-11-10 cs.SI cs.AI cs.CL 79%

Simulating Misinformation Vulnerabilities With Agent Personas

David Farr, Lynnette Hui Xian Ng, Stephen Prochaska, Iain J. Cruickshank, Jevin West

机构 * School of Information Science University of Washington(信息科学学院华盛顿大学) School of Computer Science Carnegie Mellon University(计算机科学学院卡内基梅隆大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to Winter Simulation Conference 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.18469 2025-11-10 cs.CL cs.LG 79%

Iterative Self-Tuning LLMs for Enhanced Jailbreaking Capabilities

Chung-En Sun, Xiaodong Liu, Weiwei Yang, Tsui-Wei Weng, Hao Cheng, Aidan San, Michel Galley, Jianfeng Gao

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments Accepted to NAACL 2025 Main (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21067 2025-11-10 cs.SE cs.CY 75%

Designing for Novice Debuggers: A Pilot Study on an AI-Assisted Debugging Tool

Oka Kurniawan, Erick Chandra, Christopher M. Poskitt, Yannic Noller, Kenny Tsu Wei Choo, Cyrille Jegourel

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

Comments Accepted by the 25th Koli Calling International Conference on Computing Education Research (Koli Calling 2025)

Journal ref Proc. Koli Calling '25, Article No. 41, pages 1-7. ACM, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05295 2025-11-10 cs.DS cs.CL cs.DM cs.LG 73%

Language Generation and Identification From Partial Enumeration: Tight Density Bounds and Topological Characterizations

Jon Kleinberg, Fan Wei

机构 * Department of Computer Science and Information Science, Cornell University(计算机科学与信息科学系,康奈尔大学) Department of Mathemaics, Duke University(数学系,杜克大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05114 2025-11-10 cs.LG 70%

Usando LLMs para Programar Jogos de Tabuleiro e Variações

Álvaro Guglielmin Becker, Lana Bertoldo Rossato, Anderson Rocha Tavares

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted for presentation at the I Escola Regional de Aprendizado de Máquina e Inteligência Artificial da Região Sul, 2025, in Portuguese language

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04521 2025-11-10 eess.IV cs.CV 67%

Generative Autoregressive Transformers for Model-Agnostic Federated MRI Reconstruction

Valiyeh A. Nezhad, Gokberk Elmas, Bilal Kabas, Fuat Arslan, Emine U. Saritas, Tolga Çukur

机构 * Department of Electrical and Electronics Engineering(电子工程系) National Magnetic Resonance Research Center(国家磁共振研究中心) Bilkent University(比尔肯特大学)

专题命中 其他LLM :foundation model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15252 2025-11-10 cs.CR cs.CL cs.IR 57%

Retrieval-Augmented Review Generation for Poisoning Recommender Systems

Shiyi Yang, Xinshu Li, Guanglin Zhou, Chen Wang, Xiwei Xu, Liming Zhu, Lina Yao

机构 * University of New South Wales and CSIRO’s Data61(新南威尔士大学和CSIRO的Data61)

专题命中 其他LLM :foundation model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏