arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-09-04 至 2025-09-04 共收录 109 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 10 篇

2509.03161 2025-09-04 cs.CL cs.AI 73%

Domain Adaptation of LLMs for Process Data

Rafael Seidi Oyamada, Jari Peeperkorn, Jochen De Weerdt, Johannes De Smedt

机构 * Research Centre for Information Systems Engineering(信息系统工程研究中心)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19200 2025-09-04 cs.AI cs.CL 73%

The Ramon Llull's Thinking Machine for Automated Ideation

Xinran Zhao, Boyuan Zheng, Chenglei Si, Haofei Yu, Ken Liu, Runlong Zhou, Ruochen Li, Tong Chen, Xiang Li, Yiming Zhang, Tongshuang Wu

机构 * CMU(卡内基梅隆大学) OSU(俄亥俄州立大学) Stanford(斯坦福大学) UIUC(伊利诺伊大学香槟分校) UT Dallas(德克萨斯大学达拉斯分校)

专题命中 领域大模型 :LLM(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments 21 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19608 2025-09-04 cs.AI cs.LG 62%

ChordPrompt: Orchestrating Cross-Modal Prompt Synergy for Multi-Domain Incremental Learning in CLIP

Zhiyuan Wang, Bokui Chen

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University, China(清华大学深圳国际研究生院,清华大学,中国)

专题命中 领域大模型 :language model(abstract);分类 cs.AI、cs.LG

Comments Accepted by the European Conference on Machine Learning and Principles and Practice of Knowledge Discovery in Databases (ECML-PKDD 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02942 2025-09-04 cs.IR cs.LG 57%

RankGraph: Unified Heterogeneous Graph Learning for Cross-Domain Recommendation

Renzhi Wu, Junjie Yang, Li Chen, Hong Li, Li Yu, Hong Yan

机构 * Meta MRS

专题命中 领域大模型 :foundation model(abstract);分类 cs.LG

Comments RecSys 2025

Journal ref RecSys 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12278 2025-09-04 cs.CV 50%

Towards a Universal Synthetic Video Detector: From Face or Background Manipulations to Fully AI-Generated Content

Rohit Kundu, Hao Xiong, Vishal Mohanty, Athula Balachandran, Amit K. Roy-Chowdhury

机构 * Google, Mountain View, USA(谷歌(Mountain View, USA)) University of California, Riverside(加州大学河滨分校)

专题命中 领域大模型 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 7 篇

2501.09997 2025-09-04 cs.CL cs.AI 90%

Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models

Qiang Liu, Xinlong Chen, Yue Ding, Bowen Song, Weiqiang Wang, Shu Wu, Liang Wang

机构 * New Laboratory of Pattern Recognition (NLPR), State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences (CASIA)(模式识别新实验室、多模态人工智能系统国家重点实验室、自动化研究所、中国科学院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02805 2025-09-04 cs.LG 79%

Challenges in Understanding Modality Conflict in Vision-Language Models

Trang Nguyen, Jackson Michaels, Madalina Fiterau, David Jensen

机构 * Manning College of Information \& Computer Sciences, University of Massachusetts Amherst, Amherst, U.S.

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03518 2025-09-04 cs.LG 78%

Can LLMs Lie? Investigation beyond Hallucination

Haoran Huan, Mihir Prabhudesai, Mengning Wu, Shantanu Jaiswal, Deepak Pathak

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 知识编辑与模型理解 :LLM(abstract,comments);large language model(abstract);language model(abstract);分类 cs.LG

Comments Website at https://llm-liar.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02170 2025-09-04 cs.CL cs.AI 73%

Avoidance Decoding for Diverse Multi-Branch Story Generation

Kyeongman Park, Nakyeong Yang, Kyomin Jung

机构 * Seoul National University(首尔国立大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02879 2025-09-04 econ.TH 67%

Artificial or Human Intelligence?

Eric Gao

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.06509 2025-09-04 cs.CV cs.AI cs.LG 62%

Aligning Machine and Human Visual Representations across Abstraction Levels

Lukas Muttenthaler, Klaus Greff, Frieda Born, Bernhard Spitzer, Simon Kornblith, Michael C. Mozer, Klaus-Robert Müller, Thomas Unterthiner, Andrew K. Lampinen

机构 * Google DeepMind Machine Learning Group(谷歌DeepMind机器学习组) Technische Universität Berlin(技术大学柏林) BIFOLD Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究所) Max Planck Institute for Human Cognitive and Brain Sciences(人类认知与脑科学Max Planck研究所) Max Planck Institute for Human Development(人类发展Max Planck研究所) TUD Dresden University of Technology(德累斯顿技术大学) Anthropic Department of Artificial Intelligence, Korea University(人工智能系,韩国大学) Max Planck Institute for Informatics(信息Max Planck研究所)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

Comments 91 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20757 2025-09-04 cs.CL 57%

GUARD: Glocal Uncertainty-Aware Robust Decoding for Effective and Efficient Open-Ended Text Generation

Yuanhao Ding, Esteban Garces Arias, Meimingwei Li, Julian Rodemann, Matthias Aßenmacher, Danlu Chen, Gaojuan Fan, Christian Heumann, Chongsheng Zhang

机构 * Henan University(河南大学) Department of Statistics, LMU Munich(慕尼黑大学统计系) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) CISPA Helmholtz Center for Information Security, Saarbrücken(萨尔布吕肯亥姆霍尔兹信息安全中心) University of California, San Diego(加州大学圣地亚哥分校)

专题命中 知识编辑与模型理解 :LLM(abstract);分类 cs.CL

Comments Accepted at Findings of the Association for Computational Linguistics: EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 7 篇

2509.02638 2025-09-04 cs.CY 88%

Exploring the interplay between Planetary Boundaries and Sustainable Development Goals using Large Language Models

Lamyae Rhomrasi, Pilar Manchón, Ricardo Vinuesa, Francesco Fuso-Nerini, J. Alberto Conejero, Javier García-Martínez, Sergio Hoyas

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15048 2025-09-04 cs.AI 77%

MorphAgent: Empowering Agents through Self-Evolving Profiles and Decentralized Collaboration

Siyuan Lu, Jiaqi Shao, Bing Luo, Tao Lin

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01684 2025-09-04 cs.LG cs.AI 73%

Reinforcement Learning for Machine Learning Engineering Agents

Sherry Yang, Joy He-Yueya, Percy Liang

机构 * Stanford University(斯坦福大学)

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.06336 2025-09-04 physics.acc-ph cs.AI 70%

Towards Agentic AI on Particle Accelerators

Antonin Sulc, Thorsten Hellert, Raimund Kammering, Hayden Hoschouer, Jason St. John

机构 * Helmholtz Zentrum Berlin(柏林海德堡中心) LBNL(劳伦斯伯克利国家实验室) DESY(德意志电子同步辐射研究中心) FNAL(费米国家加速器实验室)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

Comments 5 pages, 3 figures, Machine Learning and the Physical Sciences at Workshop at the 38th conference on Neural Information Processing Systems (NeurIPS)

Journal ref Machine Learning and the Physical Sciences Workshop at the 38th conference on Neural Information Processing Systems (NeurIPS) December 15, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03212 2025-09-04 cs.CV 67%

AIVA: An AI-based Virtual Companion for Emotion-aware Interaction

Chenxi Li

专题命中 其他LLM :large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11128 2025-09-04 cs.CL cs.SD eess.AS 57%

FELLE: Autoregressive Speech Synthesis with Token-Wise Coarse-to-Fine Flow Matching

Hui Wang, Shujie Liu, Lingwei Meng, Jinyu Li, Yifan Yang, Shiwan Zhao, Haiyang Sun, Yanqing Liu, Haoqin Sun, Jiaming Zhou, Yan Lu, Yong Qin

机构 * College of Computer Science, Nankai University(南开大学计算机科学学院) Microsoft Corporation(微软公司)

专题命中 其他LLM :language model(abstract);分类 cs.CL

Comments Accepted by ACM Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01030 2025-09-04 cs.IR 50%

Identifying Origins of Place Names via Retrieval Augmented Generation

Alexis Horde-Vo, Matt Duckham, Estrid He, Rafe Benli

专题命中 其他LLM :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏