arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-02 至 2025-12-02 共收录 29 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 29 篇

2409.09785 2025-12-02 cs.CL cs.AI cs.LG cs.SD eess.AS 90%

Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition

基于大语言模型的生成性错误校正:语音识别、说话人标注和情感识别的挑战与基线

Chao-Han Huck Yang, Taejin Park, Yuan Gong, Yuanchao Li, Zhehuai Chen, Yen-Ting Lin, Chen Chen, Yuchen Hu, Kunal Dhawan, Piotr Żelasko, Chao Zhang, Yun-Nung Chen, Yu Tsao, Jagadeesh Balam, Boris Ginsburg, Sabato Marco Siniscalchi, Eng Siong Chng, Peter Bell, Catherine Lai, Shinji Watanabe, Andreas Stolcke

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,comments);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出GenSEC挑战,通过大语言模型提升语音识别、说话人标注和情感识别任务的准确性与实用性。

Comments IEEE SLT 2024. The initial draft version has been done in December 2023. Post-ASR Text Processing and Understanding Community and LlaMA-7B pre-training correction model: https://huggingface.co/GenSEC-LLM/SLT-Task1-Llama2-7b-HyPo-baseline

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20804 2025-12-02 cs.CR cs.AI cs.LG 90%

AED: Automatic Discovery of Effective and Diverse Vulnerabilities for Autonomous Driving Policy with Large Language Models

AED: 利用大语言模型自动发现自主驾驶策略的有效且多样的漏洞

Le Qiu, Zelai Xu, Qixin Tan, Wenhao Tang, Chao Yu, Yu Wang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 AED 利用大语言模型自动发现自动驾驶策略的有效且多样的漏洞,提升漏洞发现的多样性和有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01748 2025-12-02 cs.LG 89%

SA-ADP: Sensitivity-Aware Adaptive Differential Privacy for Large Language Models

SA-ADP:面向大语言模型的敏感性感知自适应差分隐私

Stella Etuk, Ashraf Matrawy

机构 * School of Information Technology Carleton University(信息科技学院卡尔顿大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 SA-ADP通过根据个体PII的敏感性分配噪声,实现了大语言模型的隐私保护与效用之间的平衡。

Comments It is a 5-page paper with 5 figures and 1 Table

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19405 2025-12-02 cs.LG 89%

Learning Robust Social Strategies with Large Language Models

利用大语言模型学习稳健的社会策略

Dereck Piche, Mohammed Muqeeth, Milad Aghajohari, Juan Duque, Michael Noukhovitch, Aaron Courville

机构 * Mila University of Montreal(蒙特利尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本研究通过改进的对手学习算法,使大语言模型在多智能体协作中更稳健,提升集体收益并防止被贪婪智能体利用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20494 2025-12-02 cs.CL 88%

Adversarial Confusion Attack: Disrupting Multimodal Large Language Models

对抗混淆攻击:破坏多模态大语言模型

Jakub Hoscilowicz, Artur Janicki

机构 * Warsaw University of Technology(华沙技术大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本文提出对抗混淆攻击,通过生成扰动破坏多模态大语言模型的可靠性,展示其在不同模型上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00014 2025-12-02 cs.HC cs.AI 87%

Cultural Prompting Improves the Empathy and Cultural Responsiveness of GPT-Generated Therapy Responses

文化提示提高了GPT生成治疗回应的共情力和文化适应性

Serena Jinchen Xie, Shumenghui Zhai, Yanjing Liang, Jingyi Li, Xuehong Fan, Trevor Cohen, Weichao Yuwen

专题命中 其他LLM :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本研究通过文化提示技术提升GPT生成治疗回应的文化适应性和共情力,为多样化人群的AI治疗干预提供改进方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00293 2025-12-02 cs.LG cs.AI 86%

FiCoTS: Fine-to-Coarse LLM-Enhanced Hierarchical Cross-Modality Interaction for Time Series Forecasting

FiCoTS: 细到粗的LLM增强层次跨模态交互用于时间序列预测

Yafei Lyu, Hao Zhou, Lu Zhang, Xu Yang, Zhiyong Liu

机构 * School of Advanced Interdisciplinary Sciences, University of Chinese Academy Sciences(中国科学院大学先进交叉学科学院) MAIS, Institute of Automation, Chinese Academy of Science(中国科学院自动化研究所MAIS) Great Bay University(大亚大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 FiCoTS通过细到粗的LLM增强层次跨模态交互框架,提升多模态时间序列预测的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00656 2025-12-02 cs.CL cs.CY 86%

Sycophancy Claims about Language Models: The Missing Human-in-the-Loop

语言模型中的趋炎附势主张:缺失的人工智能循环

Jan Batzner, Volker Stocker, Stefan Schmid, Gjergji Kasneci

机构 * Weizenbaum Institute(韦岑鲍姆研究所) Technical University Berlin(柏林技术大学) Technical University Munich(慕尼黑技术大学)

专题命中 其他LLM :language model(title,abstract);LLM(abstract,comments);large language model(abstract);分类 cs.CL

AI总结 本文探讨了大型语言模型中趋炎附势现象的测量挑战,提出五个核心操作化定义,并指出当前研究缺乏对人类感知的评估,为未来研究提供建议。

Comments NeurIPS 2025 Workshop on LLM Evaluation and ICLR 2025 Workshop on Bi-Directional Human-AI Alignment

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01787 2025-12-02 eess.AS cs.AI cs.SD 85%

AHAMask: Reliable Task Specification for Large Audio Language Models without Instructions

AHAMask: 不依赖指令的大型音频语言模型可靠任务规范

Yiwei Guo, Bohan Li, Hankun Wang, Zhihan Li, Shuai Wang, Xie Chen, Kai Yu

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.AI

AI总结 AHAMask通过掩码LALMs解码器中的注意力头,实现无需指令的可靠任务规范,实验表明其性能可媲美或超越指令方法。

Comments 15 pages, 10 tables, 6 figures. This is the camera ready version for AAAI 2026, plus an appendix for supplementary experimental details and results

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01560 2025-12-02 cs.DL 85%

Estimating the prevalence of LLM-assisted text in scholarly writing

估计学者写作中LLM辅助文本的普及率

Andrew Gray

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本研究通过分析Dimensions数据库,发现LLM可能参与了2024年超过10%的已发表论文,呼吁加强披露要求以维护学术出版的完整性。

Comments 19 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01609 2025-12-02 cs.SE 82%

GPTrace: Effective Crash Deduplication Using LLM Embeddings

GPTrace: 利用LLM嵌入实现高效的崩溃去重

Patrick Herter, Vincent Ahlrichs, Ridvan Açilan, Julian Horsch

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract)

AI总结 GPTrace利用大型语言模型嵌入技术,通过计算崩溃相关数据的向量并输入聚类算法,实现高效的崩溃去重,优于传统方法。

Comments Original publication in 2026 IEEE/ACM 48th International Conference on Software Engineering (ICSE '26), April 12-18, 2026, Rio de Janeiro, Brazil. ACM, New York, NY, USA, 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13339 2025-12-02 cs.LG 79%

Statistically Accurate and Robust Generative Prediction of Rock Discontinuities with A Tabular Foundation Model

基于表格基础模型的统计准确且鲁棒的岩体结构预测

Han Meng, Gang Mei, Hong Tian, Nengxiong Xu, Jianbing Peng

机构 * School of Engineering(工程学院) Technology, China University of Geosciences (Beijing)(技术学院,中国地质大学(北京)) School of Geological Engineering(地质工程学院) Geomatics, Chang'an University(测绘学,长安大学) School of Engineering, China University of Geosciences (Wuhan)(工程学院,中国地质大学(武汉))

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

AI总结 本文提出利用表格基础模型实现对岩体结构的统计准确且鲁棒的生成预测,通过小数据学习有效捕捉复杂分布模式,实验表明优于传统方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01335 2025-12-02 cs.CR cs.AI cs.CL 79%

EmoRAG: Evaluating RAG Robustness to Symbolic Perturbations

EmoRAG:评估RAG对符号扰动的鲁棒性

Xinyun Zhou, Xinfeng Li, Yinan Peng, Ming Xu, Xuanwang Zhang, Miao Yu, Yidong Wang, Xiaojun Jia, Kun Wang, Qingsong Wen, XiaoFeng Wang, Wei Dong

机构 * ZJU Hangzhou China(浙江大学杭州校区) NTU Singapore(南洋理工大学) Hengxin Tech. Singapore(新加坡恒心科技) NUS Singapore(国立新加坡大学) NJU Nanjing China(南京大学) PKU Beijing China(北京大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 EmoRAG研究揭示RAG系统对细微表情符号扰动的鲁棒性问题,发现单个表情符号可导致检索严重误导,并提出针对性防御措施。

Comments Accepted to ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07338 2025-12-02 cs.AI cs.LG 79%

DeepPersona: A Generative Engine for Scaling Deep Synthetic Personas

DeepPersona: 一个用于扩展深度合成人设的生成引擎

Zhen Wang, Yufan Zhou, Zhongyan Luo, Lyumanshan Ye, Adam Wood, Man Yao, Saab Mansour, Luoshang Pan

机构 * UCSD(加州大学圣地亚哥分校) KU Leuven(鲁汶大学) SJTU(上海交通大学) University of Michigan(密歇根大学) Denison University(德尼森大学) Amazon(亚马逊) Meta

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 DeepPersona通过双阶段分类学引导方法生成深度合成人设,提升LLM个性化和人类模拟的准确性与多样性。

Comments add an author[Update], 12 pages, 5 figures, accepted at LAW 2025 Workshop (NeurIPS 2025) Project page: https://deeppersona-ai.github.io/

Journal ref LAW 2025 Workshop, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03657 2025-12-02 cs.CV 78%

Dynamic Multimodal Prototype Learning in Vision-Language Models

视觉-语言模型中的动态多模态原型学习

Xingyu Zhu, Shuo Wang, Beier Zhu, Miaoge Li, Yunfan Li, Junfeng Fang, Zhicai Wang, Dongsheng Wang, Hanwang Zhang

机构 * University of Science and Technology of China(中国科学技术大学) Nanyang Technological University(南洋理工大学) The Hong Kong Polytechnic University(香港理工大学) Sichuan University(四川大学) National University of Singapore(新加坡国立大学) Shenzhen University(深圳大学)

专题命中 其他LLM :language model(title,abstract)

AI总结 本文提出ProtoMM框架,通过动态更新视觉粒子和多模态原型学习,提升视觉-语言模型在测试时的适应性能。

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00004 2025-12-02 cs.IR cs.AI cs.LG 76%

Enhancing Talent Search Ranking with Role-Aware Expert Mixtures and LLM-based Fine-Grained Job Descriptions

通过角色感知专家混合与基于大语言模型的细粒度职位描述增强人才搜索排名

Jihang Li, Bing Xu, Zulong Chen, Chuanfei Xu, Minping Chen, Suyu Liu, Ying Zhou, Zeyi Wen

机构 * HKUST (GZ)(香港科技大学(广州)) Alibaba Group(阿里巴巴集团) HKUST(香港科技大学) Zhijiang Lab(浙江实验室) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室(深圳))

专题命中 其他LLM :LLM(title);分类 cs.AI、cs.LG

AI总结 本文提出基于大语言模型和角色感知专家混合的方法,提升人才搜索排名效果,通过细粒度职位描述提取和行为建模,提升CTR和CVR,实现招聘效率和成本节约。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15082 2025-12-02 eess.SY cs.SY 75%

Make Your AUV Adaptive: An Environment-Aware Reinforcement Learning Framework For Underwater Tasks

让您的水下无人航行器适应:一种环境感知的强化学习框架用于水下任务

Yimian Ding, Jingzehua Xu, Guanwen Xie, Shuai Zhang, Yi Li

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出一种环境感知的强化学习框架,通过动态捕捉流场数据和利用大语言模型优化,提升水下AUV的适应性和任务性能。

Comments This paper has been accepted by IROS 2025. Yimian Ding and Jingzehua Xu contributed equally to this work, and Jingzehua Xu is also the corresponding author of this paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00136 2025-12-02 cs.CR cs.SE 75%

An Empirical Study on the Security Vulnerabilities of GPTs

对GPTs安全漏洞的实证研究

Tong Wu, Weibin Wu, Zibin Zheng

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本研究通过实证分析揭示GPTs的安全漏洞,设计攻击套件并提出防御机制,旨在提升其安全性和应用责任感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00673 2025-12-02 cs.CL 70%

A Comparison of Human and ChatGPT Classification Performance on Complex Social Media Data

对复杂社交媒体数据中人类与ChatGPT分类性能的比较

Breanna E. Green, Ashley L. Shea, Pengfei Zhao, Drew B. Margolin

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文比较了ChatGPT和人类在复杂社交媒体数据分类中的性能,发现GPT-4在处理微妙语言时存在困难,需谨慎使用。

Comments About 15 pages, draft version of accepted conference full paper. Published paper to follow

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00392 2025-12-02 cs.CL 70%

A Taxonomy of Errors in English as she is spoke: Toward an AI-Based Method of Error Analysis for EFL Writing Instruction

英语口语中的错误分类:一种基于AI的EFL写作教学错误分析方法

Damian Heywood, Joseph Andrew Carrier, Kyu-Hong Hwang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究提出一种基于AI的英语写作错误分析系统,利用大型语言模型对写作错误进行分类和纠正,展示AI在EFL教学中的潜力。

Comments Metadata at "Replication Data for: A Taxonomy of Errors in English as she is spoke: An AI-Based System for Error Analysis for EFL Writing Instruction", https://doi.org/10.7910/DVN/N5O7C4, Harvard Dataverse, V1

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01610 2025-12-02 cs.MA 67%

Agent-Kernel: A MicroKernel Multi-Agent System Framework for Adaptive Social Simulation Powered by LLMs

Agent-Kernel: 一种基于LLM的自适应社会模拟微内核多智能体系统框架

Yuren Mao, Peigen Liu, Xinjian Wang, Rui Ding, Jing Miao, Hui Zou, Mingjie Qi, Wanxiang Luo, Longbin Lai, Kai Wang, Zhengping Qian, Peilun Yang, Yunjun Gao, Ying Zhang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Agent-Kernel是一种基于LLM的自适应社会模拟微内核多智能体系统框架,通过模块化架构提升模拟的适应性与可重用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01567 2025-12-02 eess.SP eess.IV 67%

In-Context Learning for Deep Joint Source-Channel Coding Over MIMO Channels

基于上下文学习的深度联合信源信道编码 over MIMO信道

Meng Hua, Wenjing Zhang, Chenghong Bian, Deniz Gunduz

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出基于Transformer的ICL框架,用于改进MIMO系统中图像传输的深度联合信源信道编码,通过联合学习提升编码、解码和估计性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14724 2025-12-02 cs.SI 67%

Measuring the disruptiveness of conceptual papers in the field of marketing

测量营销领域概念性论文的颠覆性

Jennifer JooYeon Lee, Hyunuk Kim

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过引用次数和颠覆评分对比,揭示概念性论文在营销领域更具颠覆性与影响力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00412 2025-12-02 cs.CV cs.LG 57%

CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion

CraftSVG: 通过布局引导扩散实现多对象文本到SVG合成

Ayan Banerjee, Nityanand Mathur, Josep Llados, Umapada Pal, Anjan Dutta

机构 * Computer Vision Center, Universitat Autònoma de Barcelona(巴塞罗那自治大学计算机视觉中心) Smallest AI Indian Statistical Institute, Kolkata(印度统计研究所) Institute for People Centred Artificial Intelligence, University of Surrey(以人为中心的人工智能研究所,萨里大学)

专题命中 其他LLM :LLM(abstract);分类 cs.LG

AI总结 SVGCraft通过布局引导扩散模型实现多对象文本到SVG合成,提升场景生成的抽象性、可识别性和细节表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00712 2025-12-02 cs.LG 57%

Exploiting Function-Family Structure in Analog Circuit Optimization

利用函数族结构进行模拟电路优化

Zhuohua Liu, Kaiqi Huang, Qinxin Mei, Yuanqi Hu, Wei W. Xing

机构 * School of Integrated Circuit Science and Engineering, Beihang University(集成电路科学与工程学院,北航) School of Mechatronic Control Engineering, Shenzhen University(机械电子控制工程学院,深大) School of Mathematical and Physical Science, University of Sheffield(数学与物理科学学院,谢菲尔德大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 本文提出电路先验网络CPN,利用预训练表格模型和直接预期改进方法,在模拟电路优化中实现高精度和高效性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00379 2025-12-02 q-bio.BM cs.LG 57%

EnzyCLIP: A Cross-Attention Dual Encoder Framework with Contrastive Learning for Predicting Enzyme Kinetic Constants

EnzyCLIP:一种基于对比学习的跨注意力双编码框架,用于预测酶动力学常数

Anas Aziz Khan, Md Shah Fahad, Priyanka, Ramesh Chandra, Guransh Singh

机构 * SCOPE Vellore Institute of Technology(维洛雷理工学院) BIT Department of Computer Science(计算机科学系) Department of Bioengineering and Biotechnology(生物工程与生物技术系)

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 EnzyCLIP通过对比学习和跨注意力机制,结合蛋白质序列和底物分子结构预测酶动力学参数,提升Kcat和Km预测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00267 2025-12-02 cs.AI 57%

Trification: A Comprehensive Tree-based Strategy Planner and Structural Verification for Fact-Checking

Trification: 一种基于树的综合策略规划与结构验证事实核查系统

Anab Maulana Barik, Shou Ziyi, Yang Kaiwen, Yang Qi, Shen Xin

机构 * Huawei Celia Team(华为Celia团队)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 Trification提出了一种基于树的综合策略规划与结构验证的事实核查框架,通过生成依赖图并动态调整验证策略,提高了事实核查的准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24888 2025-12-02 q-bio.GN 50%

Gosling Designer: a Platform to Democratize Construction and Sharing of Genomics Data Visualization Tools

Gosling Designer: 一个民主化基因组数据可视化工具构建与分享的平台

Sehi L'Yi, John Conroy, Priya Misner, David Kouřil, Astrid van den Brandt, Lisa Choy, Nezar Abdennur, Nils Gehlenborg

专题命中 其他LLM :LLM(abstract)

AI总结 Gosling Designer 是一个用于构建和分享基因组数据可视化工具的综合性平台,旨在解决现有工具在灵活性、创作难度、数据管理和协作障碍方面的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03528 2025-12-02 gr-qc 50%

Braneworld Black Bounce to Transversable Wormhole Analytically Connected to an asymptotically $AdS_5$ Boundary

分支宇宙黑洞反弹与可穿越虫洞分析地连接到渐近AdS5边界

T. M. Crispim, G. Alencar, Milko Estrada

专题命中 其他LLM :prompting(abstract)

AI总结 该研究通过分支宇宙框架,解析地连接了黑洞反弹与可穿越虫洞到渐近AdS5边界,展示了规则几何在不同黑洞类型中的应用。

Comments The authors have decided to withdraw this submission as part of an internal review process

详情

展开后加载摘要…

URL PDF HTML 收藏