arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-24 至 2025-11-24 共收录 115 信号源:cs.CL, cs.AI, cs.LG

1. 效率与部署 22 篇

2511.17155 2025-11-24 cs.CV 50%

UI-Styler: Ultrasound Image Style Transfer with Class-Aware Prompts for Cross-Device Diagnosis Using a Frozen Black-Box Inference Network

UI-Styler: 超声图像风格迁移与跨设备诊断的类感知提示使用冻结黑盒推理网络

Nhat-Tuong Do-Tran, Ngoc-Hoang-Lam Le, Ching-Chun Huang

专题命中 效率与部署 :prompting(abstract)

AI总结 UI-Styler通过类感知提示策略和模式匹配机制,提升超声图像跨设备诊断的准确性和性能。

Comments Project page: https://dotrannhattuong.github.io/UIStyler, Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16998 2025-11-24 cs.CV 50%

VLM-Augmented Degradation Modeling for Image Restoration Under Adverse Weather Conditions

增强型退化建模用于恶劣天气下的图像修复

Qianyi Shao, Yuanfan Zhang, Renxiang Xiao, Liang Hu

专题命中 效率与部署 :language model(abstract)

AI总结 本文提出了一种结合视觉-语言模型和隐式记忆库的统一模型,用于在恶劣天气下高效恢复图像,提升了修复精度和计算效率。

Journal ref Proc. 2025 30th International Conference on Automation and Computing (ICAC), pp. 1-6, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16897 2025-11-24 cs.GT 50%

Near-Optimal Dropout-Robust Sortition

近优Dropout鲁棒排序

Maya Pal Gambhir, Bailey Flanigan, Aaron Roth

专题命中 效率与部署 :prompting(abstract)

AI总结 本文提出了一种高效的损失最小化算法,用于在成员退出的情况下保持小组的代表性和适当规模,同时平衡鲁棒性、损失和公平性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16816 2025-11-24 stat.AP stat.CO 50%

Trust-Aware Multimodal Data Fusion for Yield Estimation: A Case Study of the 2020 Beirut Explosion

具有信任意识的多模态数据融合用于产量估计:贝鲁特2020爆炸的案例研究

Lekha Patel, Craig Ulmer, Stephen J. Verzi, Daniel J. Krofcheck, Indu Manickam, Asmeret Naugle, Jaideep Ray

专题命中 效率与部署 :language model(abstract)

AI总结 本文提出一种基于贝叶斯分数后验框架的多模态数据融合方法,用于估计爆炸产量,通过信任权重校准不同观测数据,提升不确定性量化和抗偏差能力。

Comments 19 pages, 4 figures, supplementary material, journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13428 2025-11-24 cs.RO 50%

VLM-SFD: VLM-Assisted Siamese Flow Diffusion Framework for Dual-Arm Cooperative Manipulation

VLM-SFD:基于视觉语言模型的双臂协作操作Siamese流扩散框架

Jiaming Chen, Yiyu Jiang, Aoshen Huang, Yang Li, Wei Pan

机构 * Department of Computer Science, The University of Manchester(计算机科学系,曼彻斯特大学) School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学)

专题命中 效率与部署 :language model(abstract)

AI总结 VLM-SFD通过双编码器-解码器架构和视觉语言模型,提升双臂协作操作的模仿学习效率与泛化能力。

Comments Accepted by IEEE RA-L

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 领域大模型 5 篇

2511.17012 2025-11-24 cs.CL cs.AI 88%

Supervised Fine Tuning of Large Language Models for Domain Specific Knowledge Graph Construction:A Case Study on Hunan's Historical Celebrities

为构建特定领域知识图谱对大型语言模型进行监督微调:以湖南历史名人案例研究

Junjie Hao, Chun Wang, Ying Qiao, Qiuyue Zuo, Qiya Song, Hua Ma, Xieping Gao

机构 * College of Information Science and Engineering(信息科学与工程学院) Hunan Provincial Key Laboratory of Philosophy and Social Sciences of Yuelushan Cultural and Digital Communication (Artificial Intelligence and International Communication AIIC)(湖南省级哲学社会科学岳麓文化与数字传播重点实验室(人工智能与国际传播AIIC)) College of Computer Science and Electronic Engineering(计算机科学与电子工程学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本研究通过监督微调提升大型语言模型在湖南历史名人领域知识提取能力,验证了参数高效方法在低资源环境下的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09613 2025-11-24 cs.CL cs.AI 88%

Task-Aligned Tool Recommendation for Large Language Models

面向大型语言模型的任务对齐工具推荐

Hang Gao, Yongfeng Zhang

机构 * Rutgers University(罗切斯特大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种以精度为导向的工具推荐方法,旨在为大型语言模型提供定制化的工具集,以提高解决复杂问题的效率。

Comments IJCNLP-AACL 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17201 2025-11-24 cs.CV 71%

Continual Alignment for SAM: Rethinking Foundation Models for Medical Image Segmentation in Continual Learning

持续对齐用于SAM:重新思考面向持续学习的医学图像分割的基础模型

Jiayi Wang, Wei Dai, Haoyu Wang, Sihan Yang, Haixia Bi, Jian Sun

机构 * Xi’an Jiaotong University(西安交通大学) School of Information and Communications Engineering, Xi’an Jiaotong University(信息与通信工程学院) School of Mathematics and Statistics, Xi’an Jiaotong University(数学与统计学学院)

专题命中 领域大模型 :foundation model(title)

AI总结 CA-SAM通过引入对齐层,实现了在持续学习中高效适应医学图像分割,提升性能并减少计算开销。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17198 2025-11-24 cs.AI cs.CV 57%

Designing Domain-Specific Agents via Hierarchical Task Abstraction Mechanism

通过层次化任务抽象机制设计领域专用代理

Kaiyu Li, Jiayu Wang, Zhi Wang, Hui Qiao, Weizhan Zhang, Deyu Meng, Xiangyong Cao

机构 * Xi’an Jiaotong University(西安交通大学) China Telecom Shaanxi Branch(中国电信陕西分公司)

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

AI总结 本文提出层次化任务抽象机制,设计领域专用代理EarthAgent,通过任务结构对齐提升复杂地理空间分析的规划能力。

Comments Page: https://earth-insights.github.io/EarthAgent

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15278 2025-11-24 cs.CV 50%

MindShot: A Few-Shot Brain Decoding Framework via Transferring Cross-Subject Prior and Distilling Frequency Domain Knowledge

MindShot: 一种通过跨受试者先验知识转移和频域知识蒸馏的少样本脑解码框架

Shuai Jiang, Zhu Meng, Haiwen Li, Delong Liu, Fei Su, Zhicheng Zhao

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Beijing Key Laboratory of Network System(北京网络系统与网络文化重点实验室)

专题命中 领域大模型 :pretraining(abstract)

AI总结 MindShot通过跨受试者先验知识转移和频域知识蒸馏,实现少样本脑解码,提升解码适应性和效率。

Comments Accepted by KBS

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 知识编辑与模型理解 7 篇

2511.17170 2025-11-24 cs.CL cs.AI 90%

Hallucinate Less by Thinking More: Aspect-Based Causal Abstention for Large Language Models

通过更多思考减少幻觉:基于方面的因果回避用于大语言模型

Vy Nguyen, Ziqi Xu, Jeffrey Chan, Estrid He, Feng Xia, Xiuzhen Zhang

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于方面的因果回避方法,通过分析LLM知识的内部多样性,实现早期回避以减少幻觉。

Comments Accepted to AAAI 2026 (Main Technical Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07318 2025-11-24 cs.CL cs.AI cs.LG 75%

When Bias Pretends to Be Truth: How Spurious Correlations Undermine Hallucination Detection in LLMs

当偏见伪装成真相:虚假相关性如何损害大语言模型中的幻觉检测

Shaowen Wang, Yiqi Dong, Ruinian Chang, Tansheng Zhu, Yuebo Sun, Kaifeng Lyu, Jian Li

机构 * Institute for Interdisciplinary Information Sciences(交叉信息学院)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文揭示了由训练数据中虚假相关性导致的幻觉问题,指出现有检测方法在面对此类幻觉时失效,并强调需新方法应对由此引发的幻觉。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17048 2025-11-24 cs.CV 71%

RoomPlanner: Explicit Layout Planner for Easier LLM-Driven 3D Room Generation

RoomPlanner: 一种显式布局规划器,用于更易由LLM驱动的3D房间生成

Wenzhuo Sun, Mingjian Liang, Wenxuan Song, Xuelian Cheng, Zongyuan Ge

专题命中 知识编辑与模型理解 :LLM(title)

AI总结 RoomPlanner通过显式布局规划和高效优化策略,实现了快速生成高质量3D室内场景。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17481 2025-11-24 cs.CV 67%

Counterfactual World Models via Digital Twin-conditioned Video Diffusion

通过数字孪生条件的视频扩散实现反事实世界模型

Yiqing Shen, Aiza Maksutova, Chenjia Li, Mathias Unberath

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 CWMDT通过构建数字孪生和应用大语言模型,实现反事实世界模型,提升对视频正向模拟的控制能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16769 2025-11-24 cs.HC 67%

Trust in AI emerges from distrust in humans: A machine learning study on decision-making guidance

对人工智能的信任源于对人类的不信任:一项关于决策指导的机器学习研究

Johan Sebastián Galindez-Acosta, Juan José Giraldo-Huertas

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 本研究通过机器学习探讨了人工智能在决策指导中的信任机制,发现对人类的不信任会促使人们转向AI,且AI在事实性情景中更受青睐。

Comments 36 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03725 2025-11-24 cs.CV 67%

Disentangled Concepts Speak Louder Than Words: Explainable Video Action Recognition

解构概念胜于言语:可解释的视频动作识别

Jongseo Lee, Wooil Lee, Gyeong-Moon Park, Seong Tae Kim, Jinwoo Choi

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 DANCE框架通过分离运动动态、物体和场景的概念类型,提升视频动作识别的可解释性与性能。

Comments NeurIPS 2025 Spotlight paper. Project page: https://jong980812.github.io/DANCE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17462 2025-11-24 q-fin.PM 50%

Scaling Conditional Autoencoders for Portfolio Optimization via Uncertainty-Aware Factor Selection

通过不确定性感知因子选择实现条件自编码器的规模扩展以用于投资组合优化

Ryan Engel, Yu Chen, Pawel Polak, Ioana Boier

专题命中 知识编辑与模型理解 :foundation model(abstract)

AI总结 本文提出通过不确定性感知因子选择扩展条件自编码器,提升投资组合优化的风险调整后绩效。

Comments 9 pages, 6 figures. Published in Proceedings of the 6th ACM International Conference on AI in Finance (ICAIF '25)

Journal ref ICAIF '25: Proceedings of the 6th ACM International Conference on AI in Finance, pages 123-131, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他LLM 8 篇

2504.08016 2025-11-24 q-bio.NC cs.AI cs.CL 90%

Emergence of psychopathological computations in large language models

大语言模型中精神病理计算的出现

Soo Yong Lee, Hyunjin Hwang, Taekwan Kim, Yuyeong Kim, Kyuri Park, Jaemin Yoo, Denny Borsboom, Kijung Shin

机构 * KAIST, Kim Jaechul Graudate School of AI(KAIST人工智能研究生院) KAIST, School of Electrical Engineering(KAIST电子工程学院) UCL, Mental Health Neuroscience Department(伦敦大学学院心理健康神经科学系) UvA, Informatics Institute(乌得勒支大学信息学院) UvA, Department of Psychology(乌得勒支大学心理学系)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本研究通过建立计算理论框架,证明大语言模型中已出现精神病理学的网络计算结构,并揭示其可能带来的安全风险。

Comments pre-print

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17301 2025-11-24 cs.CL cs.AI 87%

Large Language Models for Sentiment Analysis to Detect Social Challenges: A Use Case with South African Languages

基于大型语言模型的 sentiment 分析用于检测社会挑战:以南非语言为例的用例

Koena Ronny Mabokela, Tim Schlippe, Matthias Wölfel

机构 * University of Johannesburg, South Africa(约翰内斯堡大学) IU International University of Applied Sciences(国际应用科学大学) Karlsruhe University of Applied Sciences(卡尔斯鲁厄应用科学大学)

专题命中 其他LLM :language model(title,abstract);large language model(title);分类 cs.CL、cs.AI

AI总结 本研究利用大型语言模型对南非多种语言的社交媒体帖子进行情感分析,以检测社会挑战,发现不同模型和语言组合在情感分类上存在显著差异,融合模型能有效提升性能。

Comments Published in the Proceedings of The Southern African Conference on AI Research (SACAIR 2024), Bloemfontein, South Africa, 2-6 December 2024. ISBN: 978-0-7961-6069-0

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17315 2025-11-24 cs.CL 77%

Humanlike Multi-user Agent (HUMA): Designing a Deceptively Human AI Facilitator for Group Chats

人样多用户代理(HUMA):为群聊设计一种欺骗性人类AI助理工具

Mateusz Jacniacki, Martí Carmona Serrat

机构 * Soofte Research(Soofte研究机构)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 HUMA是一种基于LLM的多用户对话促进者,通过事件驱动架构模拟人类交互,实现在自然群聊中与人类相当的质量并难以识别。

Comments 9 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14675 2025-11-24 cs.CL 77%

LLMs as mediators: Can they diagnose conflicts accurately?

大语言模型作为调解者:它们能准确诊断冲突吗?

Özgecan Koçak, Phanish Puranam, Afşar Yegin

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文研究了大语言模型在诊断冲突根源方面的准确性,发现GPT 4在特定条件下对因果分歧的判断存在偏差,而GPT 3.5表现较弱。

Comments 27 pages, 2 appendices, 21 tables (incl appendices)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03283 2025-11-24 cs.SE cs.AI 70%

CATCODER: Repository-Level Code Generation with Relevant Code and Type Context

CATCODER: 基于相关代码和类型上下文的仓库级代码生成

Zhiyuan Pan, Xing Hu, Xin Xia, Xiaohu Yang

机构 * The State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 CatCoder通过整合相关代码和类型上下文,提升仓库级代码生成的性能和可扩展性。

Comments Revised and extended version; To be published in ACM Transactions on Software Engineering and Methodology

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17154 2025-11-24 hep-ex 67%

Proposal of an AI-Based Support Assistant for the ALICE-FIT Detector Setup at CERN

面向CERN ALICE-FIT探测器设置的基于AI的支持助手提案

Ignacy Mermer, Jakub Muszyński, Jakub Możaryn, Krystian Rosłon

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出基于AI的助手,用于支持CERN ALICE-FIT探测器操作,通过结合LLMs和RAG管道,提供上下文感知的诊断与解决方案建议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16985 2025-11-24 cs.CL 57%

ARQUSUMM: Argument-aware Quantitative Summarization of Online Conversations

ARQUSUMM: 基于论证的在线对话定量摘要

An Quang Tang, Xiuzhen Zhang, Minh Ngoc Dinh, Zhuang Li

机构 * An Quang Tang(独立研究者) Xiuzhen Zhang(独立研究者) Minh Ngoc Dinh(独立研究者) Zhuang Li(独立研究者)

专题命中 其他LLM :LLM(abstract);分类 cs.CL

AI总结 ARQUSUMM通过基于论证理论的LLM少样本学习和论证结构意识聚类算法,实现对在线对话中论点主张-理由结构的定量摘要,提升摘要的文本质量和量化准确性。

Comments Paper accepted to AAAI2026 Main Technical Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12072 2025-11-24 hep-ph 50%

High spin kaons

高自旋K介子

Ya-Rong Wang, Hao Chen, Xiao-Hai Liu, Cheng-Qun Pang

专题命中 其他LLM :prompting(abstract)

AI总结 本文基于修改后的Godfrey-Isgur模型和³P₀模型,研究了高自旋K介子的质量谱和衰变特性,并识别了可能指导未来实验研究的关键衰变通道。

Comments 13 pages, 5 figures, 9 tables. Accepted by Phys. Rev. D

Journal ref Phys. Rev. D 112, 094039 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏