arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-14 至 2026-01-14 共收录 12 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12 篇

2601.08402 2026-01-14 cs.CL cs.AI 90%

PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors

PATS:基于大语言模型导师的个性感知教学策略

Donya Rooein, Sankalan Pal Chowdhury, Mariia Eremeeva, Yuan Qin, Debora Nozza, Mrinmaya Sachan, Dirk Hovy

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 PATS通过结合学生性格特征优化大语言模型的教学策略,提升教学效果和学生参与度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08668 2026-01-14 cs.CL 88%

Analyzing Bias in False Refusal Behavior of Large Language Models for Hate Speech Detoxification

分析大型语言模型在仇恨言论净化中的虚假拒绝行为偏见

Kyuri Im, Shuzhou Yuan, Michael Färber

机构 * TU Dresden(德累斯顿理工大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 研究分析了大型语言模型在仇恨言论净化中的虚假拒绝偏见,并提出通过中英互译策略减少此类拒绝行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13874 2026-01-14 cs.AI 88%

Geometry of Knowledge Allows Extending Diversity Boundaries of Large Language Models

知识的几何结构使大语言模型的多样性边界得以扩展

Mateusz Bystroński, Doheon Han, Nitesh V. Chawla, Tomasz Kajdanowicz

机构 * Wrocław University of Science and Technology(沃拉布大学科学与技术学院) University of Notre Dame(诺特丹大学)

专题命中 其他LLM :language model(title,abstract);large language model(title);LLM(abstract);分类 cs.AI

AI总结 基于知识的几何结构,通过流形条件调节扩展大语言模型的语义多样性边界,提升创造性发散思维。

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10808 2026-01-14 cs.CL 88%

ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios

ActiveLLM: 基于大语言模型的文本少样本场景中的主动学习

Markus Bayer, Justin Lutz, Christian Reuter

机构 * PEASEC Technical University of Darmstadt(PEASEC技术大学达姆施塔特)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 ActiveLLM利用大语言模型提升少样本场景下的分类性能,优于传统方法及ADAPET、PERFECT和SetFit等少样本学习方法。

Comments 20 pages, 10 figures, 7 tables

Journal ref Transactions of the Association for Computational Linguistics 14 (2026) 1-22

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08012 2026-01-14 cs.SE 85%

Towards Verifiably Safe Tool Use for LLM Agents

面向LLM代理工具使用的可验证安全性

Aarya Doshi, Yining Hong, Congying Xu, Eunsuk Kang, Alexandros Kapravelos, Christian Kästner

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract)

AI总结 本文提出一种基于STPA的框架,通过形式化规范数据流和工具序列,实现LLM代理的可验证安全性,减少对人工标注的依赖。

Comments 4 pages, 1 figure; accepted to ICSE NIER 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08139 2026-01-14 cs.CV cs.AI 79%

Subspace Alignment for Vision-Language Model Test-time Adaptation

子空间对齐用于视觉-语言模型测试时适应

Zhichen Zeng, Wenxuan Bao, Xiao Lin, Ruizhong Qiu, Tianxin Wei, Xuying Ning, Yuchen Yan, Chen Luo, Monica Xiao Cheng, Jingrui He, Hanghang Tong

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Amazon(亚马逊)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 SubTTA通过子空间对齐提升视觉-语言模型测试时适应性能,有效解决模态差距和视觉噪声问题。

Comments 17 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08780 2026-01-14 cs.IT eess.SP math.IT 78%

LWM-Spectro: A Foundation Model for Wireless Baseband Signal Spectrograms

LWM-Spectro:一种用于无线基带信号频谱图的基础模型

Namhyun Kim, Sadjad Alikhani, Ahmed Alkhateeb

专题命中 其他LLM :foundation model(title,abstract)

AI总结 LWM-Spectro是一种基于变换器的基础模型,通过预训练大规模I/Q数据生成时频频谱图,用于无线信号的通用表示学习,能够有效提升调制分类和SNR/移动性识别等下游任务的性能。

Comments 6 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08219 2026-01-14 cs.LG 77%

A Preliminary Agentic Framework for Matrix Deflation

一个初步的代理框架用于矩阵消去

Paimon Goulart, Evangelos E. Papalexakis

机构 * University of California, Riverside(加州大学河滨分校)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出了一种基于代理的矩阵消去方法,利用LLM和VLM实现无阈值的消去,通过上下文学习和排列优化,在不同数据集上取得有竞争力的结果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08673 2026-01-14 cs.AI cs.CY 70%

Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock

为何AI对齐失败是结构性的:学习的人类互动结构与AGI作为内生性演化冲击

Didier Sornette, Sandro Claudio Lera, Ke Wu

机构 * Institute of Risk Analysis, Prediction and Management (Risks-X)(风险分析、预测与管理研究所) Southern University of Science and Technology (SUSTech)(南方科技大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文探讨AI对齐失败的结构性问题,指出AGI作为内生性演化冲击,其风险源于放大人类智能、权力与矛盾,而非对抗意图。

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07939 2026-01-14 cs.SE cs.AI 57%

SECite: Analyzing and Summarizing Citations in Software Engineering Literature

SECite: 分析和总结软件工程文献中的引用

Shireesh Reddy Pyreddy, Khaja Valli Pathan, Hasan Masum, Tarannum Shaila Zaman

机构 * Dept. of Computer Science SUNY Polytechnic Institute(计算机科学系圣尼古拉学院) Dept. of Information Systems University of Maryland Baltimore County(信息系统系马里兰大学巴尔的摩县)

专题命中 其他LLM :LLM(abstract);分类 cs.AI

AI总结 SECite通过分析文献引用中的情感倾向,结合生成式AI生成摘要,提供了一种评估学术贡献的综合框架。

Comments Accepted at IEEE CCWC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08325 2026-01-14 cs.RO 50%

ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation

ActiveVLA: 向视觉-语言-动作模型注入主动感知以实现精确的3D机器人操控

Zhenyang Liu, Yongchong Gu, Yikai Wang, Xiangyang Xue, Yanwei Fu

机构 * Fudan University(复旦大学) Shanghai Innovation Institute(上海创新研究院) Nanyang Technological University(南洋理工大学)

专题命中 其他LLM :language model(abstract)

AI总结 ActiveVLA通过引入主动感知能力,提升机器人在复杂环境中的高精度3D操控性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07999 2026-01-14 cs.SD eess.AS 50%

VoxCog: Towards End-to-End Multilingual Cognitive Impairment Classification through Dialectal Knowledge

VoxCog: 通过方言知识实现端到端多语言认知障碍分类

Tiantian Feng, Anfeng Xu, Jinkook Lee, Shrikanth Narayanan

机构 * Ming Hsieh Department of Electrical and Computer Engineering, University of Southern California(明斯希德电气与计算机工程系,南加州大学) Dornsife College of Letters, Arts and Sciences, University of Southern California(多恩斯菲学院,南加州大学)

专题命中 其他LLM :foundation model(abstract)

AI总结 VoxCog通过整合方言知识,实现了端到端的多语言认知障碍分类,其语音基础模型在AD和MCI检测中表现出优于多模态集成和大语言模型的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏