arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12265 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12265 篇

2509.17965 2026-03-11 eess.AS 67%

Benchmarking Humans and Machines on Complex Multilingual Speech Understanding Tasks

在复杂多语言语音理解任务中评估人类与机器

Sai Samrat Kankanala, Ram Chandra, Sriram Ganapathy

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究在复杂多语言语音理解任务中评估人类与机器的听觉注意力能力,发现人类在母语中表现更优,而机器在单说话人条件下表现接近人类,但在双说话人设置中表现欠佳。

Comments 5 Pages, 1 Figure, 2026 IEEE International Conference on Acoustics, Speech, and Signal Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01555 2026-03-11 eess.IV cs.CV 67%

MGCR-Net:Multimodal Graph-Conditioned Vision-Language Reconstruction Network for Remote Sensing Change Detection

MGCR-Net:多模态图条件视觉-语言重建网络用于遥感变化检测

Chengming Wang, Guodong Fan, Jinjiang Li, Min Gan, C. L. Philip Chen

机构 * School of Computer Science and Technology, Shandong Technology and Business University(山东科技职业大学计算机科学与技术学院) School of Computer Science and Technology, Qingdao University(青岛大学计算机科学与技术学院) School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 MGCR-Net通过多模态图条件视觉-语言重建机制提升遥感变化检测的语义交互能力。

Journal ref IEEE Transactions on Geoscience and Remote Sensing, vol. 64, pp. 1-15, 2026, Art no. 4701515

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19112 2026-03-10 cs.CV 67%

Universal 3D Shape Matching via Coarse-to-Fine Language Guidance

通过粗到细的语言引导实现通用3D形状匹配

Qinfeng Xiao, Guofeng Mei, Bo Yang, Liying Zhang, Jian Zhang, Kit-lun Yick

机构 * Hong Kong Polytechnic University, HK SAR(香港理工大学) Fondazione Bruno Kessler, Italy(布鲁诺·凯斯勒基金会) University of Technology Sydney, Australia(悉尼科技大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 UniMatch通过粗到细的语言引导方法,实现跨类别非等距形状的通用3D匹配。

Comments Accepted by CVPR 2026

Journal ref CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07155 2026-03-10 cs.HC cs.MA 67%

NarrativeLoom: Enhancing Creative Storytelling through Multi-Persona Collaborative Improvisation

NarrativeLoom:通过多角色协作即兴创作增强创造性叙事

Yuxi Ma, Yongqian Peng, Fengyuan Yang, Siyu Zha, Chi Zhang, Zixia Jia, Zilong Zheng, Yixin Zhu

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 NarrativeLoom通过多角色协作即兴创作提升叙事原创性,通过理论指导的协作系统增强创造性输出。

Comments 19 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07134 2026-03-10 cs.HC 67%

More Than 1v1: Human-AI Alignment in Early Developmental Communities with Multimodal LLMs

多于一对一:在早期发展社区中的人工智能对齐与多模态大语言模型

Weiyan Shi, Kenny Tsu Wei Choo

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了在早期发展社区中,利用多模态大语言模型实现人机对齐的挑战,提出分层社区对齐框架,强调社区治理而非个体优化。

Comments Accepted at CHI 2026 BiAlign Workshop; OpenReview URL: https://openreview.net/forum?id=ikeH0hsBLN

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04613 2026-03-06 cs.HC 67%

Beyond Anthropomorphism: a Spectrum of Interface Metaphors for LLMs

超越拟人化:面向大语言模型的界面隐喻光谱

Jianna So, Connie Cheng, Sonia Krishna Murthy

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出将拟人化作为设计变量,通过反拟人化到超拟人化的隐喻光谱,引导界面设计从优化可用性转向鼓励批判性参与。

Comments Extended Abstracts of the 2026 CHI Conference on Human Factors in Computing Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.03175 2026-03-04 cs.CL cs.AI cs.LG 67%

Part-of-Speech Tagger for Bodo Language using Deep Learning approach

使用深度学习方法的波多语词性标注器

Dhrubajyoti Pathak, Sanjib Narzary, Sukumar Nandi, Bidisha Som

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出BodoBERT语言模型及基于深度学习的波多语词性标注模型,实现了0.8041的F1分数,并比较了阿萨姆语词性标注器。

Comments Accepted to Natural Language Engineering

Journal ref Nat. lang. process. 31 (2025) 215-229

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16114 2026-03-04 cs.SI 67%

The Content Moderator's Dilemma: Removal of Toxic Content and Distortions to Online Discourse

内容管理员的困境:有毒内容的删除与在线讨论的扭曲

Mahyar Habibi, Dirk Hovy, Carlo Schwarz

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出利用生成式大语言模型重新表述有毒推文,以保留可挽救内容并减少在线讨论的扭曲。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01505 2026-03-03 cs.RO 67%

FATE: Closed-Loop Feasibility-Aware Task Generation with Active Repair for Physically Grounded Robotic Curricula

FATE: 闭环可行性感知任务生成与主动修复的物理 grounded 机器人课程

Bingchuan Wei, Bingqi Huang, Jingheng Ma, Zeyu zhang, Sen Cui

机构 * School of Aerospace Engineering, Tsinghua University, Beijing, China(航空航天工程系,清华大学,北京,中国) Department of Automation, Tsinghua University, Beijing, China(自动化系,清华大学,北京,中国) School of Integrated Circuits, Tsinghua University, Beijing, China(集成电路学院,清华大学,北京,中国) State Key Laboratory of General Artificial Intelligence, Beijing Institute for General Artificial Intelligence (BIGAI), Beijing, China(通用人工智能国家重点实验室,北京通用人工智能研究院(BIGAI),北京,中国)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 FATE通过闭环验证与主动修复机制,生成物理 grounded 的机器人任务课程,有效减少执行失败率。

Comments 16 Pages, 4 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01319 2026-03-03 cs.HC 67%

Caught in a Mafia Romance: How Users Explore Intimate Roleplay and Narrative Exploration with Chatbots

陷入黑手党浪漫:用户如何通过聊天机器人探索亲密角色扮演与叙事探索

Julia Kieserman, Cat Mai, Sara Lignell, Lucy Qin, Athanasios Andreou, Damon McCoy, Rosanna Bellini

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究探讨用户通过聊天机器人进行亲密角色扮演和幻想探索的行为,发现用户偏好特定角色设定并对其内容的性化程度提出安全需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01314 2026-03-03 cs.HC 67%

Actor's Note: Examining the Role of AI-Generated Questions in Character Journaling for Actor Training

演员笔记:探讨AI生成问题在演员训练中的角色期刊作用

Sora Kang, Jaemin Zoh, Hyoju Kim, Hyeonseo Park, Hajin Lim, Joonhwan Lee

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Actor's Note通过AI生成问题辅助演员训练,提升角色探索与反思实践,保持艺术沉浸感。

Comments In Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems (CHI '26), April 13-17, 2026, Barcelona, Spain

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22171 2026-03-03 cs.HC 67%

A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design Ideation

早期阶段基于草图的设计构想中人类与大语言模型交互的分类

Weiyan Shi, Kenny Tsu Wei Choo

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出了一种分类方法,用于描述人类与大语言模型在早期阶段基于草图的设计构想中的交互模式,揭示了人类与AI角色的动态变化。

Comments Accepted at CHI 2026 Posters

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01022 2026-03-03 cs.CE 67%

GeoMCP: A Trustworthy Framework for AI-Assisted Analytical Geotechnical Engineering

GeoMCP:一种可信的AI辅助分析土木工程框架

Yared W. Bekele

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 GeoMCP通过将工程方法表示为结构化数据,构建了一个可信的AI辅助分析土木工程框架,确保计算透明性和安全性。

Comments 15 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05725 2026-02-27 cs.LG cs.AI cs.CL 67%

Improving Discrete Diffusion Unmasking Policies Beyond Explicit Reference Policies

超越显式参考策略的改进离散扩散解掩政策

Chunsan Hong, Seonho An, Min-Soo Kim, Jong Chul Ye

机构 * Graduate School of AI, KAIST(人工智能研究生院,韩国科学技术院) School of Computing, KAIST(计算学院,韩国科学技术院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出了一种基于学习调度器的改进离散扩散解掩策略,通过KL正则化马尔可夫决策过程优化,显著提升了在多个基准测试中的性能。

Comments Accepted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12436 2026-02-27 cs.SE 67%

Feature Request Analysis and Processing: Tasks, Techniques, and Trends

功能需求分析与处理:任务、技术与趋势

Feifei Niu, Chuanyi Li, Haosheng Zuo, Jionghan Wu, Xin Xia

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文系统分析了功能需求的研究领域,探讨了任务、技术及趋势,识别了质量保障、规范验证和基准测试等关键挑战。

Comments Accepted to: ACM Transactions on Software Engineering and Methodology

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22362 2026-02-27 cs.HC 67%

E3VA: Enhancing Emotional Expressiveness in Virtual Conversational Agents

E3VA: 提升虚拟对话代理的情感表达性

Abhishek Kulkarni, Alexander Barquero, Pavitra Lahari, Aryaan Shaikh, Sarah Brown

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 E3VA通过情感分析和自然语言处理提升虚拟对话代理的情感表达性,增强用户体验和对话质量。

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00795 2026-02-25 cs.CV 67%

DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning

DVLA-RL:基于强化学习门控的双层视觉-语言对齐用于少样本学习

Wenhao Li, Xianjing Meng, Qiangchang Wang, Zhongyi Han, Zhibin Wu, Yilong Yin

机构 * Software School, Shandong University(山东大学软件学院) Shenzhen Loop Area Institute(深圳河套学院) School of Computing and Artificial Intelligence, Shandong University of Finance and Economics(山东财经大学计算机与人工智能学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 DVLA-RL通过双层语义构建和强化学习门控注意力,实现少样本学习中视觉与语言的双层次对齐,提升泛化能力。

Comments Accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19659 2026-02-24 physics.ed-ph quant-ph 67%

Curiosity Over Hype: Modeling Motivation Language to Understand Early Outcomes in a Selective Quantum Track

好奇心胜过喧嚣:通过建模动机语言来理解选择性量子轨迹中的早期成果

Daniella Alexandra Crysti Vargas Saldana, Freddy Herrera Cueva

专题命中 其他LLM :language model(abstract);small language model(abstract)

AI总结 研究通过分析申请人的动机语言,探讨其在早期量子计算课程中的表现预测,发现好奇心相关主题与学业成绩相关,但推断测试效果有限,需进一步研究。

Comments Published in the Proceedings of IEEE ICALTER 2025. 5 pages, 7 figures

Journal ref Proceedings of the IEEE International Conference on Advanced Learning Technologies (ICALTER), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09609 2026-02-24 cs.CV 67%

Tele-Omni: a Unified Multimodal Framework for Video Generation and Editing

Tele-Omni: 一种用于视频生成与编辑的统一多模态框架

Jialun Liu, Tian Li, Xiao Cao, Yukuo Ma, Gonghu Shang, Haibin Huang, Chi Zhang, Xiangzhen Chang, Zhiyong Huang, Jiakui Hu, Zuoxin Li, Yuanzhi Liang, Cong Liu, Junqi Liu, Robby T. Tan, Haitong Tang, Qizhen Weng, Yifan Xu, Liying Yang, Xiaoyan Yang, Peng Yu, Shiwen Zhang, Xuelong Li

机构 * TeleAI

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Tele-Omni是一种统一多模态框架,通过解析文本、图像和参考视频指令,实现视频生成与编辑的灵活控制,提升时间一致性和视觉一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18549 2026-02-24 cs.HC 67%

Tower of Babel in Cross-Cultural Communication: A Case Study of #Give Me a Chinese Name# Dialogues During the "TikTok Refugees'' Event

文化沟通中的巴别塔:#Give Me a Chinese Name# 在“TikTok难民”事件中的案例研究

Jielin Feng, Zhibo Yang, Jingyi Zhao, Yujia Li, Xinwu Ye, Xingyu Lan, Siming Chen

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究通过分析TikTok难民请求中文名字的跨文化沟通事件,揭示了跨语言文化动态中的编码解码机制及影响参与度的策略。

Comments 21 pages, 6 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00920 2026-02-24 cs.SE 67%

Can Emulating Semantic Translation Help LLMs with Code Translation? A Study Based on Pseudocode

通过模拟语义翻译能否帮助LLMs进行代码翻译?基于伪代码的研究

Songqiang Chen, Congying Xu, Jingyi Chen, Jialun Cao, Jiarong Wu, Shing-Chi Cheung

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过基于伪代码的翻译方法提升LLMs在代码翻译中的表现,发现其在复杂程序处理中具有优势,但受限于伪代码的准确性。

Comments Accepted by ACM Transactions on Software Engineering and Methodology (TOSEM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18043 2026-02-23 cs.CV 67%

Spatio-temporal Decoupled Knowledge Compensator for Few-Shot Action Recognition

时空解耦的知识补偿器用于少样本动作识别

Hongyu Qu, Xiangbo Shu, Rui Yan, Hailiang Gao, Wenguan Wang, Jinhui Tang

机构 * School of Computer Science and Engineering, Nanjing University of Science and Technology(计算机科学与工程学院,南京理工大学) State Key Lab of Brain-Machine Intelligence, Zhejiang University(脑机智能国家重点实验室,浙江大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出DiST框架,通过解耦空间和时间知识,利用大语言模型学习多粒度原型,提升少样本动作识别性能。

Comments Accepted to TPAMI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17645 2026-02-20 cs.LG cs.AI cs.CL cs.CV 67%

Pushing the Frontier of Black-Box LVLM Attacks via Fine-Grained Detail Targeting

通过细粒度细节靶向推动大视觉-语言模型攻击的前沿

Xiaohan Zhao, Zhaoyi Li, Yaxin Luo, Jiacheng Cui, Zhiqiang Shen

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 通过细粒度细节靶向改进M-Attack,显著提升黑盒攻击效果,成功率达100%

Comments Code at: https://github.com/vila-lab/M-Attack-V2

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17526 2026-02-20 cs.LG cs.AI cs.CL 67%

The Anxiety of Influence: Bloom Filters in Transformer Attention Heads

影响的焦虑:变换器注意力头中的布洛姆过滤器

Peter Balogh

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究发现变换器注意力头中存在布洛姆过滤器功能,通过成员测试策略实现高效重复标记识别,且误报率随嵌入距离降低而减少。

Comments 13 pages, 8 figures, code at https://github.com/pbalogh/anxiety-of-influence v2: L3H0 reclassified as prefix-attention head following confound control. Capacity analysis updated. Duplicate-token head overlap experiment added v3: All experiments were independently validated on CPU to rule out hardware-specific computation artifacts. Results are consistent across backends

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20334 2026-02-19 cs.SE 67%

Comment Traps: How Defective Commented-out Code Augment Defects in AI-Assisted Code Generation

评论陷阱:缺陷评论代码如何增强AI辅助代码生成中的缺陷

Yuan Huang, Yukang Zhou, Xiangping Chen, Zibin Zheng

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究探讨了缺陷注释代码如何影响AI编码助手生成缺陷代码,发现其生成缺陷率高达58.17%,并指出需提升AI编码助手的鲁棒性与安全性。

Comments Accepted to the The ACM International Conference on the Foundations of Software Engineering (FSE) (FSE 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15682 2026-02-18 cs.IR 67%

The Next Paradigm Is User-Centric Agent, Not Platform-Centric Service

下一个范式是用户导向的智能体,而非平台导向的服务

Luankang Zhang, Hang Lv, Qiushi Pan, Kefen Wang, Yonghao Huang, Xinrui Miao, Yin Xu, Wei Guo, Yong Liu, Hao Wang, Enhong Chen

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文主张数字服务应从平台导向转向用户导向的智能体,通过提升隐私保护和用户控制来实现真正的用户利益。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07961 2026-02-18 stat.AP 67%

Language Markers of Emotion Flexibility Predict Depression and Anxiety Treatment Outcomes

情绪灵活性的语言标记预测抑郁症和焦虑的治疗结果

Benjamin Brindle, George A. Bonanno, Thomas Derrick Hull, Nicolas Charon, Matteo Malgaroli

专题命中 其他LLM :language model(abstract);small language model(abstract)

AI总结 研究通过分析远程治疗转录文本,发现情绪灵活性的语言标记可预测抑郁症和焦虑的治疗结果,揭示情绪动态对治疗反应的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11671 2026-02-13 cs.SE 67%

Do Not Treat Code as Natural Language: Implications for Repository-Level Code Generation and Beyond

不要将代码视为自然语言:对仓库级代码生成及更广泛领域的启示

Minh Le-Anh, Huyen Nguyen, Khanh An Tran, Nam Le Hai, Linh Ngo Van, Nghi D. Q. Bui, Bach Le

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Hydra通过结构化代码生成框架提升仓库级代码生成性能,超越现有方法并实现更高效依赖检索。

Comments Accepted to FSE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15831 2026-02-11 cs.CV 67%

UniFit: Towards Universal Virtual Try-on with MLLM-Guided Semantic Alignment

UniFit: 向基于多模态大语言模型引导的语义对齐的通用虚拟试衣迈进

Wei Zhang, Yeying Jin, Xin Li, Yan Zhang, Xiaofeng Cong, Cong Wang, Fengcai Qiao, zhichao Lian

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 UniFit通过多模态大语言模型引导的语义对齐模块,解决虚拟试衣中语义差距和数据稀缺问题,实现通用且高性能的试衣框架。

Comments accepted to AAAI-2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01918 2026-02-10 cs.HC 67%

When Workout Buddies Are Virtual: AI Agents and Human Peers in a Longitudinal Physical Activity Study

当训练伙伴是虚拟的:AI代理与人类同伴在长期身体活动研究中的应用

Alessandro Silacci, Mauro Cherubini, Arianna Boldi, Amon Rapp, Maurizio Caon

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了AI代理与人类同伴在长期身体活动中的影响,发现AI能提供更稳定的鼓励,而人类同伴则增强社会存在感,两者互补促进持续锻炼。

详情

展开后加载摘要…

URL PDF HTML 收藏