arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-23 至 2025-12-23 共收录 280 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 18 篇

2412.00457 2025-12-23 cs.CL 70%

Non-native speakers of English or ChatGPT: Who thinks better?

非英语母语者或ChatGPT:谁更擅长思考?

Mohammed Q. Shormani

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究通过比较非英语母语者与ChatGPT在处理中心嵌套英语结构的能力,发现人类大脑在语言处理上仍具独特优势。

Comments 16 pages, 2 figures

Journal ref F1000Research, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19855 2025-12-23 cs.CL 70%

Artificial intelligence contribution to translation industry: looking back and forward

人工智能对翻译行业的影响:回顾与展望

Mohammed Q. Shormani, Yehia A. Al-Sohbani

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究回顾了人工智能在翻译行业45年的发展历程,分析了机器翻译、低资源语言等热点问题,并指出需进一步研究以解决多方言和文化语域等挑战。

Comments 30 pages, 13 figures

Journal ref Discover Artificial Intelligence, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18445 2025-12-23 cs.LG 70%

On the Universality of Transformer Architectures; How Much Attention Is Enough?

Transformer架构的通用性;注意力是否足够?

Amirreza Abbasi, Mohsen Hooshmand

机构 * Department of Computer Science and Information Technology(计算机科学与信息技术系) Institute for Advanced Studies in Basic Sciences (IASBS)(基础科学高级研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文探讨Transformer架构的通用性问题,分析其表达能力及未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18362 2025-12-23 cs.CL 70%

SRS-Stories: Vocabulary-constrained multilingual story generation for language learning

SRS-Stories: 词汇受限的多语言故事生成用于语言学习

Wiktor Kamzela, Mateusz Lango, Ondrej Dusek

机构 * Poznan University of Technology, Institute of Computer Science(波兹南技术大学计算机科学学院) Charles University, Faculty of Mathematics and Physics(查尔斯大学数学与物理系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SRS-Stories通过间隔重复系统生成多语言故事,利用已知词汇教授新词并复习旧词,提升语言学习效果。

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12880 2025-12-23 cs.CR 67%

Universal Jailbreak Suffixes Are Strong Attention Hijackers

通用劫持后缀是强大的注意力劫持者

Matan Ben-Tov, Mor Geva, Mahmood Sharif

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了基于后缀的劫持攻击,发现通用后缀能更有效地劫持LLM的上下文化过程,并提出通过提升后缀普遍性来增强攻击效果,同时提供缓解方法。

Comments Accepted at TACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19020 2025-12-23 cs.CV cs.LG 57%

CETCAM: Camera-Controllable Video Generation via Consistent and Extensible Tokenization

CETCAM: 通过一致且可扩展的分词实现相机可控的视频生成

Zelin Zhao, Xinyu Gong, Bangya Liu, Ziyang Song, Jun Zhang, Suhui Wu, Yongxin Chen, Hao Zhang

机构 * Georgia Institute of Technology(佐治亚理工学院) ByteDance(字节跳动) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) The Hong Kong Polytechnic University(香港理工大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 CETCAM通过一致且可扩展的分词方案实现相机可控的视频生成,提高了几何一致性和视觉真实性,同时支持额外的控制模式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18996 2025-12-23 cs.IR cs.SE 50%

Modular Layout Synthesis (MLS): Front-end Code via Structure Normalization and Constrained Generation

模块化布局合成(MLS):通过结构规范化和约束生成实现前端代码

Chong Liu, Ming Zhang, Fei Li, Hao Zhou, Xiaoshuang Chen, Ye Yuan

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出模块化布局合成(MLS),通过结构规范化和约束生成,提升前端代码的模块化和重用性,适用于多种现代开发框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18813 2025-12-23 cs.CV 50%

Revealing Perception and Generation Dynamics in LVLMs: Mitigating Hallucinations via Validated Dominance Correction

揭示LVLMs中的感知与生成动态:通过验证主导修正缓解幻觉

Guangtao Lyu, Xinyi Cheng, Chenghao Xu, Qi Liu, Muli Yang, Fen Fang, Huilin Chen, Jiexi Yan, Xu Yang, Cheng Deng

机构 * School of Electronic Engineering, Xidian University(电子科技大学电子工程学院) School of Computer Science and Technology, Xidian University(电子科技大学计算机科学与技术学院) Hohai university(河海大学) Institute for Infocomm Research (I 2 R), A*STAR(信息通信研究所(I2R)) School of Foreign Languages, Xidian University(电子科技大学外国语言学院)

专题命中 其他LLM :language model(abstract)

AI总结 本文提出VDC策略,通过验证主导修正缓解LVLMs中的幻觉问题,揭示感知与生成的动态机制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18772 2025-12-23 cs.CV 50%

In-Context Audio Control of Video Diffusion Transformers

上下文音频控制视频扩散变换器

Wenze Liu, Weicai Ye, Minghong Cai, Quande Liu, Xintao Wang, Xiangyu Yue

机构 * MMLab, The Chinese University of Hong Kong(香港中文大学MMLab) Kling Team, Kuaishou Technology(快手科技Kling团队)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文提出ICAC框架,通过三维注意力机制实现音频驱动的视频生成,解决时间同步信号在视频扩散模型中的整合问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06716 2025-12-23 cs.SE cs.PF 50%

Efficiently Ranking Software Variants with Minimal Benchmarks

通过最小基准高效排序软件变体

Théo Matricon, Mathieu Acher, Helge Spieker, Arnaud Gotlieb

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出BISS方法,通过测试套件优化技术减少基准测试成本,同时保持变体排名的稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏