arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-23 至 2025-12-23 共收录 18 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 18 篇

2505.18947 2025-12-23 cs.CV 88%

OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model

OpenHOI: 基于多模态大语言模型的开放世界手-物体交互合成

Zhenhao Zhang, Ye Shi, Lingxiao Yang, Suting Ni, Qi Ye, Jingya Wang

机构 * ShanghaiTech University(上海科技大学) Zhejiang University(浙江大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

AI总结 OpenHOI通过多模态大语言模型实现开放世界手-物体交互合成,能生成长周期操控序列并处理复杂语言指令。

Comments Accepted by NeurIPS 2025 as Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19378 2025-12-23 cs.CL 86%

HATS: High-Accuracy Triple-Set Watermarking for Large Language Models

HATS: 高精度三集合水印技术用于大语言模型

Zhiqing Hu, Chenxu Zhao, Jiazhong Lu, Xiaolei Liu

机构 * Institute of Computer Application, China Academy of Engineering Physics, Mianyang, China(计算机应用研究所,工程物理科学院,绵阳,中国) National Interdisciplinary Research Center of Engineering Physics, Mianyang, China(工程物理跨学科研究中心,绵阳,中国) School of Cybersecurity (Xin Gu Industrial College), Chengdu University of Information Technology, Mianyang, China(网络安全学院(新谷工业学院),信息科技大学,绵阳,中国)

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract);分类 cs.CL

AI总结 HATS通过三集合划分和统计检测技术,实现大语言模型的高精度水印检测,同时保持文本可读性。

Comments Camera-ready version of the paper accepted for oral presentation at the 11th International Conference on Computer and Communications (ICCC 2025)

Journal ref In Proceedings of the 11th International Conference on Computer and Communications, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12715 2025-12-23 cs.CV cs.RO 78%

AsyMoE: Leveraging Modal Asymmetry for Enhanced Expert Specialization in Large Vision-Language Models

AsyMoE:利用模态不对称性提升大视觉-语言模型专家专业化

Heng Zhang, Haichuan Hu, Yaomin Shen, Weihao Yu, Yilei Yuan, Haochen You, Guo Cheng, Zijian Zhang, Lubin Gan, Huihui Wei, Hao Zhang, Jin Huang

专题命中 其他LLM :language model(title,abstract)

AI总结 AsyMoE通过三个专门专家组解决视觉-语言模态不对称问题,提升大模型专家专业化性能。

Comments This submission has been withdrawn by the authors due to a fundamental error in the methodology that affects the validity of the main results

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23941 2025-12-23 cs.LG q-bio.NC 77%

Brain-language fusion enables interactive neural readout and in-silico experimentation

脑语言融合实现交互式神经读取和计算机仿真实验

Victoria Bosch, Daniel Anthes, Adrien Doerig, Sushrut Thorat, Peter König, Tim Christian Kietzmann

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 CorText通过将神经活动嵌入LLM潜在空间,实现脑数据与自然语言的交互式解码和跨类别泛化,推动脑活动与语言之间的生成性接口发展。

Comments v2

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18189 2025-12-23 cs.AI cs.SC 77%

NL2CA: Auto-formalizing Cognitive Decision-Making from Natural Language Using an Unsupervised CriticNL2LTL Framework

NL2CA: 一种基于无监督批评树的自然语言自动形式化认知决策方法

Zihao Deng, Yijia Li, Renrui Zhang, Peijun Ye

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 NL2CA通过自动形式化自然语言描述的认知决策规则,实现无监督学习下的认知代理构建与优化,提升认知建模的可解释性和可扩展性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13901 2025-12-23 cs.CL 77%

RAID: Refusal-Aware and Integrated Decoding for Jailbreaking LLMs

RAID:针对对抗解码的拒绝意识与集成解码用于对抗大型语言模型

Tuan T. Nguyen, John Le, Thai T. Vu, Willy Susilo, Heath Cooper

机构 * University of Wollongong(沃林根大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 RAID通过嵌入空间正则化和拒绝意识正则化器,有效提升对抗解码攻击的成功率,同时降低计算成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20277 2025-12-23 cs.LG cs.AI 73%

HVAdam: A Full-Dimension Adaptive Optimizer

HVAdam:一种全维度自适应优化器

Yiheng Zhang, Shaowu Wu, Yuanzhuo Xu, Jiajun Wu, Shang Xu, Steve Drew, Xiaoguang Niu

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 HVAdam提出一种具有连续可调自适应性的新型优化器,通过增量延迟更新机制提升收敛性与鲁棒性,在多种任务中超越现有优化器。

Comments Accepted at AAAI2025

Journal ref In Proceedings of the AAAI Conference on Artificial Intelligence, volume 39, pp. 22623-22631, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03476 2025-12-23 cs.CL 70%

Delta-KNN: Improving Demonstration Selection in In-Context Learning for Alzheimer's Disease Detection

Delta-KNN: 提高阿尔茨海默病检测中基于上下文学习的演示选择

Chuyuan Li, Raymond Li, Thalia S. Field, Giuseppe Carenini

机构 * Department of Computer Science(计算机科学系) Vancouver Stroke Program and Division of Neurology, Faculty of Medicine(温哥华卒中计划和神经病学系,医学院) The University of British Columbia(不列颠哥伦比亚大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 Delta-KNN通过Delta评分和KNN检索器提升基于上下文学习的阿尔茨海默病检测性能,实现新状态的最先进结果。

Journal ref In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00457 2025-12-23 cs.CL 70%

Non-native speakers of English or ChatGPT: Who thinks better?

非英语母语者或ChatGPT:谁更擅长思考?

Mohammed Q. Shormani

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究通过比较非英语母语者与ChatGPT在处理中心嵌套英语结构的能力,发现人类大脑在语言处理上仍具独特优势。

Comments 16 pages, 2 figures

Journal ref F1000Research, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19855 2025-12-23 cs.CL 70%

Artificial intelligence contribution to translation industry: looking back and forward

人工智能对翻译行业的影响:回顾与展望

Mohammed Q. Shormani, Yehia A. Al-Sohbani

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究回顾了人工智能在翻译行业45年的发展历程,分析了机器翻译、低资源语言等热点问题,并指出需进一步研究以解决多方言和文化语域等挑战。

Comments 30 pages, 13 figures

Journal ref Discover Artificial Intelligence, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18445 2025-12-23 cs.LG 70%

On the Universality of Transformer Architectures; How Much Attention Is Enough?

Transformer架构的通用性;注意力是否足够?

Amirreza Abbasi, Mohsen Hooshmand

机构 * Department of Computer Science and Information Technology(计算机科学与信息技术系) Institute for Advanced Studies in Basic Sciences (IASBS)(基础科学高级研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文探讨Transformer架构的通用性问题,分析其表达能力及未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18362 2025-12-23 cs.CL 70%

SRS-Stories: Vocabulary-constrained multilingual story generation for language learning

SRS-Stories: 词汇受限的多语言故事生成用于语言学习

Wiktor Kamzela, Mateusz Lango, Ondrej Dusek

机构 * Poznan University of Technology, Institute of Computer Science(波兹南技术大学计算机科学学院) Charles University, Faculty of Mathematics and Physics(查尔斯大学数学与物理系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SRS-Stories通过间隔重复系统生成多语言故事,利用已知词汇教授新词并复习旧词,提升语言学习效果。

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12880 2025-12-23 cs.CR 67%

Universal Jailbreak Suffixes Are Strong Attention Hijackers

通用劫持后缀是强大的注意力劫持者

Matan Ben-Tov, Mor Geva, Mahmood Sharif

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了基于后缀的劫持攻击,发现通用后缀能更有效地劫持LLM的上下文化过程,并提出通过提升后缀普遍性来增强攻击效果,同时提供缓解方法。

Comments Accepted at TACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19020 2025-12-23 cs.CV cs.LG 57%

CETCAM: Camera-Controllable Video Generation via Consistent and Extensible Tokenization

CETCAM: 通过一致且可扩展的分词实现相机可控的视频生成

Zelin Zhao, Xinyu Gong, Bangya Liu, Ziyang Song, Jun Zhang, Suhui Wu, Yongxin Chen, Hao Zhang

机构 * Georgia Institute of Technology(佐治亚理工学院) ByteDance(字节跳动) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) The Hong Kong Polytechnic University(香港理工大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 CETCAM通过一致且可扩展的分词方案实现相机可控的视频生成,提高了几何一致性和视觉真实性,同时支持额外的控制模式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18996 2025-12-23 cs.IR cs.SE 50%

Modular Layout Synthesis (MLS): Front-end Code via Structure Normalization and Constrained Generation

模块化布局合成(MLS):通过结构规范化和约束生成实现前端代码

Chong Liu, Ming Zhang, Fei Li, Hao Zhou, Xiaoshuang Chen, Ye Yuan

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出模块化布局合成(MLS),通过结构规范化和约束生成,提升前端代码的模块化和重用性,适用于多种现代开发框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18813 2025-12-23 cs.CV 50%

Revealing Perception and Generation Dynamics in LVLMs: Mitigating Hallucinations via Validated Dominance Correction

揭示LVLMs中的感知与生成动态:通过验证主导修正缓解幻觉

Guangtao Lyu, Xinyi Cheng, Chenghao Xu, Qi Liu, Muli Yang, Fen Fang, Huilin Chen, Jiexi Yan, Xu Yang, Cheng Deng

机构 * School of Electronic Engineering, Xidian University(电子科技大学电子工程学院) School of Computer Science and Technology, Xidian University(电子科技大学计算机科学与技术学院) Hohai university(河海大学) Institute for Infocomm Research (I 2 R), A*STAR(信息通信研究所(I2R)) School of Foreign Languages, Xidian University(电子科技大学外国语言学院)

专题命中 其他LLM :language model(abstract)

AI总结 本文提出VDC策略,通过验证主导修正缓解LVLMs中的幻觉问题,揭示感知与生成的动态机制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18772 2025-12-23 cs.CV 50%

In-Context Audio Control of Video Diffusion Transformers

上下文音频控制视频扩散变换器

Wenze Liu, Weicai Ye, Minghong Cai, Quande Liu, Xintao Wang, Xiangyu Yue

机构 * MMLab, The Chinese University of Hong Kong(香港中文大学MMLab) Kling Team, Kuaishou Technology(快手科技Kling团队)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文提出ICAC框架,通过三维注意力机制实现音频驱动的视频生成,解决时间同步信号在视频扩散模型中的整合问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06716 2025-12-23 cs.SE cs.PF 50%

Efficiently Ranking Software Variants with Minimal Benchmarks

通过最小基准高效排序软件变体

Théo Matricon, Mathieu Acher, Helge Spieker, Arnaud Gotlieb

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出BISS方法,通过测试套件优化技术减少基准测试成本,同时保持变体排名的稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏