arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-08-04 至 2026-08-04 共收录 713 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 55 篇

2608.00119 2026-08-04 cs.CV cs.AI 新提交 70%

Counting the Cost of War Under Satellite Embargo: Zero-Shot Estimation of Impacted Infrastructure

卫星禁运下的战争损失统计:受影响基础设施的零样本估计

Saleh Sakib Ahmed, M. Sohel Rahman

机构 * Bangladesh University of Engineering and Technology(孟加拉国工程技术大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 针对冲突区人道主义救援中受影响建筑估计受卫星数据禁运阻碍的问题,提出基于打击前地图的零样本几何投影方法,引入两项技术创新,在2026年中东冲突数据上验证了深度增强大视觉语言模型的优势,确立了混合零样本危机映射范式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.05633 2026-08-04 cs.AI 版本更新 70%

Answer Presence Drives RAG Rewriting Gains

答案存在驱动RAG重写收益

Yuejie Li, Yueying Hua, Ke Yang, Li Zhang, Yueping He, Yueping He, Ruiqi Li, Bolin Chen, Tao Wang, Bowen Li, Chengjun Mao

机构 * Ant Group(蚂蚁集团)

专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 通过受控干预审计,发现检索增强问答中重写器带来的性能提升主要由黄金答案字符串出现在重写上下文中驱动,而非证据质量改善。

Comments The authors have withdrawn this manuscript after identifying errors in the experimental analysis reported in Sections 3 and 4. These errors affect the reported relationship between answer presence and RAG rewriting gains and undermine the paper's main conclusions. Therefore, the results and conclusions in the current version should not be relied upon

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02613 2026-08-04 cs.MA cs.AI cs.CY 版本更新 70%

Exploring Silicon-Based Societies: An Early Study of the Moltbook Agent Community

探索基于硅的社会:对Moltbook代理社区的早期研究

Yu-Zheng Lin, Bono Po-Jen Shih, Hsuan-Ying Alessandra Chien, Shalaka Satam, Jesus Horacio Pacheco, Naima Kaabouch, Sicong Shao, Soheil Salehi, Pratik Satam

机构 * 1 Department of Electrical Computer Engineering, University of Arizona 3 Department of Systems Industrial Engineering, University of Arizona 5 School of Electrical Engineering Computer Science, University of North Dakota 2 The Rock Ethics Institute \& The Leonhard Center for the Enhancement of Engineering Education, The Pennsylvania State University 4 Department of Industrial Engineering, University of Sonora 3 Telecommunication Laboratories, Chunghwa Telecom Co., Ltd. Corresponding Email

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文通过分析Moltbook平台,研究自主代理的社会结构形成,揭示了代理如何通过数据驱动的方法系统性地组织集体空间。

Comments 11 pages, 3 figures. Improves clarity and exposition and corrects minor errors. Technical content and conclusions remain unchanged

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01870 2026-08-04 cs.AI cond-mat.dis-nn math-ph math.MP math.NT 版本更新 70%

Testing Transformer Learnability on the Arithmetic Sequence of Rooted Trees

在根树的算术序列上测试Transformer的可学习性

Alessandro Breccia, Federica Gerace, Marco Lippi, Gabriele Sicuro, Pierluigi Contucci

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 该研究探讨了Transformer模型在根树算术序列上的学习能力,发现模型能部分捕捉序列的内部规律和相关性,表明可学习性可能延伸至算术结构本身。

Comments 33 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02064 2026-08-04 cs.LG cs.AI cs.CL 新提交 67%

Geometry-Guided Layerwise FFN Width Allocation in Transformers

Transformer中基于几何引导的分层前馈网络宽度分配

Timur Mudarisov, Mikhail Burtsev, Radu State

机构 * University of Luxembourg(卢森堡大学) London Institute of Mathematical Sciences(伦敦数学科学研究所)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 该研究针对Transformer中FFN宽度恒定的问题,提出基于几何特征的分层宽度分配方法,在多个预训练语言模型实验中,其可降低验证损失且效果优于均匀宽度和余弦衰减方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00721 2026-08-04 cs.RO 新提交 67%

OmniAI: A Surface-Adaptive Aerial Projection Interface for Human--Drone Interaction

OmniAI:一种面向人机交互的表面自适应空中投影界面

Nikita Kuzmin, Yuhua Jin, Georgii Demianchuk, Mariya Lezina, Fawad Mehboob, Ivan Valuev, Nikolai Lutsenko, Miguel Altamirano Cabrera, Dzmitry Tsetserukou

机构 * Skolkovo Institute of Science and Technology(斯科尔科沃科学技术研究院) School of Science and Engineering, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)理工学院)

专题命中 其他LLM :LLM(abstract,abstract_cn)

AI总结 OmniAI是一种具现化空中智能体,通过自适应投影实现人机交互,采用RGB-D与RANSAC检测投影表面,支持语音、手势控制,为情境感知人机交互提供移动空间AR界面。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.04015 2026-08-04 eess.SP 版本更新 67%

GenED-SC: Generative Editing Semantic Communication with Integrated Multi-Modal LLMs

GenED-SC:集成多模态大模型的生成式编辑语义通信

Shuoyao Wang, Suzhi Bi, Mingze Gong, Zhanpeng Wang, Li Ping Qian, Qiang Ye

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 提出一种两阶段语义图像传输框架,结合JSCC判别传输与MLLM生成编辑,在低信噪比下提升语义保真度和感知质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.01436 2026-08-04 cs.CY cs.AI cs.CL cs.HC 新提交 62%

Same violence, different answer: how AI responds to coercive control against women across languages

相同的暴力,不同的回应:人工智能如何对不同语言中针对女性的胁迫控制做出反应

Lyu Chang, Sònia Estradé Albiol, Núria Vergés Bosch

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究分析7种语言模型对9种语言下女性受伴侣胁迫控制场景的回应,发现非英语开发者系统在母语中易失效,前沿系统可实现统一保护,主张为各语言设定AI保护的最低标准。

Comments 16 pages, 1 figure, 2 tables. Supplementary methods, coding manual, and data workbook included as ancillary files

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24310 2026-08-04 cs.CL cs.LG 62%

Discovering Lexical Gaps Using Embeddings from Multilingual LLMs

利用多语言大语言模型的嵌入发现词汇空缺

Yoonwon Jung, Aaron S. Cohen, Benjamin K. Bergen

机构 * Department of Cognitive Science, University of California San Diego(加州大学圣地亚哥分校认知科学系)

专题命中 其他LLM :LLM(abstract_cn);分类 cs.CL、cs.LG

AI总结 提出一种数据驱动框架,通过多语言大语言模型的上下文嵌入计算语义相似度,以识别跨语言词汇空缺,在韩英和英韩方向上分别达到0.81和0.76的AUC。

Comments CoNLL 2026

Journal ref Yoonwon Jung, Aaron S. Cohen, and Ben Bergen. 2026. Discovering Lexical Gaps Using Embeddings from Multilingual LLMs. In Proceedings of the 30th Conference on Computational Natural Language Learning, pages 641-660

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09935 2026-08-04 cs.CL cs.AI 版本更新 62%

Unpacking Hateful Memes: Presupposed Context and False Claims

解析仇恨表情包:预设语境与虚假主张

Weibin Cai, Jiayu Li, Reza Zafarani

机构 * Syracuse University(Syracuse大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

AI总结 本文针对仇恨表情包检测中忽略仇恨本质的问题,提出含PCM与FACT模块的SHIELD框架,实验显示其检测性能优于现有最优方法,且可用于假新闻检测等任务。

Comments Accepted to SDM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.01793 2026-08-04 cs.LG 新提交 57%

ReFP-AD: Rectified Flow Preconditioning for Energy-Based Anomaly Detection

ReFP-AD:用于基于能量的异常检测的整流流预处理

Camile Lendering, Erkut Akdag, Joaquín Figueira, Egor Bondarev

机构 * Eindhoven University of Technology(埃因霍温理工大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 该研究针对统一异常检测中高维token空间EBM训练不稳定问题,提出ReFP-AD方法,经实验在MVTec-AD和VisA数据集上取得优于基线的异常检测性能。

Comments Accepted at the European Conference on Computer Vision (ECCV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.01414 2026-08-04 cs.AI cs.CR 新提交 57%

No Single Neuron of Failure: Distributed Safety Alignment Against White-Box Attacks

无单一失效神经元:针对白盒攻击的分布式安全对齐

Simiao Xie, Chuancheng Shi, Shangze Li, Wenhua Wu, Fei Shen, Ying Zhou, Zhiyong Wang, Tat-Seng Chua

机构 * The University of Sydney(悉尼大学) National University of Singapore(新加坡国立大学)

专题命中 其他LLM :foundation model(abstract);分类 cs.AI

AI总结 针对现有安全对齐方法的单点失效问题,提出分布式安全对齐(DSA),通过冗余编码安全能力提升模型对抗白盒神经元级攻击的鲁棒性,且保留通用语言与多模态效用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.01238 2026-08-04 cs.CL cs.MM 新提交 57%

Evaluating VLMs on Multimodal Aristotelian Persuasion Tasks

评估视觉语言模型(VLMs)在亚里士多德式多模态说服任务上的表现

Khondoker Ittehadul Islam

机构 * Saarland University(萨尔大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL

AI总结 该研究采用ImageArg数据集评估VLMs在亚里士多德式多模态说服任务的表现,发现Qwen系列模型在相关检测任务上性能提升并发布代码。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00284 2026-08-04 cs.RO cs.AI 新提交 57%

Hybrid Attention Estimation Pipeline for Adaptive HRI Using an Expressive Robotic Head

基于富有表现力的机器人头部的自适应人机交互混合注意力估计流程

Pablo Moraes, Monica Rodriguez, Christopher Peters, Hiago Sodre, Tobias Doernbach, Bruna Guterres, Ricardo Grando

机构 * Technological University of Uruguay(乌拉圭科技大学) Ostfalia University of Applied Sciences(奥斯特法利亚应用科技大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 该研究提出结合几何与语义感知层的混合注意力估计流程,通过有限状态机调节自适应人机交互,经10人40次试验验证其交互启动可靠、暂停行为一致且输出信息非冗余。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00247 2026-08-04 cs.SD cs.AI 版本更新 57%

Adaptive Perturbation Selection for Contrastive Audio Decoding

对比音频解码的自适应扰动选择

Aaron Isidore Grace, Zhouyuan Huo, Weiran Wang

机构 * Google(谷歌) University of Iowa(爱荷华大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI

AI总结 针对大型音频语言模型幻觉问题,提出自适应选择最优音频扰动作为对比解码负分支的方法,在时序、存在性等任务上提升准确率。

Comments In submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00463 2026-08-04 cs.CV cs.MM cs.SD 新提交 50%

Scene2Sound: Auditory-Grounded Soundscape Generation for 3D Gaussian Worlds

Scene2Sound:面向3D高斯世界的听觉锚定声景生成

Masaki Yoshida, Ren Togo, Takahiro Ogawa, Miki Haseyama

机构 * Hokkaido University(北海道大学)

专题命中 其他LLM :language model(abstract)

AI总结 Scene2Sound是一个无训练框架,通过听觉锚定和高斯集匹配为3D高斯世界生成空间一致的声景,在保留音频质量的同时解决了现有方法空间不一致的问题,经实验和用户研究验证有效。

Comments 14 pages. Project page: https://masaki-lmd.github.io/scene2sound/

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03434 2026-08-04 hep-th gr-qc nlin.CD 版本更新 50%

Berry Picking: Random Wave Chaos Hierarchy for BPS Microstate Geometries

Berry采摘:BPS微观态几何的随机波混沌层次结构

Vladan Djukić, Milica Stepanović, Mihailo Čubrović

专题命中 其他LLM :LLM(abstract_cn)

AI总结 研究不同超引力背景下探测波与测地线的混沌强度,发现波混沌随超对称减少和喉道变长变强,测地线运动则相反,还通过相关测试及计算解释二者差异,指出BPS混沌层次在体和场论中作用不同。

Comments 40 pages, 9 figures; this version: additional explanations in section 5, improved figures, additional discussion in section 6

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18309 2026-08-04 cs.GR cs.CV cs.SD eess.AS 50%

GCDance: Genre-Controlled Music-Driven 3D Full Body Dance Generation

GCDance:基于音乐的风格控制3D全身舞蹈生成

Xinran Liu, Xu Dong, Shenbin Qian, Diptesh Kanojia, Wenwu Wang, Zhenhua Feng

机构 * School of Computer Science and Electronic Engineering, University of Surrey(萨里大学计算机科学与电子工程学院) Department of Music and Media, University of Surrey(萨里大学音乐与媒体系) Department of Informatics, University of Oslo(奥斯陆大学信息学系) Centre for Vision, Speech and Signal Processing, University of Surrey(萨里大学视觉、语音与信号处理中心) School of Artificial Intelligence and Computer Science, Jiangnan University(江南大学人工智能与计算机科学学院)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文提出GCDance框架,通过音乐和文本条件生成风格化的3D舞蹈,利用文本控制机制和音乐基础模型提升生成质量,实验验证其优于现有方法。

Journal ref IEEE Transactions on Multimedia, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28568 2026-08-04 cs.CV 版本更新 50%

XSPA: Crafting Imperceptible X-Shaped Sparse Adversarial Perturbations for Transferable Attacks on VLMs

XSPA:构建不可察觉的X形稀疏对抗扰动以对视觉语言模型的可迁移攻击

Chengyin Hu, Jiaju Han, Xuemeng Sun, Qike Zhang, Luwei Yang, Lehan Sun, Jiahuan Long, Yiwei Wei, Jiujiang Guo

专题命中 其他LLM :language model(abstract)

AI总结 本文提出XSPA攻击,通过限制扰动为两条相交对角线,测试VLMs在稀疏扰动下的鲁棒性,实验表明其能显著破坏跨任务语义。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25129 2026-08-04 cs.CV 版本更新 50%

AirSplat: Alignment and Rating for Robust Feed-Forward 3D Gaussian Splatting

AirSplat:基于鲁棒前馈3D高斯散射的对齐与评分

Minh-Quan Viet Bui, Jaeho Moon, Munchurl Kim

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文提出AirSplat框架,通过自一致性姿态对齐和基于评分的透明度匹配技术,提升无姿态视角合成的重建质量。

Comments Project page: https://kaist-viclab.github.io/airsplat-site, accepted to ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18545 2026-08-04 cond-mat.dis-nn cond-mat.stat-mech cond-mat.str-el quant-ph 版本更新 50%

Large deviations in the many-body localization transition: The case of the random-field XXZ chain

多体局域化转变中的大偏差:随机场XXZ链的情况

Greivin Alfaro Miranda, Fabien Alet, Giulio Biroli, Leticia F. Cugliandolo, Nicolas Laflorencie, Marco Tarzia

专题命中 其他LLM :prompting(abstract)

AI总结 本研究针对随机场XXZ自旋链,采用平均场无序玻璃态系统类比方法,识别多体局域化转变的三个区域,推导有限大小相图,揭示有限尺寸下MBL相失稳源于罕见短程共振路径。

Comments Updated published version. Implemented corrections from the referees with the improvement of some of the Figures

Journal ref Phys. Rev. B 114, 014210 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03697 2026-08-04 eess.AS 版本更新 50%

Improving ASR Fairness for Cleft Lip and Palate Speech: A Study on Severity-Aware Data Mixing

改善唇腭裂语音的自动语音识别公平性:一项关于严重程度感知数据混合的研究

Susmita Bhattacharjee, Jagabandhu Mishra, H. S. Shekhawat, Ravi Jasuja, S. R. Mahadeva Prasanna

专题命中 其他LLM :foundation model(abstract)

AI总结 该研究针对唇腭裂(CLP)语音的ASR公平性问题,提出严重程度感知的CLP与正常语音混合策略,在AIISH和NMCPC数据集上使GMM-HMM、Whisper等模型的词错误率显著降低,提升了ASR公平性。

Comments Submitted to Speech Communication

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03827 2026-08-04 stat.AP 版本更新 50%

Equity in Focus : Investigating Gender Disparities in Glioblastoma via Propensity Score Matching

聚焦公平性:通过倾向得分匹配研究胶质母细胞瘤中的性别差异

Solomon Eshun

专题命中 其他LLM :prompting(abstract)

AI总结 本研究采用倾向得分匹配(PSM)分析癌症基因组图谱(TCGA)数据,发现调整混杂因素后男性胶质母细胞瘤(GBM)发病率高于女性,为性别公平性及针对性治疗提供依据。

Comments This preprint is withdrawn because the scope and direction of this research have changed substantially, and the current version is no longer representative of the work the author intends to disseminate

详情

展开后加载摘要…

URL PDF HTML 收藏