arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12265 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12265 篇

2601.04854 2026-04-07 cs.CL cs.AI cs.LG 67%

Projected Autoregression: Autoregressive Language Generation in Continuous State Space

投影自回归:在连续状态空间中进行自回归语言生成

Oshri Naparstek

机构 * IBM Research(IBM研究院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出投影自回归方法,通过连续预测和离散投影实现语言生成,展示连续状态空间在生成文本结构和动态方面的独特优势。

Comments In preperation to Neurips 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13036 2026-04-07 cs.HC cs.SI 67%

Uncovering the Internet's Hidden Values: An Empirical Study of Desirable Behavior Using Highly-Upvoted Content on Reddit

揭示互联网的隐藏价值:利用Reddit上高赞内容研究可取行为的实证研究

Agam Goyal, Charlotte Lambert, Yoshee Jain, Eshwar Chandrasekharan

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 通过分析Reddit上16000条高赞评论,提取了2016和2022年不同社区的64和72个宏观、中观和微观价值,发现现有计算模型仅能捕捉82%的提取价值,揭示了需更细致的可取行为模型。

Comments ICWSM'26: 16 pages, 3 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03612 2026-04-07 cs.CR 67%

Perceptual Gaps: ASCII Art and Overlapping Audio as CAPTCHA

感知缺口:ASCII艺术与重叠音频作为CAPTCHA

Choon-Hou Rafael Chong

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了利用人类高度专业化神经处理的CAPTCHA任务,提出基于ASCII艺术和重叠音频的两种CAPTCHA类型,测试结果显示现有LLM无法有效解决,表明其在当前有效但可能面临AI发展挑战。

Comments 8 pages, 3 figures. Research paper proposing novel CAPTCHA methods using ASCII art and overlapping audio

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17677 2026-04-06 cs.CL cs.AI cs.LG 67%

Adaptive Guidance for Retrieval-Augmented Masked Diffusion Models

适应性引导的检索增强掩码扩散模型

Jaemin Kim, Jong Chul Ye

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出ARAM框架,通过动态校准引导尺度提升检索增强扩散模型在知识密集型问答中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22340 2026-04-03 cs.HC 67%

Conversational Successes and Breakdowns in Everyday Smart Glasses Use

日常智能眼镜中的对话成功与失败

Xiuqi Tommy Zhu, Xiaoan Liu, Casper Harteveld, Smit Desai, Eileen McGivney

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究通过一个月的协作自民族志方法,分析了非显示智能眼镜在日常场景中的对话表现,揭示其独特优势与改进方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.01361 2026-04-03 cs.CV 67%

IGLOSS: Image Generation for Lidar Open-vocabulary Semantic Segmentation

IGLOSS:面向激光雷达开放词汇语义分割的图像生成

Nermin Samet, Gilles Puy, Renaud Marlet

机构 * LIGM, Ecole des Ponts, Univ Gustave Eiffel, CNRS(LIGM, 巴黎高科桥梁学院, 古斯塔夫·埃菲尔大学, 法国国家科学研究中心)

专题命中 其他LLM :language model(abstract);foundation model(abstract)

AI总结 本文提出一种新的方法,用于3D汽车激光雷达数据的零样本开放词汇语义分割。通过文本生成图像创建原型,利用2D视觉基础模型提取3D网络特征,实现点云标注,方法在nuScenes和SemanticKITTI上达到最新水平。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15997 2026-04-03 cs.LG cs.AI cs.CL 67%

The Geometric Anatomy of Capability Acquisition in Transformers

Transformer中能力获取的几何解剖

Jayadev Billa

机构 * San Jose, CA, USA(美国加利福尼亚州圣何塞) Yahoo(雅虎) Nuance(纽昂斯) BBN(BBN科技)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究揭示了Transformer模型在训练过程中能力获取的几何变化与行为变化的关系,发现任务难度和模型规模影响几何变化与行为表现的先后顺序,且仅在困难任务中存在明显的先决阶段。

Comments 19 pages (13 pages main, 6 pages appendix), 13 tables, 8 figures. v4: significant rewrite with additional experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00395 2026-04-02 cs.CV 67%

Advancing Complex Video Object Segmentation via Tracking-Enhanced Prompt: The 1st Winner for 5th PVUW MOSE Challenge

通过跟踪增强提示推进复杂视频对象分割:第五届PVUW MOSE挑战赛第一名

Jinrong Zhang, Canyang Wu, Xusheng He, Weili Guan, Jianlong Wu, Liqiang Nie

机构 * Harbin Institute of Technology, Shenzhen, China(哈尔滨工业大学(深圳)) Shenzhen Loop Area Institute, China(深圳市环区研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出TEP方法,通过跟踪增强提示解决SAM3在处理小目标和语义主导对象时的不足,取得PVUW挑战赛优异成绩。

Comments 1st Place Solution for the 5th PVUW MOSE Challenge (CVPR 2026 Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09697 2026-04-02 cs.LG cs.AI cs.CL 67%

Mousse: Rectifying the Geometry of Muon with Curvature-Aware Preconditioning

Mousse:通过曲率感知预条件化校正μon的几何

Yechen Zhang, Shuhao Xing, Junhao Huang, Kai Lv, Yunhua Zhou, Xipeng Qiu, Qipeng Guo, Kai Chen

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai AI Laboratory(上海人工智能实验室) Fudan University(复旦大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出Mousse优化器,通过曲率感知预条件化改进谱方法的结构稳定性,实验证明其在大规模语言模型训练中显著提升效率。

Comments 17 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12052 2026-04-01 cs.CV cs.GR 67%

A Text-to-3D Framework for Joint Generation of CG-Ready Humans and Compatible Garments

为生成兼容服装的CG-ready人类和服装联合生成的文本到3D框架

Zhiyao Sun, Yu-Hui Wen, Ho-Jui Fang, Sheng Ye, Matthieu Lin, Tian Lv, Yong-Jin Liu

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出Tailor框架,通过语义解析、几何感知服装生成和一致纹理合成,实现高保真定制3D人像和物理兼容服装,优于现有方法。

Comments Project page: https://human-tailor.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01405 2026-03-31 cs.HC 67%

Feedback by Design: Understanding and Overcoming User Feedback Barriers in Conversational Agents

反馈由设计:理解并克服对话代理中用户反馈的障碍

Nikhil Sharma, Zheng Zhang, Daniel Lee, Namita Krishnan, Guang-Jie Ren, Ziang Xiao, Yunyao Li

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究探讨了用户反馈障碍,通过Grice准则识别出四个障碍,并提出设计原则以提升反馈质量。

Comments Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems, 23 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07992 2026-03-31 econ.EM 67%

Fake Date Tests: Can We Trust In-sample Accuracy of LLMs in Macroeconomic Forecasting?

虚假日期测试:我们能否信任LLM在宏观经济预测中的样本内准确性?

Alexander Eliseev, Sergei Seleznev

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过设计虚假日期测试,检验LLM在宏观经济预测中的样本内准确性是否能推广到样本外表现,发现现代LLM存在样本内预测的先行偏见。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.27332 2026-03-31 cs.CV 67%

Unsafe by Reciprocity: How Generation-Understanding Coupling Undermines Safety in Unified Multimodal Models

因互惠而不安全:生成-理解耦合如何在统一多模态模型中损害安全性

Kaishen Wang, Heng Huang

机构 * University of Maryland, College Park(马里兰大学帕克分校)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究探讨了统一多模态模型中生成与理解之间的互惠关系可能成为安全漏洞的根源,提出RICE攻击方法,揭示了跨功能互惠对安全性的负面影响。

Comments 7 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24454 2026-03-26 cs.CV 67%

Unleashing Vision-Language Semantics for Deepfake Video Detection

释放视觉-语言语义以进行深度伪造视频检测

Jiawen Zhu, Yunqi Miao, Xueyi Zhang, Jiankang Deng, Guansong Pang

机构 * Singapore Management University(新加坡管理学院) The University of Warwick(沃里克大学) Nanyang Technological University(南洋理工大学) Imperial College London(伦敦帝国理工学院)

专题命中 其他LLM :language model(abstract);prompting(abstract)

AI总结 本文提出VLAForge框架,通过ForgePerceiver增强视觉感知并引入身份感知VLA评分,利用跨模态语义提升深度伪造检测的判别能力。

Comments 14 pages, 7 figures, accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.21944 2026-03-24 cs.CV 67%

Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection

Group3D: 基于多模态大语言模型的开放词汇3D目标检测语义分组

Youbin Kim, Jinho Park, Hogun Park, Eunbyung Park

机构 * Department of Artificial Intelligence, Sungkyunkwan University(成均馆大学人工智能系) Department of Artificial Intelligence, Yonsei University(延世大学人工智能系)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Group3D通过将语义约束整合到实例构建过程中,解决多视角开放词汇3D目标检测中的几何过合并问题,实现语义与几何一致性融合的检测框架。

Comments 24 pages, 7 figures, Project page: https://ubin108.github.io/Group3D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.20847 2026-03-24 cs.SE 67%

Engineering Pitfalls in AI Coding Tools: An Empirical Study of Bugs in Claude Code, Codex, and Gemini CLI

人工智能编码工具中的工程陷阱:对Claude代码、Codex和Gemini CLI中bug的实证研究

Ruixin Zhang, Wuyang Dai, Hung Viet Pham, Gias Uddin, Jinqiu Yang, Song Wang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过分析超过3800个公开报告的bug,揭示了构建AI编码工具中的常见故障模式和工程挑战,发现功能相关bug占比超67%,主要症状包括API错误、终端问题和命令失败。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02775 2026-03-24 cs.HC 67%

Expecting Too Much, Getting Too Little: Exploring the Challenges and Design Opportunities of Asynchronous AI Interviewers

期望过高,收获过低:探索异步AI面试官的挑战与设计机会

Md Nazmus Sakib, Naga Manogna Rayasam, Sanorita Dey

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 研究探讨了异步AI面试系统对用户自主性的影响,通过分析和访谈揭示了期望不匹配问题,并设计了支持用户自主性的界面。

Comments This paper has been accepted at CSCW 2026 and will appear in the Proceedings of the ACM on Human-Computer Interaction

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.18811 2026-03-20 cs.RO 67%

V-Dreamer: Automating Robotic Simulation and Trajectory Synthesis via Video Generation Priors

V-Dreamer:通过视频生成先验自动机器人仿真与轨迹合成

Songjia He, Zixuan Chen, Hongyu Ding, Dian Shao, Jieqi Shi, Chenxu Li, Jing Huo, Yang Gao

机构 * Nanjing University(南京大学) Northwestern Polytechnical University(西北工业大学)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 V-Dreamer通过视频生成先验自动构建仿真环境和可执行轨迹,利用大语言模型和3D生成模型生成物理合理的3D场景,并通过Sim-to-Gen模块实现高效轨迹合成。

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17272 2026-03-20 cs.CR cs.ET 67%

Network and Device Level Cyber Deception for Contested Environments Using RL and LLMs

网络与设备层面的网络空间欺骗用于对抗环境中的RL和LLMs

Abhijeet Sahu, Shuva Paul, Richard Macwan

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨利用RL和LLMs构建网络与设备层面的网络欺骗方法,以提高对抗环境中的欺骗策略准确性和成本效益。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17192 2026-03-19 cs.CY 67%

Narrative Frames: A New Approach to Analysing Metaphors in AI Ethics and Policy Discourse

叙事框架:一种分析人工智能伦理与政策 discourse 中隐喻的新方法

Daniel Stone

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出Narrative Frames框架,通过归纳编码和交叉参考,系统分析AI政策 discourse 中隐喻,解决现有方法定义不一致的问题,为研究者和政策制定者提供共同词汇。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17055 2026-03-19 cs.CV 67%

PaAgent: Portrait-Aware Image Restoration Agent via Subjective-Objective Reinforcement Learning

PaAgent:通过主观-客观强化学习实现的面向人物图像修复代理

Yijian Wang, Qingsen Yan, Jiantao Zhou, Duwei Dai, Wei Dong

机构 * School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院) Shenzhen Research Institute of Northwestern Polytechnical University(西北工业大学深圳研究院) State Key Laboratory of Internet of Things for Smart City, University of Macau(澳门大学智慧城市物联网国家重点实验室) National-Local Joint Engineering Research Center of Biodiagnosis and Biotherapy, the Second Affiliated Hospital of Xi’an Jiaotong University(西安交通大学生物诊断与生物治疗国家地方联合工程研究中心) College of Information and Control Engineering, Xi’an University of Architecture and Technology(西安建筑科技大学信息与控制工程学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 PaAgent通过结合自进化的人物银行和检索增强生成技术,提升图像修复任务中对复杂场景的感知能力,通过主观-客观强化学习策略优化修复工具选择,实验验证其在多种修复基准上的优越性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16633 2026-03-18 cs.HC 67%

Why We Need to Destroy the Illusion of Speaking to A Human: Critical Reflections On Ethics at the Front-End for LLMs

为什么我们需要摧毁与人类交谈的幻觉:关于LLMs前端伦理的批判性反思

Sarah Diefenbach, Daniel Ullrich

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨了基于LLM的聊天机器人与人类对话的相似性及其伦理挑战,提出改进AI前端设计的伦理原则。

Comments CHI 2026 Conference on Human-Computer Interaction

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16196 2026-03-18 cs.RO 67%

PanguMotion: Continuous Driving Motion Forecasting with Pangu Transformers

PanguMotion: 基于Pangu Transformer的连续驾驶运动预测

Quanhao Ren, Yicheng Li, Nan Song

机构 * School of Data Science, Fudan University(复旦大学数据科学学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出PanguMotion框架,利用Pangu-1B的Transformer模块提升连续驾驶场景的运动预测能力,通过Argoverse 2数据集的RealMotion策略生成连续序列以提升预测准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.12282 2026-03-18 cs.IR 67%

Algorithmic Trust and Compliance: Benchmarking Brand Notability for UK iGaming Entities in Generative Search Engines

算法信任与合规:在生成搜索引擎中评估英国在线博彩实体的品牌知名度

Julen Oruesagasti

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨生成搜索引擎中品牌知名度的评估方法,分析合规信号如何影响大语言模型的权威性,提出新的优化框架。

Comments Technical Report. Produced by Interamplify Research Division (UK)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18444 2026-03-18 cs.CV 67%

SineProject: Machine Unlearning for Stable Vision Language Alignment

SineProject:用于稳定视觉语言对齐的机器反遗忘

Arpit Garg, Hemanth Saratchandran, Simon Lucey

机构 * Australian Institute for Machine Learning(澳大利亚机器学习研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 SineProject通过在冻结的投影器中加入正弦调制的可训练参数,提升Jacobian的谱条件数,稳定反遗忘过程中的视觉语言对齐,减少无害查询拒绝并实现目标信息遗忘。

Comments Accepted at CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15034 2026-03-17 cs.CL cs.AI cs.LG 67%

Interpretable Predictability-Based AI Text Detection: A Replication Study

可解释的预测性AI文本检测:一项复制研究

Adam Skurla, Dominik Macko, Jakub Simko

机构 * Faculty of Information Technology, Brno University of Technology(布拉格技术学院) Kempelen Institute of Intelligent Technologies(智能技术研究所)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文复制并扩展了AuTexTification 2023共享任务中用于机器生成文本作者归属的系统,测试了新多语言模型并添加了26个文档级风格学特征,通过SHAP分析探讨特征影响,使用Qwen等新生成模型提升性能,多语言配置在性能上不逊于语言特定模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.14662 2026-03-17 cs.HC 67%

ViDscribe: Multimodal AI for Customizing Audio Description and Question Answering in Online Videos

ViDscribe:多模态AI用于定制在线视频的音频描述和问答

Maryam Cheema, Sina Elahimanesh, Pooyan Fazli, Hasti Seifi

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 ViDscribe通过整合AI生成的音频描述和六种用户自定义选项,提升视障和低视力用户在YouTube视频中的交互体验,研究显示定制化描述提高了效果和沉浸感。

Comments CHI EA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.13876 2026-03-17 cs.MA cs.CY 67%

How do Role Models Shape Collective Morality? Exemplar-Driven Moral Learning in Multi-Agent Simulation

角色榜样如何塑造集体道德?多智能体模拟中的范例驱动道德学习

Junjie Liao, Huacong Tang, Zhou Ziheng, Yizhou Wang, Fangwei Zhong

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过多智能体模拟研究角色榜样对集体道德的影响,发现身份驱动的从众行为能迅速改变初始倾向,推动价值观趋同。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10232 2026-03-12 cs.RO 67%

Hierarchical Task Model Predictive Control for Sequential Mobile Manipulation Tasks

分层任务模型预测控制用于连续移动操作任务

Xintong Du, Siqi Zhou, Angela P. Schoellig

机构 * Learning Systems and Robotics Lab(学习系统与机器人实验室) Technical University of Munich(慕尼黑技术大学) University of Toronto Institute for Aerospace Studies(多伦多大学航空航天研究所) University of Toronto Robotics Institute(多伦多大学机器人研究所) Munich Institute of Robotics and Machine Intelligence(慕尼黑机器人与机器智能研究所) Vector Institute for Artificial Intelligence(人工智能向量研究所)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出了一种分层任务模型预测控制框架,通过利用机器人冗余性,提高了任务序列执行的性能和反应性,实验显示在任务变化和参考变化情况下,轨迹跟踪性能提升了42%。

Comments 8 pages, Published in IEEE Robotics and Automation Letters ( Volume: 9, Issue: 2, February 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10228 2026-03-12 cs.CR 67%

Paladin: A Policy Framework for Securing Cloud APIs by Combining Application Context with Generative AI

Paladin:通过结合应用上下文与生成式AI实现云API安全的策略框架

Shriti Priya, Julian James Stephen, Arjun Natarajan

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Paladin通过结合应用上下文和生成式AI,为云API安全提供了一种策略框架,能够有效防止资源消耗、敏感业务流程访问和认证断裂等威胁。

详情

展开后加载摘要…

URL PDF HTML 收藏