arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12241 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12241 篇

2604.19429 2026-04-22 cs.HC cs.CY 75%

Discerning Authorship in Online Health Communities: Experience, Trust, and Transparency Implications for Moderating AI

在在线健康社区中辨别作者身份:经验、信任与透明度对调节AI的影响

Yefim Shulman, Agnieszka Kitkowska, Mark Warner

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究探讨在线健康社区中用户辨别AI生成建议作者身份的能力,发现透明度与信任的重要性,尽管用户难以区分AI与人类生成内容,但健康状况有显著影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10575 2026-04-14 cs.HC 75%

NexusAI: Enabling Design Space Exploration of Ideas through Cognitive Abstraction and Functional Decomposition

NexusAI: 通过认知抽象和功能分解实现想法的设计空间探索

Anqi Wang, Bingqian Wang, Huiyang Chen, Keqing Jiao, Lei Han, Xin Tong, Pan Hui

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 NexusAI通过认知抽象和功能分解解决LLM生成想法的结构性不透明问题,提升设计空间探索效率和创造性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10222 2026-04-14 cs.CY 75%

Morally Programmed LLMs Reshape Human Morality

道德编程的大型语言模型重塑人类道德

Pengzhao Lyu, Yeun Joon Kim, Yingyue Luna Luan, Jungmin Choi

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究通过道德编程的LLM与人类交互,发现其能系统性地改变人类道德倾向,且影响持续两周,揭示了道德原则嵌入LLM的伦理困境。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09444 2026-04-13 cs.HC 75%

Confidence Without Competence in AI-Assisted Knowledge Work

人工智能辅助知识工作中缺乏能力的自信

Elena Eleftheriou, George Pallis, Marios Constantinides

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了不同LLM交互设计对深度思考的影响,发现未来自我解释能提高理解和学习效果,而引导提示能带来最大学习收益。

Comments 25 pages, 13 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.09120 2026-04-13 cs.SE 75%

The Role of LLMs in Collaborative Software Design

大型语言模型在协作软件设计中的作用

Victoria Jackson, Yoonha Cha, Rafael Prikladnicki, André van der Hoek

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究探讨了LLM在软件设计协作中的影响,发现共享实例促进理解,而并行使用可能导致上下文漂移,专业人员会审查LLM响应以获得设计洞察,但早期锚定可能限制探索。

Comments accepted into the 2nd HumanAISE workshop 2026, to be published in the FSE Companion '26

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05876 2026-04-09 eess.SY cs.SY 75%

Context-Aware Model Predictive Control for Microgrid Energy Management via LLMs

基于LLMs的上下文感知模型预测控制用于微电网能源管理

Ruixiang Wu, Jiahao Ai, Tinko Sebastian Bartels, Tongxin Li

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出InstructMPC框架,利用LLM和可调最后一层映射,将非结构化操作上下文转化为MPC控制器的预测扰动轨迹,通过理论分析和实验验证,证明了在微电网中整合语义信息提升控制效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24895 2026-04-08 cs.HC 75%

PII Shield: A Browser-Level Overlay for User-Controlled Personal Identifiable Information (PII) Management in AI Interactions

PII Shield:一种浏览器层面的叠加层,用于在AI交互中实现用户控制的个人可识别信息(PII)管理

Max Holschneider, Saetbyeol LeeYouk

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出PII Shield,通过企业级红actions管道,为用户提供直观的免费AI交互隐私保护体验,引入本地实体匿名化和干扰第三方分析的'烟雾'机制,以平衡用户数据使用与隐私保护。

Comments Accepted at the Proceedings of the CHI 2026 Workshop: Ethics at the Front-End

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09335 2026-04-08 cs.SE 75%

Can ChatGPT Generate Realistic Synthetic System Requirement Specifications? Results of a Case Study

ChatGPT能否生成逼真的合成系统需求规范?案例研究的结果

Alex R. Mattukat, Florian M. Braun, Horst Lichter

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文探讨了ChatGPT能否在不接触真实系统需求规范的情况下生成逼真的合成规范,通过系统方法生成300个跨10行业的规范,并发现62%的专家认为其合理,但存在矛盾陈述和缺陷,强调需结合专家评估。

Comments This is the accepted version of a paper that will appear in the proceedings of the 21st International Conference on Evaluation of Novel Approaches of Software Engineering (ENASE 2026). The final published version will be available from Science and Technology Publications (SCITEPRESS). 15 pages, 3 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.04283 2026-04-07 cs.CR 75%

Semantics Over Syntax: Uncovering Pre-Authentication 5G Baseband Vulnerabilities

语义超越语法:揭示预认证5G基带漏洞

Qiqing Huang, Xingyu Wang, Wanda Guo, Guofei Gu, Hongxin Hu

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究5G基带在预认证前的语义不一致问题,提出ConSeT框架通过提取规范约束生成语义违规测试用例,发现7个未知漏洞及29个崩溃点。

Comments To appear in the 35th USENIX Security Symposium (USENIX Security 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.04028 2026-04-07 eess.SP 75%

Enhancing 6G Wireless Intelligence: Do LLMs Work for CSI Prediction?

增强6G无线智能:LLMs能否用于CSI预测?

Mohsen Kazemian, Jürgen Jasperneite

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出基于LLM的OTFS信道预测框架,结合物理描述符提升动态环境下的信道预测精度,实验表明其在NMSE上优于传统深度学习和无物理描述的LLM方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18985 2026-04-06 cs.SI 75%

Simulating Online Social Media Conversations on Controversial Topics Using AI Agents Calibrated on Real-World Data

利用真实数据校准的AI代理模拟争议性话题的在线社交媒体对话

Elisa Composta, Nicolo' Fontana, Francesco Corso, Francesco Pierri

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了基于LLM的代理在模拟微博客社交网络中的行为,探讨了其在不同场景下生成内容、互动及意见演变的机制,发现其生成内容在语气和毒性方面不如真实数据多样,需更精细的认知建模以模拟人类行为。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.01086 2026-04-03 cs.DS cs.IT math.IT math.ST stat.TH 75%

Asymptotically Optimal Sequential Testing with Heterogeneous LLMs

异质大语言模型的渐近最优顺序检验

Guokai Li, Alys Liang, Mo Liu, Murray Lei, Stefanus Jasin, Fenghua Yang, Preet Baxi

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了使用异质大语言模型进行贝叶斯二元顺序假设检验问题,通过分析信息率和成本,证明在误差容忍度趋近于零时,最优策略等价于使用最多两个模型,并构造了基于信念的混合策略以逼近下限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00518 2026-04-03 cs.CY 75%

Do Agents Repair When Challenged -- or Just Reply? Challenge, Repair, and Public Correction in a Deployed Agent Forum

智能体在受挑战时会修复还是仅回应?挑战、修复与公共纠正在一个部署的智能体论坛中的探讨

Luyang Zhang, Yi-Yun Chu, Jialu Wang, Beibei Li, Ramayya Krishnan

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 研究比较了部署的智能体论坛Moltbook与Reddit社区,发现Moltbook讨论线程较少,挑战回应率低,修复机制缺失,表明社区维持互动过程对规范教学和修订的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19851 2026-04-02 cs.CV 75%

Towards Online Multi-Modal Social Interaction Understanding

面向在线多模态社交交互理解

Xinpeng Li, Shijian Deng, Bolin Lai, Weiguo Pian, James M. Rehg, Yapeng Tian

机构 * University of Texas at Dallas(德克萨斯大学达拉斯分校) Georgia Institute of Technology(佐治亚理工学院) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文提出在线多模态社交交互理解问题,通过历史信息实现多模态社交交互理解,采用多模态大语言模型框架,引入多对话者预测和社交感知视觉提示,实现三个任务的最优性能。

Comments Accepted to Transactions on Machine Learning Research (TMLR). Project page: https://sampson-lee.github.io/online-mmsi-project-page

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.29982 2026-04-01 cs.GT 75%

Performative Scenario Optimization

表现性场景优化

Quanyan Zhu, Zhengye Han

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出了一种表现性场景优化框架,用于决策依赖的约束问题,通过反馈循环考虑决策对数据生成过程的影响,证明了其存在性并展示了在AI安全应用中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28655 2026-03-31 cs.CR 75%

Safeguarding LLMs Against Misuse and AI-Driven Malware Using Steganographic Canaries

通过隐写术可以ary文件保护LLM免受滥用及AI驱动的恶意软件

Md Raz, Venkata Sai Charan Putrevu, Meet Udeshi, Prashanth Krishnamurthy, Farshad Khorrami, Ramesh Karri

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出基于隐写术canary文件的框架,用于检测未经授权的LLM处理,通过两种模式实现符号和语言隐写术的结合,有效识别敏感文档。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.26635 2026-03-30 cs.MA 75%

Deception and Communication in Autonomous Multi-Agent Systems: An Experimental Study with Among Us

自主多智能体系统中的欺骗与交流:以Among Us为例的实证研究

Maria Milkowski, Tim Weninger

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文通过Among Us游戏研究自主智能体的欺骗与交流行为,发现智能体主要使用指示性语言,伪装者倾向使用解释和否认等行为,欺骗多表现为 equivocation 而非直接谎言,揭示了自主交流中真理与效用之间的根本矛盾。

Comments 8 pages + references, 9 figures. Accepted at AAMAS 2026

Journal ref Proceedings of the 25th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2026), IFAAMAS, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03282 2026-03-30 cs.CV cs.GR cs.HC 75%

MIBURI: Towards Expressive Interactive Gesture Synthesis

MIBURI:迈向表达性交互手势合成

M. Hamza Mughal, Rishabh Dabral, Vera Demberg, Christian Theobalt

机构 * Max Planck Institute for Informatics, SIC(马克斯·普朗克信息学研究所,SIC) Saarland University(萨尔大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 MIBURI提出首个实时因果框架,生成与实时口语对话同步的表达性全身手势和面部表情,通过多级离散标记和辅助目标提升自然度与多样性。

Comments CVPR 2026 (Main). Project page: https://vcai.mpi-inf.mpg.de/projects/MIBURI/

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17946 2026-03-26 cs.HC 75%

"I Use ChatGPT to Humanize My Words": Affordances and Risks of ChatGPT to Autistic Users

我用ChatGPT让话语更有人性化:ChatGPT对自闭症用户的影响与风险

Renkai Ma, Ben Zefeng Zhang, Chen Chen, Fan Yang, Xiaoshan Huang, Haolun Wu, Lingyao Li

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文通过分析3984条自闭症用户社交媒体帖子,探讨ChatGPT在缓解自闭症用户执行功能障碍、情绪调节等方面的作用,同时指出其对用户心理健康带来的风险。

Comments Accepted to ACM Interactive Health '2026 extended abstract

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23880 2026-03-26 cs.SI 75%

ProcureGym: A Multi-Agent Markov Game Framework for Modeling National Volume-based Drug Procurement

ProcureGym: 一个用于建模国家基于体积的药品采购的多智能体马尔可夫游戏框架

Jia Wang, Qian Xu, Xuanwen Ding, Zhuangqi Li, Chao He, Bao Liu, Zhongyu Wei

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出ProcureGym,通过真实数据模拟中国基于体积的药品采购,评估RL、LLM等智能体模型,发现RL在收益和利润上表现更优,揭示最大有效报价和采购量对战略结果的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23852 2026-03-26 cs.SE 75%

APISENSOR: Robust Discovery of Web API from Runtime Traffic Logs

APISENSOR: 从运行时流量日志中鲁棒地发现Web API

Yanjing Yang, Chenxing Zhong, Ke Han, Zeru Cheng, Jinwei Xu, Xin Zhou, He Zhang, Bohan Liu

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 APISENSOR通过无监督方法从混合运行时流量中恢复准确API,提升了API发现的鲁棒性和准确性,优于现有方法。

Comments 14 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19747 2026-03-23 cs.HC 75%

ConSearcher: Supporting Conversational Information Seeking in Online Communities with Member Personas

ConSearcher:通过成员人设支持在线社区中的对话式信息检索

Shiwei Wu, Xinyue Chen, Yuheng Liu, Xingbo Wang, Qingyu Guo, Longfei Chen, Chuhan Shi, Zhenhui Peng

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出ConSearcher,一种基于LLM的对话式信息检索工具,通过动态生成成员人设提升在线社区的信息获取效率,实验显示其在信息获取和用户参与度上优于现有方法,但存在过度个性化的问题。

Comments 25 pages, 7figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19026 2026-03-20 cs.CV 75%

Rethinking MLLM Itself as a Segmenter with a Single Segmentation Token

重新思考MLLM本身作为分割器:仅用一个分割标记

Anqi Zhang, Xiaokang Ji, Guangyu Gao, Jianbo Jiao, Chi Harold Liu, Yunchao Wei

机构 * Beijing Institute of Technology(北京理工大学) University of Birmingham(伯明翰大学) Beijing Jiaotong University(北京交通大学) Beijing Academy of Artificial Intelligence(北京人工智能研究院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出SELF1E方法,通过保留原始图像分辨率并利用残差特征提升分割精度,无需外部解码器即可实现与专业解码器相当的分割性能。

Comments Paper is accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24310 2026-03-18 eess.AS 75%

Code-switching Speech Recognition Under the Lens: Model- and Data-Centric Perspectives

代码切换语音识别的多视角审视:模型与数据为中心的视角

Hexin Liu, Haoyang Zhang, Qiquan Zhang, Xiangyu Zhang, Dongyuan Shi, Eng Siong Chng, Haizhou Li

专题命中 其他LLM :large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文从模型和数据两个角度分析代码切换语音识别的挑战,提出简化等价约束理论(SECT)提升ASR性能和语言质量,通过数据增强和提示策略改进代码切换文本生成。

Comments 14 pages, 4 figures, 10 tables, accepted to IEEE TASLP. Copyright has been transferred to IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15691 2026-03-18 cs.SE 75%

VibeContract: The Missing Quality Assurance Piece in Vibe Coding

VibeContract: vibe编码中缺失的质量保证环节

Song Wang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出VibeContract框架,通过将高层次自然语言意图分解为任务序列并生成任务级合同,以提升LLM生成代码的正确性、鲁棒性和可维护性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10387 2026-03-12 cs.CR 75%

Don't Let the Claw Grip Your Hand: A Security Analysis and Defense Framework for OpenClaw

不要让爪子抓住你的手:OpenClaw平台的安全分析与防御框架

Zhengyang Shan, Jiayun Xin, Yue Zhang, Minghui Xu

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文分析了OpenClaw平台的安全问题,提出HITL防御机制,显著提升系统防御率至19%-92%。

Comments 12 pages, 2 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09544 2026-03-11 cs.CR 75%

Compartmentalization-Aware Automated Program Repair

具有 compartmentalization 意识的自动化程序修复

Jia Hu, Youcheng Sun, Pierre Olivier

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文提出了一种专门针对 compartment 接口安全的 APR 框架,通过整合 fuzzer、补丁生成和验证组件,自动化修复 cross-compartment 接口漏洞。

Comments Accepted to appear in ICSE's Journal Ahead Workshop (JAWs) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07442 2026-03-10 cs.RO 75%

LITHE: Bridging Best-Effort Python and Real-Time C++ for Hot-Swapping Robotic Control Laws on Commodity Linux

LITHE:连接最佳努力Python与实时C++以实现机器人控制律的热插拔

He Kai Lim, Tyler R. Clites

机构 * Department of Mechanical and Aerospace Engineering, University of California Los Angeles(加州大学洛杉矶分校机械与航空航天工程系)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 LITHE通过轻量级架构实现Python与C++的实时控制律热插拔,提升机器人系统在动态环境中的适应能力。

Comments 8 pages, 5 figures. Submitted to IEEE/RSJ International Conference on Intelligent Robots & Systems (IROS) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02895 2026-03-04 cs.AR cs.PL 75%

SpecLoop: An Agentic RTL-to-Specification Framework with Formal Verification Feedback Loop

SpecLoop:一种带有形式验证反馈循环的代理RTL到规范框架

Fu-Chieh Chang, Yu-Hsin Yang, Hung-Ming Huang, Yun-Chia Hsu, Yin-Yu Lin, Ming-Fang Tsai, Chun-Chih Yang, Pei-Yuan Wu

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 SpecLoop通过形式验证驱动的反馈循环提升RTL到规范生成的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12386 2026-03-04 cs.HC 75%

Hey Dashboard!: Supporting Voice, Text, and Pointing Modalities in Dashboard Onboarding

Hey Dashboard!:在仪表盘引导中支持语音、文本和指向模态

Vaishali Dhanoa, Gabriela Molina León, Eve Hoggan, Eduard Gröller, Marc Streit, Niklas Elmqvist

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 DIANA通过语音、文本和指向等多种模态,为仪表盘引导提供自助导航和分析支持,提升用户在复杂仪表盘中的使用效率。

Journal ref Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems (CHI '26), April 13--17, 2026, Barcelona, Spain

详情

展开后加载摘要…

URL PDF HTML 收藏