arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-25 至 2026-02-25 共收录 15 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 15 篇

2602.14733 2026-02-25 cs.HC 89%

More than Decision Support: Exploring Patients' Longitudinal Usage of Large Language Models in Real-World Healthcare-Seeking Journeys

超越决策支持:探索患者在真实世界医疗寻求旅程中大型语言模型的纵向使用

Yancheng Cao, Yishu Ji, Chris Yue Fu, Sahiti Dharmavaram, Meghan Turchioe, Natalie C Benda, Lena Mamykina, Yuling Sun, Xuhai "Orson" Xu

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract)

AI总结 本文通过为期四周的日记研究,探讨患者在真实世界医疗寻求旅程中如何将LLM作为动态伴侣,改变传统患者-提供者关系中的权力动态,并提出未来LLM应作为纵向边界伴侣持续调解患者与临床医生之间的互动。

Comments Conditionally accepted to CHI Conference on Human Factors in Computing Systems (CHI'26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19297 2026-02-25 cs.AI 88%

Automated Generation of Microfluidic Netlists using Large Language Models

利用大语言模型自动化生成微流体网表

Jasper Davidson, Skylar Stockham, Allen Boston, Ashton Snelgrove, Valerio Tenace, Pierre-Emmanuel Gaillardon

机构 * Department of Electrical and Computer Engineering, University of Utah(电气与计算机工程系,犹他大学) Primis AI, Inc.(Primis AI 公司)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文利用大语言模型自动化生成微流体器件的结构网表,实现了从自然语言描述到Verilog代码的转换,并在典型微流体设计中验证了其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20486 2026-02-25 cs.HC cs.AI 85%

Hybrid LLM-Embedded Dialogue Agents for Learner Reflection: Designing Responsive and Theory-Driven Interactions

混合LLM嵌入对话代理用于学习者反思:设计响应性且理论驱动的互动

Paras Sharma, YuePing Sha, Janet Shufor Bih Epse Fofang, Brayden Yan, Jess A. Turner, Nicole Balay, Hubert O. Asare, Angela E. B. Stewart, Erin Walker

机构 * University of Pittsburgh(匹兹堡大学) Carnegie Mellon University(卡内基梅隆大学) Bowie State University(鲍威州立大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种混合对话系统,结合LLM的响应能力与理论驱动的规则框架,以支持学习者反思,但发现其在提高反思深度的同时也带来了一些挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02866 2026-02-25 cs.SE cs.AI cs.AR cs.CR 83%

LM-Fix: Lightweight Bit-Flip Detection and Rapid Recovery Framework for Language Models

LM-Fix:一种轻量级位翻转检测与快速恢复框架用于语言模型

Ahmad Tahmasivand, Noureldin Zahran, Saba Al-Sayouri, Mohammed Fouda, Khaled N. Khasawneh

机构 * The National Institutes of Health(美国国家卫生研究院)

专题命中 其他LLM :language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 LM-Fix通过轻量级检测和局部修复技术,有效提升大型语言模型在生产环境中的可靠性。

Comments Accepted at IEEE ICCD 2025. Code: https://github.com/ata990/lm-fix. Detects over 94 percent single-bit flips (near 100 percent multi-bit) with about 1 to 7.7 percent overhead; recovery is over 100x faster than a full reload. Keywords: LLMs, bit-flip, fault injection, reliability, security, Rowhammer, SDC, Jailbreaking, Attack, Defense, GPU DRAM faults

Journal ref Proc. IEEE Int. Conf. on Computer Design (ICCD), 2025, pp. 432-440

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21175 2026-02-25 cs.CV 78%

Seeing Through Words: Controlling Visual Retrieval Quality with Language Models

通过文字看见:利用语言模型控制视觉检索质量

Jianglin Lu, Simon Jenni, Kushal Kafle, Jing Shi, Handong Zhao, Yun Fu

机构 * Adobe Research(Adobe研究部) Northeastern University(东北大学)

专题命中 其他LLM :language model(title,abstract)

AI总结 通过生成语言模型扩展短查询并控制图像质量,提升视觉检索效果和可控性

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21091 2026-02-25 econ.GN q-fin.EC 75%

Can Interest-Bearing Positions Solve the Long-Horizon Problem in Prediction Markets?

带利息位置能否解决预测市场的长期限问题?

Caleb Maresca

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究带利息位置如何缓解预测市场长期事件的流动性与准确性问题,发现其显著提升市场参与度并减少定价偏差。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20493 2026-02-25 cs.NI cs.MA 75%

AWCP: A Workspace Delegation Protocol for Deep-Engagement Collaboration across Remote Agents

AWCP:用于远程代理深度协作的工件委托协议

Xiaohang Nie, Zihan Guo, Youliang Chen, Yuanjian Zhou, Weinan Zhang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 AWCP通过工件委托协议实现远程代理的深度协作,提供开源实现以提升代理间协作的互操作性。

Comments 16 pages, 7 figure, tech report of Agent Workspace Collaboration Protocol

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20284 2026-02-25 cs.SE 75%

PhantomRun: Auto Repair of Compilation Errors in Embedded Open Source Software

PhantomRun: 针对嵌入式开源软件编译错误的自动修复

Han Fu, Andreas Ermedahl, Sigrid Eldh, Kristian Wiklund, Philipp Haller, Cyrille Artho

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 PhantomRun利用大型语言模型自动修复嵌入式开源软件的编译错误,成功修复了45%的CI失败案例。

Comments 13 pages, 5 figures, Mining Software Repositories 2026 (MSR 2026) , Rio de Janeiro, Brazil, 13-14 April 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20656 2026-02-25 cs.DC 71%

Lagom: Unleashing the Power of Communication and Computation Overlapping for Distributed LLM Training

Lagom:释放通信与计算重叠在分布式大模型训练中的潜力

Guanbin Xu, ZhenGuo Xu, Yuzhe Li, Youhui Bai, Ping Gong, Chaoyi Ruan, Cheng Li

专题命中 其他LLM :LLM(title)

AI总结 Lagom通过统一成本模型和优先级搜索算法,优化分布式大模型训练中的通信与计算重叠,实现速度提升。

Comments 6 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21082 2026-02-25 cs.CL 70%

Beyond the Star Rating: A Scalable Framework for Aspect-Based Sentiment Analysis Using LLMs and Text Classification

超越星级评分:一种基于LLMs和文本分类的可扩展的基于方面的情感分析框架

Vishal Patil, Shree Vaishnavi Bacha, Revanth Yamani, Yidan Sun, Mayank Kejriwal

机构 * University of Southern California(南加州大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出一种结合LLMs和传统机器学习的方法,用于大规模基于方面的情感分析,通过分析470万条评论,展示了其在餐饮体验、菜系和地理区域中的应用效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00044 2026-02-25 cs.CY cs.AI 70%

When LLMs Imagine People: A Human-Centered Persona Brainstorm Audit for Bias and Fairness in Creative Applications

当LLMs想象人们:一种以人为中心的人格脑暴审计方法,用于创意应用中的偏见和公平性

Hongliu Cao, Eoin Thomas, Rodrigo Acuna Agost

机构 * Amadeus Nice France (2026)(法国Nice阿马德斯)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种人格脑暴审计方法,用于评估LLM在创意应用中的偏见和公平性,通过量化偏见并分析多个交集身份和社会角色,揭示模型生成中潜在的不公平现象。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00795 2026-02-25 cs.CV 67%

DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning

DVLA-RL:基于强化学习门控的双层视觉-语言对齐用于少样本学习

Wenhao Li, Xianjing Meng, Qiangchang Wang, Zhongyi Han, Zhibin Wu, Yilong Yin

机构 * Software School, Shandong University(山东大学软件学院) Shenzhen Loop Area Institute(深圳河套学院) School of Computing and Artificial Intelligence, Shandong University of Finance and Economics(山东财经大学计算机与人工智能学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 DVLA-RL通过双层语义构建和强化学习门控注意力,实现少样本学习中视觉与语言的双层次对齐,提升泛化能力。

Comments Accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17364 2026-02-25 cs.CL cs.AI 62%

Bridging Gaps in Natural Language Processing for Yorùbá: A Systematic Review of a Decade of Progress and Prospects

弥合自然语言处理在约鲁巴语中的差距:对十年进展与前景的系统综述

Toheeb Aduramomi Jimoh, Tabea De Wille, Nikola S. Nikolov

机构 * Department of Computer Science and Information Systems, University of Limerick(计算机科学与信息系统系,利默里克大学)

专题命中 其他LLM :language model(abstract);分类 cs.CL、cs.AI

AI总结 本文通过系统综述分析了近十年约鲁巴语NLP的发展,揭示了资源匮乏、声调复杂性等挑战,并总结了基于规则等主要技术,旨在推动该语言在NLP中的应用进展。

Journal ref Natural Language Processing Journal, 13 (2025), 100194

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20947 2026-02-25 cs.LG cs.CV 57%

Estimation of Confidence Bounds in Binary Classification using Wilson Score Kernel Density Estimation

利用Wilson分数核密度估计法估计二分类中的置信界

Thorbjørn Mosekjær Iversen, Zebin Duan, Frederik Hagelskjær

专题命中 其他LLM :foundation model(abstract);分类 cs.LG

AI总结 本文提出一种基于Wilson分数核密度估计的二分类置信界估计方法,适用于各种特征提取器,具有较低的计算复杂度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12402 2026-02-25 cs.LG math.OC stat.ML 57%

Cautious Weight Decay

谨慎的权重衰减

Lizhang Chen, Jonathan Li, Kaizhao Liang, Baiyu Su, Cong Xie, Nuo Wang Pierse, Chen Liang, Ni Lao, Qiang Liu

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 CWD是一种对优化器无关的权重衰减方法,通过仅对符号一致的参数应用衰减,提升语言模型和ImageNet任务的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏