Prompting Decision Transformer for Few-Shot Policy Generalization
专题命中 其他LLM :prompting(title);分类 cs.AI、cs.LG
Comments ICML 2022. Project page: https://mxu34.github.io/PromptDT/
AI 大模型
大语言模型、预训练、指令微调、后训练和语言模型应用。
专题命中 其他LLM :prompting(title);分类 cs.AI、cs.LG
Comments ICML 2022. Project page: https://mxu34.github.io/PromptDT/
专题命中 其他LLM :prompting(title);分类 cs.CL、cs.AI
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
Comments 15 pages
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
Comments Accepted to appear in the main conference of EMNLP 2021
专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI
Comments Accepted at IEEE Affective Computing and Intelligent Interaction (ACII) 2021
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
专题命中 其他LLM :prompting(title);分类 cs.AI、cs.LG
Comments 12 pages, 5 figures, The Web Conference 2020, ACM WWW 2020
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
Comments 21 pages
专题命中 其他LLM :language model(title);分类 cs.CL、cs.LG
机构 * School of Interactive Computing, College of Computing Georgia Institute of Technology, USA(交互计算学院、计算机学院、佐治亚理工学院)
专题命中 其他LLM :language model(title,journal_ref);分类 cs.CL
Journal ref Published at the Conference on Language Models (COLM) 2025
AutoSaddler:基于智能体执行轨迹的持久化更新的自动工具优化
机构 * KAIST(韩国科学技术院) ; Southern University of Science and Technology(南方科技大学) ; Microsoft(微软) ; POSTECH(浦项科技大学)
专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.CL、cs.AI、cs.LG
AI总结 AutoSaddler是基于智能体执行轨迹的自动工具优化框架,通过离线学习结合故障诊断等技术,在多个基准上提升智能体性能9.0-10.0个百分点,为构建更可靠的智能体系统提供了新方向。
Comments 44 pages, 15 figures. Project website and code: https://aka.ms/AutoSaddler-website
利用多模态大语言模型探索程序流程图在代码生成中的潜力
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)
AI总结 研究利用多模态大语言模型探索程序流程图对代码生成的潜力,通过从示例代码生成流程图并与问题陈述一起提供给模型进行代码生成实验,发现结合流程图可提升性能,不同细节程度的流程图效果有差异,还比较了与少样本学习的效果。
Comments 21 pages, Accepted at the 26th IEEE International Conference on Software Quality, Reliability, and Security (QRS 2026), Regular Papers Track
APPROVE:结合大型语言模型的面向视觉端用户的机器人编程
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)
AI总结 APPROVE是一种结合LLMs的多模态端用户机器人编程框架,通过Blockly可视化程序并支持用户确认、修改或拒绝,还可存储复用已确认功能,解决了现有系统透明度不足、意图对齐差、复用有限的问题,提升了用户信任与编程灵活性。
Comments Accepted for publication in Procedia CIRP, Proceedings of the 20th CIRP Conference on Intelligent Computation in Manufacturing Engineering (ICME 2026)
具有增强对抗样本迁移性的视角不变攻击
机构 * The Hong Kong Polytechnic University(香港理工大学)
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)
AI总结 针对对抗样本跨模型迁移性带来的安全威胁,提出视角不变攻击(PIA)及其扩展PIA-Mix,通过多自由度顶点采样策略提升对抗样本迁移性,实验显示其性能优于当前最优基于迁移的攻击方法。
Journal ref IEEE Transactions on Information Forensics and Security, vol. 21, pp. 6818-6831, 2026
AlphaSeek:面向多源金融数据的轨迹级自迭代因子挖掘框架
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)
AI总结 针对现有量化因子挖掘方法的缺陷,提出AlphaSeek框架,整合自动化方向发现等技术,在CSI300上取得优异策略性能,且因子具跨市场迁移性。
单轮文本提示模式分类法
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)
AI总结 提出一种单轮文本交互的提示模式分类法,通过可复现方法识别30种规范模式,按两个维度组织。
Comments 25 pages, 2 figures
估算AI使用产生的温室气体排放:企业级测量框架
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)
AI总结 针对企业AI排放核算缺乏公认方法的问题,提出标准化企业级AI排放核算框架,以支撑减排目标设定与脱碳决策。
学习选择视觉上下文示例
机构 * University of Cincinnati(辛辛那提大学) ; University of California, Los Angeles(加利福尼亚大学洛杉矶分校)
专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 本文提出LSD方法,通过强化学习构建最优演示集,提升多模态大语言模型在视觉回归任务中的表现,揭示了学习选择在视觉上下文学习中的必要性。
Comments 21 pages, 12 figure, accepted to Computer Vision and Pattern Recognition Conference (CVPR) 2026 Findings Track
Journal ref In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (pp. 9455-9465) 2026
Baikal:面向数据湖深度研究的结构化搜索框架
专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.CL、cs.AI、cs.LG
AI总结 Baikal是面向数据湖深度研究的结构化搜索框架,通过聚类证据为语义区域并自适应搜索,在HybridQA和TAT-QA数据湖上较基线方法提升报告得分28%和36%,展现了结构化语义探索的价值。
内部叙事参数化情感状态
机构 * Applied Computational Psychiatry Lab(应用计算精神病学实验室) ; Max Planck UCL Centre for Computational Psychiatry and Ageing Research(马克斯·普朗克UCL计算精神病学与衰老研究中心) ; Queen Square Institute of Neurology and Mental Health(圣夸克广场神经病学与心理健康研究所) ; Neuroscience Department(神经科学系) ; Division of Psychiatry(精神病学系)
专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 本文通过量化参与者内部叙事的大语言模型表示及其子空间,研究了叙事与情感状态之间的关系,发现特定症状的描述性思维能够预测标准化的抑郁评分,并强调保持症状间的协方差对构建效度至关重要。
基于语音AI的可扩展与个性化口头评估
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)
AI总结 本文提出利用语音AI进行个性化口头评估,通过Viva系统实现高效评估,降低成本,并发现需分解模块、约束行为、保持随机化、多模型评分等改进方法。
Comments 34 pages, 6 figures, 9 tables. Author's version of a paper published in Communications of the ACM
Journal ref Communications of the ACM (2026)
仅向ChatGPT提供审核规则是不够的:策略即提示的审核及其对社区治理的潜在影响
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)
AI总结 研究社交媒体内容审核模式变化,探讨“策略即提示”方法在内容审核中应用,阐述其技术与治理特性、局限性及风险,提出有效提示治理考量,指出仅编写提示不适用于确保社区治理。
Comments Accepted at the Mensch und Computer (MuC) 2026 Workshop "Human-Centered Content Moderation: Expertise, Context & Evaluation"
从图形用户界面测试到对话交互:特定应用语音助手的新视角
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)
AI总结 研究针对移动平台语音助手与应用行为契合度低的问题,提出由语言模型驱动、利用图形用户界面测试代码自动开发特定应用语音助手的方法,经安卓原型验证,可复用测试代码实现语音助手合成并指明相关研究方向。
通过安全感知工具描述减轻MCP服务器中的污点式漏洞
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract)
AI总结 研究MCP服务器污点式漏洞,提出SPELLSMITH,通过分析服务器高风险能力,结合工具描述等识别潜在风险,利用协议属性和大语言模型自我反思能力迭代优化输出,有效减轻污点式漏洞利用。
基于梯度的任意语音识别模型的语音到文本对齐:从连接主义时间分类到语音大语言模型
机构 * Machine Learning and Human Language Technology Group, RWTH Aachen University(机器学习与人类语言技术组,亚琛工业大学) ; AppTek GmbH(AppTek公司)
专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 研究基于梯度的任意可微语音识别模型的语音到文本对齐方法,通过对教师强制令牌对数概率取梯度并解码,无需训练、改模型及对齐头,适用于各模型家族,在多模型上评估,结果显示该方法能产生可用对齐,有优有劣。
数学中的技术转向
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract)
AI总结 探讨交互式定理证明器、自动定理证明器和大语言模型等技术对数学实践的影响,包括对数学知识创造与共享、社会层面如协作和信任关系的改变,以及引发的关于形式化本质和人机认知劳动分工等问题。
探测而非提示:用于多元RAG中元数据过滤的隐藏状态探测器
机构 * National University of Kyiv-Mohyla Academy(基辅莫希拉学院国家大学)
专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 研究改进多跳问答检索,用基于小型开源语言模型隐藏状态训练的确定性探测器取代专有提取器,探测器在准确率上有优势且输出空间固定,介绍了使其工作的设计选择及低成本输出方式。