Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments
扩展搜索相关性:用LLM生成的判断增强应用商店排名
机构 * Apple(苹果公司)
专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG
AI总结 针对应用商店排名中专家文本相关性标签稀缺的问题,通过微调LLM生成数百万标签,结合行为相关性优化排序器,显著提升Pareto前沿和转化率。
AI 大模型
大语言模型、预训练、指令微调、后训练和语言模型应用。
扩展搜索相关性:用LLM生成的判断增强应用商店排名
机构 * Apple(苹果公司)
专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG
AI总结 针对应用商店排名中专家文本相关性标签稀缺的问题,通过微调LLM生成数百万标签,结合行为相关性优化排序器,显著提升Pareto前沿和转化率。
人类心理测量问卷误判LLM行为
机构 * Graduate School of Data Science, Seoul National University(首尔国立大学数据科学研究生院) ; Department of Communication, Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学通信系人工智能交叉学科项目)
专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.AI
AI总结 通过比较LLM在Likert问卷和生成概率上的价值与人格特征,发现问卷存在系统性偏差,提出基于生成概率的评估方法更准确。
Comments 38 pages, 6 figures
主动型智能体真的需要LLM来决定何时唤醒和锚定什么吗?
机构 * Purdue University(普渡大学) ; Microsoft(微软) ; Michigan State University(密歇根州立大学) ; Georgia Institute of Technology(佐治亚理工学院)
专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.AI
AI总结 提出用时间图学习(TGL)模型替代LLM作为主动智能体的触发器,通过图更新而非文本处理用户活动,实现高效、低延迟的触发决策。
Comments 31 pages, 5 figures, 7 tables
先验知识还是搜索?LLM代理在硬件感知代码优化中的研究
机构 * Applied AI Institute(应用人工智能研究所) ; ITMO University(ITMO大学) ; AI Talent Hub(AI人才中心)
专题命中 其他LLM :LLM(title,title_cn);分类 cs.AI、cs.LG
AI总结 该研究探讨了在硬件感知代码优化中,LLM代理是依赖于先验知识还是搜索过程,通过三个受控实验发现LLM在纯黑盒优化中表现为贪婪优化器,在零样本内核生成中输入大小信息无明显影响,而在反馈循环内核优化中CUDA单调改进而TVM IR主动退化,表明LLM在代码优化任务中高度依赖预训练先验而非反馈或代理结构。
用于文本条件回归的LLM流程
机构 * Secondmind ; Wayve
专题命中 其他LLM :LLM(title,title_cn);分类 cs.CL、cs.LG
AI总结 本文提出结合轻量级神经过程与边际LLM预测,以解决文本条件回归中的误差累积问题,提升预测校准和局部一致性。
具有增强步骤级转换的分层强化学习用于大语言模型代理
机构 * Beijing University of Posts and Telecommunications(北京邮电大学) ; China Mobile Group Design Institute Co., Ltd(中国移动集团设计院有限公司) ; Singapore Management University(新加坡管理大学)
专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
AI总结 本文提出STEP-HRL框架,通过仅基于单步转换而非完整交互历史实现步骤级学习,提升LLM代理在复杂任务中的性能和泛化能力,同时减少token使用。
Comments Accepted to ACL 2026 Main Conference
机构 * Algorithms and Complexity Group TU Wien(算法与复杂性组维也纳技术大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG
Comments 23 page; deeper analysis of streamliners and statistics about benchmark instances added
Journal ref F. Voboril, V. P. Ramaswamy, S. Szeider, Generating Streamlining Constraints with Large Language Models, Journal of Artificial Intelligence Research, volume 84, pages 16:1-16:19, 2025
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
Comments ICLR 2024 Workshop on Secure and Trustworthy Large Language Models
风力涡轮机维护日志标注框架:基于LLM驱动的数据校正与语义提取的可靠性智能增强
机构 * Institute for Energy Systems, School of Engineering, The University of Edinburgh(能源系统研究所,工程学院,爱丁堡大学) ; Nadara, Lisbon, Portugal(纳达拉,里斯本,葡萄牙)
专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL
AI总结 提出一种利用大语言模型自动标准化和结构化风力涡轮机维护日志的方法,通过纠正系统代码、提取故障模式与维护动作分类,将非结构化文本转化为定量可靠性指标。
Comments An adjustable template containing the Python script architecture, applied dynamic prompts, and data schemas is hosted in an open-source GitHub repository: https://github.com/mvmalyi/llm-driven-wind-turbine-maintenance-log-labelling
干扰网络:通过观点扰动和大语言模型放大社会分歧
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 该研究基于Friedkin-Johnsen模型,利用强化学习框架微调大语言模型生成干扰性文本,实现了对社交网络的战略性干扰,其结果对内容审核等领域有重要参考价值。
CausalRepair:通过双重切片弥合基于大语言模型的自动化程序修复中的因果差距
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 CausalRepair是基于最小因果上下文的对话驱动型自动化程序修复框架,通过双重切片策略构建因果相关上下文,在Defects4J数据集上修复313个漏洞,性能优于现有方法且修复成本更低。
大型语言模型(LLM)比专家自身更擅长解释专家
专题命中 其他LLM :LLM(title_cn,abstract);large language model(title);language model(title)
AI总结 该研究表明,大型语言模型可从专家行为中外部化隐性知识,提升决策质量并帮助新手接近专家表现,为波兰尼悖论提供实证支持,凸显其作为克服专家表述瓶颈的可扩展工具的潜力。
大语言模型在老年人中的反论点:认知卸载还是对道德说服的脆弱性?
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 研究发现大语言模型能显著影响老年人和年轻人的道德判断,老年人更易受说服,认知功能较低者在情感冲突困境中更易接受反论点,表明LLMs可能成为认知卸载工具,但也对认知脆弱者构成风险。
Comments This paper has been published in Computers in Human Behavior. The final published version is available at https://doi.org/10.1016/j.chb.2026.109142
WIP:Chat-Debugging:大型语言模型作为硬件调试助手
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 该研究提出将大型语言模型应用于硬件调试的Chat-Debugging方法,通过人机交互提升电气类学生的调试信心与技能,填补了物理硬件调试辅助工具的研究空白。
Comments This is the accepted version of a paper accepted for presentation at the 2026 IEEE Frontiers in Education Conference (FIE). The final version will be available via IEEE Xplore at: https://ieeexplore.ieee.org/Xplore/home.jsp
结合种群自进化大语言模型的动态模糊作业车间调度问题自动编程
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 该研究针对动态模糊作业车间调度问题,提出种群自进化框架结合大语言模型自动设计启发式调度规则,性能优于多种现有方法。
Comments 13 pages, 10 figures. Accepted for publication in IEEE Transactions on Fuzzy Systems
大规模的亲社会说服?大语言模型在不同个性化水平上的捐赠呼吁中优于人类
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 研究探讨大语言模型生成的捐赠呼吁在不同个性化水平下的有效性,发现其在捐款数额、参与度和说服力上均优于人类创作内容,但虚假个性化会带来负面影响。
一种使用大语言模型生成基于对话的语音合成课程的半自动系统:教育潜力的探索性研究
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 研究提出用大语言模型和文本转语音技术生成基于对话课程的半自动系统,经准实验探索其教育潜力。系统通过三阶段工作流程增强教育工作者,引入新方法。实验表明对话TTS在理解等方面优于单声道TTS,为TTS音频教育可接受性及课程形式设计提供依据。
从令牌到状态:LLM作为世界模型的特例及其连续路径
机构 * Paul Dubois(保罗·杜博伊斯)
专题命中 其他LLM :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 本文论证LLM是世界模型的退化特例,并提出从NTP到JEPA的连续谱系,逐步放松LLM约束,同时探讨其可扩展性挑战。
Comments 10 pages, 6 figures, 1 table
大型语言模型的投资者风险画像
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 本研究通过标准化风险问卷评估GPT、Gemini和Llama三种大型语言模型的风险偏好,发现它们普遍为长期投资者但风险容忍度不同,且赋予特定人物角色后各模型会调整其风险画像。
Comments Poster presented at the AI for Finance Symposium '25, The 6th ACM International Conference on AI in Finance (ICAIF '25)
基于大语言模型的自动安全合规线性时序逻辑生成:一种自监督框架
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)
AI总结 本文提出AutoSafeLTL框架,利用大语言模型自动生成符合安全限制的LTL规范,通过语言包含检查与自动反例引导修改机制确保逻辑一致性和语义准确性。
基于大语言模型的开源系统在现实中的应用:对GitHub安全通告中漏洞的实证研究
专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract,comments);language model(abstract,comments)
AI总结 本研究分析了295份GitHub安全通告,发现大多数漏洞映射到已知CWE,但模型中介暴露不足,建议结合CWE和OWASP视角更全面地评估LLM集成系统的漏洞。
Comments The 2nd International Workshop on Large Language Model Supply Chain Analysis (LLMSC 2026)
SPRIG:通过系统提示优化提升大语言模型性能
机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; University of Michigan(密歇根大学) ; University of Illinois Chicago(伊利诺伊大学芝加哥分校) ; LG AI Research(LG AI 研究院)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 本文提出SPRIG,一种基于编辑的遗传算法,通过优化通用提示提升大语言模型性能,发现系统提示与任务提示结合可进一步提升效果。
Comments Accepted at ICLR 2026
通过多智能体模拟探索大型语言模型中对自闭症的隐含视角
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)
AI总结 通过多智能体模拟研究大型语言模型对自闭症的隐含视角,揭示其偏见并提出双重共情问题以改进与自闭症人群的互动。
受脑启发的大型语言模型中功能网络和关键神经元探索
机构 * School of Automation, Northwestern Polytechnical University, Xi’an, China(自动化学院,西北工业大学,西安,中国) ; School of Computing, University of Georgia, Athens, USA(计算机学院,佐治亚大学,亚特兰大,美国) ; School of Physics and Information Technology, Shaanxi Normal University, Xi’an, China(物理与信息技术学院,陕西师范大学,西安,中国)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 本文受脑启发,探索LLM中的功能网络和关键神经元,发现这些网络对模型性能至关重要,通过抑制或增强网络活动可影响模型整体表现或特定任务效果。
Comments 21 pages, 18 figures
SQLBarber: 借助大型语言模型生成定制化和真实SQL工作负载的系统
机构 * Cornell University(康奈尔大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 SQLBarber利用大型语言模型生成定制化且真实的SQL工作负载,通过声明式接口和贝叶斯优化器高效生成符合目标成本分布的查询。
Comments Accepted by SIGMOD 2026; extended version with appendix
机构 * University of Texas at Dallas(德克萨斯大学达拉斯分校) ; UIUC(伊利诺伊大学)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
机构 * Singapore University of Technology and Design, Singapore(新加坡科技设计大学) ; Royal Holloway, University of London, UK(伦敦大学皇家霍洛威学院)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
Comments 12 pages, 4 figures, 8 tables, Accepted at EMNLP 2024 Findings
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
Comments 8 pages, 4 figures
Journal ref Proc. Natl. Acad. Sci. U.S.A. 122 (31) e2415697122 (2025)
机构 * University of Science and Technology of China(中国科学技术大学) ; State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室) ; Gaoling School of Artificial Intelligence(光明人工智能学院) ; Renmin University of China(中国人民大学) ; Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理重点实验室) ; Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程技术研究中心) ; Tsinghua University(清华大学) ; Huawei Technologies Co. Ltd(华为技术有限公司)
专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG
Comments 25 pages, 14 figures. Accepted by ACL 2025 (main conference)
专题命中 其他LLM :LLM(title,abstract);prompting(title,abstract);large language model(abstract);language model(abstract)
Comments Accepted by CHI'25