arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

AAAI Conference on Artificial Intelligence · 会议 · Artificial Intelligence

至 收录 9567
2608.07592 2026-08-11 cs.HC cs.CY 新提交

"Always Want to Use it for Everything": Understanding Young Adults' Perceptions of AI Dependence

“总想用它做所有事”:理解年轻人对AI依赖的看法

Ashlee Milton, Leah Ajmani, Amy Heger, Forough Poursabzi-Sangdeh, Mihaela Vorvoreanu, Jina Suh

AI总结 本研究通过问卷收集18-25岁AI聊天机器人用户反馈,确定AI依赖的三个促成因素,指出其对年轻人的心理与发展影响,并提出相关启示。

Comments To appear in proceedings of the Ninth AAAI/ACM Conference on AI, Ethics, and Society (AIES '26)

URL PDF HTML 收藏
2608.05656 2026-08-11 cs.CY cs.AI cs.HC 版本更新

Studying People to Study AI: Expert Perspectives on the Epistemic Fit and Barriers of Human Research in AI Safety & Ethics

研究人以研究AI:关于人类研究在AI安全与伦理中的认知适配性及障碍的专家观点

Jessica Y. Bo, Paula Akemi Aoyagui, Shalaleh Rismani, Dipto Das, Syed Ishtiaque Ahmed, Ashton Anderson

AI总结 本研究通过对93名AI安全与伦理领域专家的调查及17名专家的访谈,分析了人类研究在AI安全与伦理领域的认知适配性与障碍,并提出相关建议以避免流于形式的“人类洗白”。

Comments Ninth AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026)

URL PDF HTML 收藏
2608.05602 2026-08-11 cs.AI cs.HC 版本更新

Epistemic Trustworthiness in Generative AI: A Normative Framework for Warranted Reliance in High-Stakes Workflows

生成式AI中的认知可信赖性:高风险工作流程中合理依赖的规范框架

Nimisha Karnatak, Max Van Kleek, Nigel Shadbolt

AI总结 本文针对生成式AI在高风险场景的合理依赖问题,提出含认知谦逊、认知可及性、抗认知不公三条件的规范框架,通过案例分析指出现有标准的不足,为GenAI设计评估提供新方向。

Comments Accepted at AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026)

URL PDF HTML 收藏
2606.30412 2026-08-11 cs.CY cs.AI 版本更新

Can LLMs Rank? A Tale of Triads and Triage

LLM能排序吗?三元组与分诊的故事

Gaurab Pokharel, Shafkat Farabi, Patrick J. Fowler, Sanmay Das

机构 * Virginia Tech(弗吉尼亚理工大学) Washington University in St. Louis(圣路易斯华盛顿大学)

AI总结 研究LLM作为排序裁判的可靠性,提出用循环三元组计数和排序距离度量一致性,并在无家可归者住房分配和急诊分诊任务中验证。

Comments Accepted to the Ninth AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026)

URL PDF HTML 收藏
2601.08856 2026-08-11 cs.SE cs.AI 版本更新

LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns

LAUDE: 基于LLM的硬件设计单元测试生成与调试

Deeksha Nandal, Riccardo Revalor, Soham Dan, Debjit Pal

机构 * Dept. of Electrical and Computer Engineering, University of Illinois Chicago(伊利诺伊大学芝加哥分校电子与计算机工程系) Microsoft(微软公司)

AI总结 LAUDE利用大语言模型能力,实现硬件设计的单元测试生成与调试,有效检测和修复设计中的错误。

Comments 11 Pages, 9 Figures, 9 Tables Submitted to AAAI 2027

URL PDF HTML 收藏
2410.12265 2026-08-11 cs.CL

Auto-PRE: An Automatic and Cost-Efficient Peer-Review Framework for Language Generation Evaluation

Junjie Chen, Weihang Su, Zhumin Chu, Haitao Li, Yujia Zhou, Dingbo Yuan, Xudong Wang, Jun Zhou, Yiqun Liu, Min Zhang, Shaoping Ma, Qingyao Ai

Comments AAAI 2026

URL PDF HTML 收藏
2509.06586 2026-08-11 cs.CY 版本更新

Simulating Dispute Mediation with LLM-Based Agents for Legal Research

基于大语言模型智能体的争议调解模拟:面向法律研究

Junjie Chen, Haitao Li, Minghao Qin, Yujia Zhou, Yanxue Ren, Wuyue Wang, Yiqun Liu, Yueyue Wu, Qingyao Ai

AI总结 研究针对法律争议调解实证研究的局限,提出首个基于LLM的AgentMediation框架模拟争议调解,通过可控实验发现符合社会学理论的模式,为社科与AI在法律研究的融合提供平台。

Comments AAAI 2026

URL PDF HTML 收藏
2608.07434 2026-08-10 cs.CV cs.IR 新提交

Conformal Coverage Guarantees for Any Video Temporal Grounder

任意视频时间定位器的共形覆盖保证

Aseel Mohamed, Rasul Khanbayov, Erchin Serpedin, Hasan Kurban

AI总结 针对视频时间定位器的模糊性问题,提出与模型无关的COVER包装器,可保证输出区域以至少1-α的概率包含真实时刻,在多基准和定位器上验证了其覆盖度符合目标值。

Comments Submitted to AAAI 2027

URL PDF HTML 收藏
2608.07367 2026-08-10 cs.AI 新提交

People Are Not Just Their Countries. Disentangling Social Determinants of LLM Value Alignment Across Europe

人不只是其所属国家:在欧洲范围内解耦大型语言模型(LLM)价值对齐的社会决定因素

Maria-Louisa Wightman, Guillaume Bied, Tijl De Bie

AI总结 本研究依托欧洲社会调查,发现LLMs与不同社会人口学群体的价值观对齐存在差异,国籍作为单独变量的解释力与全部社会人口学变量相当,国家与社会人口学因素在解释对齐模式上互补。

Comments Accepted at AIES 2026 (9th AAAI/ACM Conference on AI, Ethics, and Society)

URL PDF HTML 收藏
2608.07297 2026-08-10 cs.CY 新提交

Data Annotation as Measurement

数据标注即测量

Emma Harvey, Allison Koenecke, Rene F. Kizilcec

AI总结 本文将数据标注重新定义为测量,通过文献综述和访谈开发标注问题诊断框架,明确标注问题来源并给出超越一致性的质量评估方法,为提升AI标注数据质量提供概念基础。

Comments AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026)

URL PDF HTML 收藏
2608.06908 2026-08-10 cs.CL cs.AI cs.CY cs.LG 新提交

Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests

针对各向异性校准WEAT:将ZCA白化作为嵌入关联测试的几何预处理步骤

Seitaro Ono, Senna Ross, Jun Saiki

AI总结 本研究提出将ZCA白化作为预处理步骤校准WEAT,可降低嵌入空间各向异性,超30%WEAT结果显著性改变,提升语义相似度,为相关偏差测量提供更可靠基础。

Comments Extended version (with appendices) of a paper accepted at the 9th AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026)

URL PDF HTML 收藏
2605.16336 2026-08-10 cs.CR cs.AI cs.CY 版本更新

On Seeding Watermarks to Detect Verbatim LLM Copy-Paste Responses

检测作业中的直接LLM复制粘贴

Aizierjiang Aiersilan, Artin Yousefi, Robert Pless

机构 * The George Washington University(乔治·华盛顿大学)

AI总结 本文提出SteganoPrompt工具,通过在作业提示中嵌入隐形指令,使LLM在响应时生成特征签名,帮助教师检测学生直接复制粘贴模型回复的行为。

Comments We are pleased to announce that this paper has been accepted by the 9th AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026). We appreciate the valuable feedback from the reviewers and look forward to sharing our findings with the community

URL PDF HTML 收藏
2502.20295 2026-08-10 cs.LG cs.AI cs.CV 版本更新

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

通过封面判断书籍:调查多模态大语言模型用于多页手写文档转录

Benjamin Gutteridge, Matthew Thomas Jackson, Toni Kukurin, Xiaowen Dong

机构 * University of Oxford(牛津大学) QuantCo

AI总结 本文研究多模态大语言模型在多页手写文档转录中的应用,提出OCR+PAGE-1和OCR+PAGE-N策略,通过共享页面内容提升转录效果。

Comments 10 pages (36 including references and appendices), 11 figures, accepted at COLM 2026, earlier version accepted at AAAI 2025 Workshop on Document Understanding and Intelligence

URL PDF HTML 收藏
2511.08322 2026-08-10 cs.CV cs.LG

Mitigating Negative Flips via Margin Preserving Training

Simone Ricci, Niccolò Biondi, Federico Pernici, Alberto Del Bimbo

Comments Accepted at AAAI2026

Journal ref Proceedings of the AAAI Conference on Artificial Intelligence, 40(11), 8721-8730, 2026

URL PDF HTML 收藏
1803.10888 2026-08-10 stat.ML cs.LG econ.EM 交叉投稿

An Empirical Analysis of Constrained Support Vector Quantile Regression for Nonparametric Probabilistic Forecasting of Wind Power

风电非参数概率预测的带约束支持向量分位数回归的实证分析

Kostas Hatalis, Shalinee Kishore, Katya Scheinberg, Alberto Lamadrid

AI总结 本文结合支持向量机、非线性分位数回归与非交叉约束提出风电非参数概率预测方法,基于全球能源预测竞赛2014年数据验证,该方法性能优于三个基准模型且避免分位数估计重叠问题。

Comments Originally published at The AAAI-17 Workshop on Artificial Intelligence for Smart Grids and Smart Buildings

Journal ref Thirty-First AAAI Conference on Artificial Intelligence, 2017

URL PDF HTML 收藏
2608.06106 2026-08-07 cs.CY 新提交

The Algorithmic Flattening of Sound: Computational Evidence and Justice Implications of AI Music Homogenization

声音的算法扁平化:AI音乐同质化的计算证据与正义意涵

Zoe Slendebroek, Danaé Metaxa

AI总结 本文审计Suno和Lyria 3在四类音乐中的同质化趋势,发现二者存在不同的同质化模式,AI与人类曲目可被MIR特征近乎完美区分,该现象关乎音乐风格的可识别性与经济回报等正义问题。

Comments To be published in AAAI/ACM AIES 2026

URL PDF HTML 收藏
2608.05238 2026-08-07 cs.LG 新提交

Decoupling Perception from Description: Computation-Grounded Representation Alignment between Multivariate Time Series and Language

感知与描述解耦:多变量时间序列与语言间的计算基础表示对齐

Xinran Feng, Yi Xie, Chao Zhang, Ruikun Li, Wanyun Ling, Ziyue Li, Chenxi Liu

AI总结 该研究针对多模态时间序列语言对齐的三难困境,提出CGTime模型,通过计算处理感知、LLM处理描述,在多变量理解任务上优于更大的通用模型。

Comments 46 pages, 9 figures, including supplementary material. Submitted to AAAI 2027. Xinran Feng and Yi Xie contributed equally

URL PDF HTML 收藏
2608.05008 2026-08-06 cs.CY 新提交

The Beginning of ChatGPT Ads

ChatGPT 广告的开端

Emma Lurie, Ro Encarnación, Sorelle A. Friedler, Danaé Metaxa

AI总结 本文首次实证研究 ChatGPT 广告,采用傀儡审计方法,发现低收入账号更易收到广告,发布了广告存档并提出未来研究建议。

Comments To be published in AAAI/ACM AIES 2026

URL PDF HTML 收藏
2608.04365 2026-08-06 cs.LG cs.CR cs.CY 新提交

Manipulation-Proof Oblivious Audits against Deceptive Model Providers

针对欺骗性模型提供者的防操纵遗忘审计

Augustin Godinot, Sofiane Azogagh, Julien Ferry, Sébastien Gambs

AI总结 本文提出新型防操纵遗忘审计协议,利用私有信息检索机制提升审计对模型提供者操纵行为的可检测性,实验验证其有效实用。

Comments This work has been accepted for publication at the 2026 AAAI/ACM Conference on AI, Ethics, and Society (AIES)

URL PDF HTML 收藏
2608.01742 2026-08-06 cs.AI cs.CL 版本更新

MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents

MemSIF:从结构化交互到面向大语言模型智能体的双轨事实记忆

YuFei Luo, Xiucheng Xu, Zhen Yang

AI总结 该研究针对大语言模型智能体的长期记忆问题,提出MemSIF框架,通过结构化交互记忆与双轨事实记忆缓解两种错位模式,在两个数据集上的五种主干模型中均取得最高总准确率。

Comments Submitted to AAAI 2027. 19 pages, 10 figures, 18 tables

URL PDF HTML 收藏
2605.28566 2026-08-06 cs.AI cs.LG 版本更新

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

思维树作为经典启发式搜索问题:形式化基础与设计模式

Guni Sharon

机构 * Guni Sharon

AI总结 本文通过经典启发式搜索术语统一分类法,将基于LLM的推理映射到搜索组件,并识别出系统搜索和前瞻性策略两种设计模式。

Comments Extended version of the SoCS 2026 paper. Includes appendices omitted from the proceedings version

Journal ref Proceedings of the Nineteenth International Symposium on Combinatorial Search (SoCS 2026), AAAI Press, 2026

URL PDF HTML 收藏
2508.13661 2026-08-06 cs.LG cs.MA 版本更新

Communication-Enhanced Tutoring for Efficient Decentralized Multi-Agent Reinforcement Learning

MACTAS: 基于自注意力的多智能体强化学习中智能体间通信方法 with 动作-价值函数分解

Maciej Wojtala, Bogusz Stefańczyk, Dominik Bogucki, Łukasz Lepak, Paweł Wawrzyński

机构 * University of Warsaw(华沙大学) IDEAS Research Institute(IDEAS研究机构) Institute of Fundamental Technological Research, Polish Academy of Sciences(波兰科学院基础技术研究所) IDEAS NCBR Warsaw University of Technology(华沙技术大学)

AI总结 MACTAS通过自注意力机制实现多智能体强化学习中的智能体间通信,结合动作-价值函数分解,提供可微通信方法,实现大规模系统的高效训练与高性能表现。

Comments Submitted for AAAI 2027

URL PDF HTML 收藏
2511.07322 2026-08-06 cs.CL cs.AI 版本更新

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

FinRpt:用于权益研究报告生成的数据集、评估系统及基于大语言模型的多智能体框架

Song Jin, Shuqi Li, Shukun Zhang, Rui Yan

AI总结 本文首次明确权益研究报告生成任务,推出含数据集、评估系统的FinRpt基准,提出多智能体框架FinRpt-Gen,实验验证其有效性,相关代码与数据集公开。

Comments AAAI 2026

URL PDF HTML 收藏
2608.03171 2026-08-05 cs.GT cs.AI 新提交

EFX Allocation In (Multi)Hypergraphs

(多)超图中的EFX分配

Thanasis Lianeas, Alkmini Sgouritsa, Minas Marios Sotiriou

AI总结 该研究在(多超)图设置下,证明围长至少为4的超图(主体具一般单调估值)及满足特定条件的多超图中,总存在EFX分配,分别可在多项式和伪多项式时间构造。

Comments 15 pages, accepted in AAAI 2026

URL PDF HTML 收藏
2608.02699 2026-08-05 cs.AI cs.CY cs.LG 新提交

Explainable AI for the EU Right to Explanation: A Systematic Review of the Law-XAI Translation Gap

面向欧盟解释权的可解释人工智能:法律与可解释人工智能之间翻译差距的系统综述

Benjamin Fresz, Elena Dubovitskaya, Marco F. Huber

AI总结 本文针对欧盟解释权开展XAI系统综述,发现法律与技术视角整合不足,提出收件人/目的框架与四阶段操作蓝图,指出需解决相关研究问题以保障解释权落地。

Comments Accepted at the 9th AAAI/ACM Conference on AI, Ethics and Society (AIES-26)

URL PDF HTML 收藏
2608.02660 2026-08-05 cs.CY cs.AI cs.HC 新提交

AI Alignment and Fiduciary Obligation

AI对齐与受托义务

Benjamin Lange

AI总结 本文提出将受托理论应用于长期AI助手部署,以开发者对用户负有的忠诚、注意、善意和坦诚四项受托义务为基础构建AI对齐标准,为AI对齐研究提供新视角。

Comments 10 pages, 1 table. Accepted at the 9th AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026)

URL PDF HTML 收藏
2608.00123 2026-08-05 cs.CL cs.AI cs.GT cs.LG 版本更新

LLM-OSDA: An Optimal-Stopping Dynamic Auction for Native Advertising in Multi-Turn LLM Conversations

LLM-OSDA:多轮大语言模型对话中原生广告的最优停止动态拍卖机制

Yan Fang, Jialin Chen, Chun Gan, Hang Yu, Mingjun Nie, Yeyu Zhang, Fengxiang He, Ching Law

AI总结 针对多轮LLM对话中原生广告的时机与分配耦合问题,提出LLM-OSDA动态拍卖机制,结合贝尔曼最优停止等,在模拟实验中使净收益提升11%且用户留存相当。

Comments 14 pages, 7 figures. Submitted to the 41st AAAI Conference on Artificial Intelligence (AAAI 2027)

URL PDF HTML 收藏
2511.13300 2026-08-05 eess.AS cs.AI cs.SD 交叉投稿

PASE: Leveraging the Phonological Prior of WavLM for Low-Hallucination Generative Speech Enhancement

PASE:利用WavLM的语音学先验实现低幻觉生成式语音增强

Xiaobin Rong, Qinwen Hu, Mansur Yesilbursa, Kamil Wojcicki, Jing Lu

AI总结 本研究针对生成式语音增强的幻觉问题,提出PASE框架,通过适配WavLM为去噪专家、双流声码器训练,实现低幻觉的语音增强,性能优于现有最优模型。

Comments Accepted by AAAI 2026

URL PDF HTML 收藏
2608.02391 2026-08-04 cs.AI cs.LG 新提交

Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training

资源受限智能体大语言模型后训练的协同协同进化方法

Zhiyuan Wang, Shengcai Liu, Jiahao Wu, Ning Lu, Hui Ouyang, Shaofeng Zhang, Haoze Lv, Ke Tang

AI总结 针对资源受限智能体LLM后训练的高内存、长耗时问题,提出CoPES方法,在GPU小时预算下,其验证准确率恢复率、内存效率均优于标准ES和LoRA-GRPO,在多任务基准上表现更优。

Comments 14 pages,9 figures, submit to AAAI 2027

URL PDF HTML 收藏
2608.00106 2026-08-04 cs.LG 新提交

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

面向智能体工作流的组合式元路由学习:一个可执行基准

Natan Vidra, Alina Kapanova, Arun Kanhai, Spurthi Setty

AI总结 本文提出一个智能体工作流的可执行基准和感知预算的组合式元路由器,在保留测试集上实现100%成功率,成本比静态策略低43%,但在词汇偏移挑战任务上表现不佳,凸显词汇泛化是主要限制。

Comments 7 pages, 1 figure; AAAI 2027 anonymous submission

URL PDF HTML 收藏