arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-01-01 至 2026-01-01 共收录 101 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 多智能体 17 篇

2512.24885 2026-01-01 cs.CL cs.GT cs.MA 57%

BEDA: Belief Estimation as Probabilistic Constraints for Performing Strategic Dialogue Acts

BEDA:将信念估计作为概率约束以执行战略对话行为

Hengli Li, Zhaoxin Yu, Qi Shen, Chenxi Li, Mengmeng Wang, Tinglang Wu, Yipeng Kang, Yuxuan Wang, Song-Chun Zhu, Zixia Jia, Zilong Zheng

机构 * Institute for Artificial Intelligence, PKU(北京大学人工智能研究所) NLCo, BIGAI(NLCo和BIGAI) Institute of Automation, CAS(中国科学院自动化研究所) School of Artificial Intelligence, BUPT(北京邮电大学人工智能学院) Department of Automation, THU(清华大学自动化系) Yuanpei College, PKU(北京大学元培学院)

专题命中 多智能体 :agent(abstract);分类 cs.CL

AI总结 BEDA通过将信念估计作为概率约束,有效提升了战略对话行为的执行效果,其在多个任务中均优于现有基线方法。

Comments Accepted by AAMAS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24383 2026-01-01 math.AP math.PR 50%

Mean-Field Limits of Deterministic and Stochastic Flocking Models with Nonlinear Velocity Alignment

确定性和随机 flocking 模型的均场极限:非线性速度对齐

Vinh Nguyen, Roman Shvydkoy, Changhui Tan

专题命中 多智能体 :agent(abstract)

AI总结 本文研究了具有非线性速度对齐的 flocking 模型的均场极限,扩展了 Cucker-Smale 理论到非线性框架,并在确定性和随机性情况下提供了收敛率的改进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20391 2026-01-01 math.OC cs.RO cs.SY eess.SY 50%

Contingency Model-based Control (CMC) for Communicationless Cooperative Collision Avoidance in Robot Swarms

基于应急模型的控制(CMC)用于机器人群通信less的协作避障

Georg Schildbach

专题命中 多智能体 :agent(abstract)

AI总结 本文提出了一种无需通信的分布式协作避障方法CMC,通过预先设计的共识规则和应急轨迹确保机器人群在受限环境中的安全运行。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 工作流自动化 7 篇

2512.23737 2026-01-01 cs.DC cs.LG 85%

Governing Cloud Data Pipelines with Agentic AI

用智能体AI治理云数据管道

Aswathnarayan Muthukrishnan Kirubakaran, Adithya Parthasarathy, Nitin Saksena, Ram Sekhar Bodala, Akshay Deshpande, Suhas Malempati, Shiva Carimireddy, Abhirup Mazumder

机构 * IEEE Senior Member, USA(IEEE高级会员,美国) Independent Researcher, USA(独立研究者,美国) Albertsons, USA(Albertsons公司,美国) Amtrak, USA(Amtrak公司,美国) Cato, USA(Cato公司,美国)

专题命中 工作流自动化 :agentic(title,abstract);agent(abstract);AI agent(abstract);分类 cs.LG

AI总结 本文提出Agentic Cloud Data Engineering,通过智能体AI实现云数据管道的治理,显著提升恢复效率、降低成本并减少人工干预。

Comments https://www.ijcstjournal.org/volume-13/issue-6/IJCST-V13I6P44.pdf

Journal ref International Journal of Computer Science Trends and Technology (IJCST), Volume 13 Issue 6,Dec 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24289 2026-01-01 cs.CL 70%

Automated Analysis of Sustainability Reports: Using Large Language Models for the Extraction and Prediction of EU Taxonomy-Compliant KPIs

欧盟可持续性报告的自动化分析:利用大型语言模型提取和预测符合欧盟分类法的KPIs

Jonathan Schmoll, Adam Jatowt

机构 * University of Innsbruck(因斯布鲁克大学)

专题命中 工作流自动化 :workflow(abstract);agentic(abstract);分类 cs.CL

AI总结 本文提出一个结构化数据集,利用大型语言模型评估欧盟分类法合规性,发现模型在预测财务KPIs上表现不佳,但能辅助识别经济活动。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24997 2026-01-01 cs.CL cs.AI 62%

Classifying long legal documents using short random chunks

使用短随机片段对长法律文件进行分类

Luis Adrián Cabrera-Diego

机构 * Jus Mundi

专题命中 工作流自动化 :workflow(abstract);分类 cs.AI、cs.CL

AI总结 本文提出基于DeBERTa V3和LSTM的法律文件分类器,通过随机片段输入和高效部署流程实现高分类精度与处理效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24511 2026-01-01 cs.DC 50%

Understanding LLM Checkpoint/Restore I/O Strategies and Patterns

理解LLM检查点/恢复I/O策略和模式

Mikaila J. Gossman, Avinash Maurya, Bogdan Nicolae, Jon C. Calhoun

专题命中 工作流自动化 :workflow(abstract)

AI总结 本文研究了LLM检查点/恢复I/O策略,通过微基准测试发现聚合和合并策略能显著提升写入吞吐量,比现有方法提高3.9至7.6倍。

Comments SCA/HPCAsia 2026 Workshops: Supercomputing Asia and International Conference on High Performance Computing in the Asia Pacific Region Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24371 2026-01-01 q-fin.MF q-fin.PM q-fin.RM 50%

Utility Maximisation with Model-independent Constraints

在无模型约束下的效用最大化

Alexander M. G. Cox, Daniel Hernandez-Hernandez

专题命中 工作流自动化 :agent(abstract)

AI总结 本文研究了在无模型约束下,代理人如何在最大化效用的同时保持投资组合估值不低于阈值,并通过数学方法分析了完全市场和Black-Scholes-Merton模型中的最优投资策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18737 2026-01-01 q-bio.QM eess.IV 50%

Photon Absorption Remote Sensing Virtual Histopathology: A Preliminary Exploration of Diagnostic Equivalence to Gold-Standard H&E Staining in Skin Cancer Excisional Biopsies

光子吸收远程传感虚拟组织病理学:皮肤癌切除活检中与金标准H&E染色诊断等效的初步探索

Benjamin R. Ecclestone, James E. D. Tweel, Marie Abi Daoud, Hager Gaouda, Deepak Dinakaran, Michael P. Wallace, Ally Khan Somani, Gilbert Bigras, John R. Mackey, Parsin Haji Reza

专题命中 工作流自动化 :workflow(abstract)

AI总结 PARS虚拟H&E成像在皮肤癌切除活检中实现了与化学H&E染色等效的诊断性能,为无标记组织学分析提供了新方法。

Comments 19 pages, 3 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10370 2026-01-01 cond-mat.mtrl-sci cond-mat.mes-hall 50%

Kinetically accessible 1D magnetic chains of transition-metal chalcogenides and halides on van der Waals surfaces

过渡金属硫化物和卤化物在范德华表面上的1D磁链的动能可及性

Canbo Zong, Deping Guo, Renhong Wang, Weihan Zhang, Jiaqi Dai, Zhongqin Zhang, Cong Wang, Xianghua Kong, Fei Pang, Zhihai Cheng, Zhong-Yi Lu, Wei Ji

专题命中 工作流自动化 :workflow(abstract)

AI总结 研究通过高通量计算发现183种动能可及的1D磁链,揭示其磁性特性及磁弹性耦合,为拓扑超导和马约拉纳零模提供基础。

Comments 22 pages, 4 figures, Supplementary Information supplied

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 软件智能体 7 篇

2512.24630 2026-01-01 cs.SE 83%

How Do Agentic AI Systems Address Performance Optimizations? A BERTopic-Based Analysis of Pull Requests

代理AI系统如何解决性能优化问题?基于BERTopic的拉取请求分析

Md Nahidul Islam Opu, Shahidul Islam, Muhammad Asaduzzaman, Shaiful Chowdhury

专题命中 软件智能体 :agentic(title,abstract);AI agent(abstract);分类 cs.SE

AI总结 本文通过BERTopic分析,揭示了代理AI系统在软件开发中处理性能优化的方式及影响因素。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24040 2026-01-01 cs.AI 83%

ROAD: Reflective Optimization via Automated Debugging for Zero-Shot Agent Alignment

ROAD: 通过自动调试实现零样本智能体对齐的反思优化

Natchaya Temyingyong, Daman Jain, Neeraj Kumarsahu, Prabhat Kumar, Rachata Phondi, Wachiravit Modecrua, Krittanon Kaewtawee, Krittin Pachtrachai, Touchapon Kraisingkorn

专题命中 软件智能体 :agent(title,abstract);multi-agent(abstract);分类 cs.AI

AI总结 ROAD通过自动调试实现零样本智能体对齐的反思优化,利用多智能体架构将无结构故障日志转化为结构化决策树协议,提升智能体性能和样本效率。

Comments 22 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23875 2026-01-01 cs.SE 83%

From Illusion to Insight: Change-Aware File-Level Software Defect Prediction Using Agentic AI

从幻觉到洞察:利用代理AI的变更感知文件级软件缺陷预测

Mohsen Hesamolhokama, Behnam Rohani, Amirahmad Shafiee, MohammadAmin Fazli, Jafar Habibi

专题命中 软件智能体 :agentic(title);agent(abstract);multi-agent(abstract);分类 cs.SE

AI总结 本文提出了一种基于代理AI的变更感知文件级软件缺陷预测方法,通过多代理辩论框架提升缺陷预测的准确性和敏感性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24636 2026-01-01 cs.SE 79%

How Do Agentic AI Systems Deal With Software Energy Concerns? A Pull Request-Based Study

代理AI系统如何处理软件能耗问题?基于拉取请求的研究

Tanjum Motin Mitul, Md. Masud Mazumder, Md Nahidul Islam Opu, Shaiful Chowdhury

专题命中 软件智能体 :agentic(title);agent(abstract);分类 cs.SE

AI总结 本文研究了代理AI系统在生成软件时对能耗问题的意识,通过分析拉取请求发现,尽管构建和运行这些系统耗能高,但生成的软件制品表现出一定的能耗意识,但优化相关的PR因影响可维护性而被接受较少。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03257 2026-01-01 cs.LG cs.AI cs.MA 73%

Triple-BERT: Do We Really Need MARL for Order Dispatch on Ride-Sharing Platforms?

Triple-BERT: 为网约车平台订单调度是否真的需要MARL?

Zijian Zhao, Sen Li

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

专题命中 软件智能体 :agent(abstract);multi-agent(abstract);分类 cs.AI、cs.LG

AI总结 Triple-BERT通过动作分解和BERT网络提升网约车平台订单调度效率,实现11.95%的性能提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23713 2026-01-01 cs.CL cs.AI 73%

PyBangla at BLP-2025 Task 2: Enhancing Bangla-to-Python Code Generation with Iterative Self-Correction and Multilingual Agents

在BLP-2025任务2中提升孟加拉语到Python代码生成:通过迭代自我纠正和多语言代理

Jahidul Islam, Md Ataullha, Saiful Azad

机构 * Department of Computer Science and Engineering(计算机科学与工程系)

专题命中 软件智能体 :agent(abstract);multi-agent(abstract);分类 cs.AI、cs.CL

AI总结 本文提出BanglaCodeAct框架,通过多代理提示和迭代自我纠正,提升孟加拉语到Python代码生成的性能,实验显示Qwen3-8B结合该框架在开发集和盲测集上分别达到94.0%和71.6%的准确率。

Comments 6 Pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07017 2026-01-01 cs.SE cs.AI 62%

Benchmarking LLMs for Fine-Grained Code Review with Enriched Context in Practice

基于实践丰富上下文的LLM细粒度代码审查基准测试

Ruida Hu, Xinchen Wang, Xin-Cheng Wen, Zhao Zhang, Bo Jiang, Pengfei Gao, Chao Peng, Cuiyun Gao

机构 * Harbin Institute of Technology(哈尔滨工业大学) ByteDance(字节跳动)

专题命中 软件智能体 :workflow(abstract);分类 cs.AI、cs.SE

AI总结 ContextCRBench通过丰富上下文的细粒度代码审查基准测试,评估LLM在代码审查中的性能,发现文本上下文比代码上下文更有效,且在工业应用中提升了审查系统性能。

详情

展开后加载摘要…

URL PDF HTML 收藏

4. GUI与网页智能体 4 篇

2512.19432 2026-01-01 cs.CL 83%

MobileWorld: Benchmarking Autonomous Mobile Agents in Agent-User Interactive and MCP-Augmented Environments

MobileWorld: 用于Agent-用户交互和MCP增强环境中的自主移动代理基准测试

Quyu Kong, Xu Zhang, Zhenyu Yang, Nolan Gao, Chen Liu, Panrong Tong, Chenglin Cai, Hanzhang Zhou, Jianan Zhang, Liangyu Chen, Zhidan Liu, Steven Hoi, Yue Wang

机构 * Tongyi Lab , Alibaba Group(通义实验室,阿里巴巴集团) HKUST (GZ)(香港科技大学(广州)) University of Florida(佛罗里达大学)

专题命中 GUI与网页智能体 :agent(title,abstract);agentic(abstract);分类 cs.CL

AI总结 MobileWorld通过201个任务和20个应用的挑战性基准测试,评估代理在用户交互和MCP增强环境中的表现,揭示了与AndroidWorld相比的显著性能下降。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24651 2026-01-01 cs.RO cs.AI cs.LG 81%

Hybrid Motion Planning with Deep Reinforcement Learning for Mobile Robot Navigation

基于深度强化学习的混合运动规划用于移动机器人导航

Yury Kolomeytsev, Dmitry Golembiovsky

机构 * Department of Computational Mathematics and Cybernetics(计算数学与自动化系) Lomonosov Moscow State University(罗蒙诺夫莫斯科国立大学)

专题命中 GUI与网页智能体 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出HMP-DRL,结合基于图的全局规划与深度强化学习,提升移动机器人在复杂环境中的导航安全性和可靠性。

Comments 22 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17221 2026-01-01 cs.CV 67%

DAVE: A VLM Vision Encoder for Document Understanding and Web Agents

DAVE: 一种用于文档理解与网络代理的视觉编码器

Brandon Huang, Hang Hua, Zhuoran Yu, Trevor Darrell, Rogerio Feris, Roei Herzig

机构 * MIT-IBM Watson AI Lab(MIT-IBM Watson AI实验室) UC Berkeley(加州大学伯克利分校) University of Wisconsin–Madison(威斯康星大学麦迪逊分校)

专题命中 GUI与网页智能体 :agent(abstract);agentic(abstract)

AI总结 DAVE是一种专为文档理解和网络代理设计的视觉编码器,通过自监督和监督预训练结合模型融合策略,提升对文档和网络任务的适应性与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22942 2026-01-01 cs.RO cs.AI 57%

AINav: Large Language Model-Based Adaptive Interactive Navigation

AINav: 基于大语言模型的自适应交互导航

Kangjie Zhou, Yao Mu, Haoyang Song, Yi Zeng, Pengying Wu, Han Gao, Chang Liu

机构 * School of Advanced Manufacturing and Robotics, Peking University(北京大学先进制造与机器人学院) AI Institute, School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机学院人工智能研究所)

专题命中 GUI与网页智能体 :planning(abstract);分类 cs.AI

AI总结 AINav通过基于大语言模型的自适应交互导航方法,实现复杂环境中的路径规划与目标达成。

Comments 13 pages, 12 figures, accepted to IEEE Robotics & Automation Magazine

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 记忆与上下文管理 7 篇

2512.23760 2026-01-01 cs.CR cs.AI 85%

Audited Skill-Graph Self-Improvement for Agentic LLMs via Verifiable Rewards, Experience Synthesis, and Continual Memory

通过可验证奖励、经验合成和持续记忆对代理LLM进行审计化技能图自改进

Ken Huang, Jerry Huang

机构 * OWASP Fairfax, VA, USA Kleiner Perkins

专题命中 记忆与上下文管理 :agentic(title,abstract);agent(abstract);AI agent(abstract);分类 cs.AI

AI总结 本文提出ASG-SI框架,通过可验证奖励、经验合成和持续记忆实现代理LLM的审计化技能图自改进,以提升安全性和可追溯性。

Comments 11 pages, 4 figures. Includes a complete runnable reference implementation and audit logging framework

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24684 2026-01-01 cs.CL cs.AI 79%

R-Debater: Retrieval-Augmented Debate Generation through Argumentative Memory

R-Debater:通过论证记忆的检索增强辩论生成

Maoyuan Li, Zhongsheng Wang, Haoyuan Li, Jiamou Liu

机构 * Wuhan College of Communication(武汉通信学院) Wuhan College of Communication University of Auckland(武汉通信学院奥克兰大学) University of Auckland(奥克兰大学)

专题命中 记忆与上下文管理 :agent(abstract);planning(abstract);agentic(abstract);分类 cs.AI、cs.CL

AI总结 R-Debater通过整合检索和结构化规划,实现了更忠实、一致且连贯的多轮辩论生成。

Comments Accepteed by AAMAS 2026 full paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23747 2026-01-01 cs.SE cs.AI cs.CL 67%

State-of-the-art Small Language Coder Model: Mify-Coder

最先进的小型语言编码器模型:Mify-Coder

Abhinav Parmar, Abhisek Panigrahi, Abhishek Kumar Dwivedi, Abhishek Bhattacharya, Adarsh Ramachandra, Aditya Choudhary, Aditya Garg, Aditya Raj, Alankrit Bhatt, Alpesh Yadav, Anant Vishnu, Ananthu Pillai, Ankush Kumar, Aryan Patnaik, Aswatha Narayanan S, Avanish Raj Singh, Bhavya Shree Gadda, Brijesh Pankajbhai Kachhadiya, Buggala Jahnavi, Chidurala Nithin Krishna, Chintan Shah, Chunduru Akshaya, Debarshi Banerjee, Debrup Dey, Deepa R., Deepika B G, Faiz ur Rahman, Gagan Gayari, Gudhi Jagadeesh Kumar Naidu, Gursimar Singh, Harshal Tyagi, Harshini K, James Mani Vathalloor, Jayarama Nettar, Jayashree Gajjam, Joe Walter Sugil George, Kamalakara Sri Krishna Tadepalli, Kamalkumar Rathinasamy, Karan Chaurasia, Karthikeyan S, Kashish Arora, Kaushal Desai, Khushboo Buwade, Kiran Manjrekar, Malikireddy Venkata Sai Likhitha, Manjunath A, Mitali Mahavir Bedmutha, Mohammed Rafee Tarafdar, Nikhil Tiwari, Nikitha K Gigi, Pavan Ravikumar, Pendyala Swarnanjali, Piyush Anand, Prakash Chandrasekar, Prasanna Bhalchandra Gawade, Prasanth Sivan, Preeti Khurana, Priyanshi Babbar, Rajab Ali Mondal, Rajesh Kumar Vissapragada, Rajeshwari Ganesan, Rajeswari Koppisetti, Ramjee R., Ramkumar Thiruppathisamy, Rani G. S., S Reka, Samarth Gupta, Sandeep Reddy Kothakota, Sarathy K, Sathyanarayana Sampath Kumar, Saurabh Kumar, Shashank Khasare, Shenbaga Devi Venkatesh Kumar, Shiva Rama Krishna Parvatham, Shoeb Shaikh, Shrishanmathi A, Shubham Pathak, Sree Samhita Koppaka, Sreenivasa Raghavan K S, Sreeram Venkatasubramanian, Suprabha Desai Bojja, Swetha R, Syed Ahmed, Chinmai Harshitha Thota, Tushar Yadav, Veeravelly Kusumitha, V V S S Prasanth Patnaik, Vidya Sri Sesetti, Vijayakeerthi K, Vikram Raj Bakshi, Vinay K K, Vinoth Kumar Loganathan, Vipin Tiwari, Vivek Kumar Shrivastav, V Venkata Sri Datta Charan, Wasim Akhtar Khan

机构 * Infosys AI Research(英矽斯人工智能研究院) Mify Team(Mify团队)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL、cs.SE

AI总结 Mify-Coder通过高效训练策略和数据优化,在保持高准确性和安全性的同时,实现了比更大模型更优的代码生成性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23982 2026-01-01 cs.SE cs.AI 62%

Coding With AI: From a Reflection on Industrial Practices to Future Computer Science and Software Engineering Education

编码与AI:从工业实践的反思到未来计算机科学与软件工程教育

Hung-Fu Chang, MohammadShokrolah Shirazi, Lizhou Cao, Supannika Koolmanojwong Mobasser

机构 * R. B. Annis School of Engineering(R. B. Annis 工程学院) University of Indianapolis(印第安纳大学) E. S. Witchger School of Engineering(E. S. Witchger 工程学院) Marian University(玛丽安大学) University of Maryland Eastern Shore(马里兰大学东部分校) The Boehm Center for Systems and Software Engineering(Boehm 系统与软件工程中心)

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.AI、cs.SE

AI总结 本文探讨了AI在工业实践中对软件开发的影响,分析了LLM工具带来的生产力提升与风险,并提出教育应转向问题解决和项目式学习以适应变化。

Comments 21 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14297 2026-01-01 cs.NI cs.AI cs.ET cs.PF hep-ex 57%

A Threshold-Triggered Deep Q-Network-Based Framework for Self-Healing in Autonomic Software-Defined IIoT-Edge Networks

基于阈值触发的深度Q网络框架用于自愈的自治软件定义IIoT边缘网络

Agrippina Mwangi, León Navarro-Hilfiker, Lukasz Brewka, Mikkel Gryning, Elena Fumagalli, Madeleine Gibescu

机构 * Energy and Resources Group at the Copernicus Institute of Sustainable Development, Utrecht University(能源与资源组,可持续发展Copernicus研究所,乌特勒支大学) Operational Technology (OT) System Engineering and Security (Engineering, Procurement, and Construction) at Ørsted Wind Power(运营技术(OT)系统工程与安全(工程、采购与建设)部,Ørsted风能) System Design Specialists (Engineering, Procurement, and Construction) at Ørsted Wind Power(系统设计专家(工程、采购与建设)部,Ørsted风能)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

AI总结 本研究提出基于阈值触发的深度Q网络框架,用于提升自治软件定义IIoT边缘网络的自愈能力,通过实时检测和缓解网络中断,提高中断恢复性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09179 2026-01-01 math.OC 50%

Reachability for multiagent control systems via Lyapunov functions

通过李雅普诺夫函数研究多智能体控制系统可达性

Giulia Cavagnari, Marc Quincampoix

专题命中 记忆与上下文管理 :agent(abstract)

AI总结 本文提出利用适应于概率测度Wasserstein空间的李雅普诺夫方法,研究多智能体控制系统可达性问题,并获得Hamilton-Jacobi方程在该空间中粘性解的新的比较结果。

Comments This is the author accepted manuscript of an article published in Communications on Pure and Applied Analysis

Journal ref Communications on Pure and Applied Analysis, vol. 26 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06141 2026-01-01 cs.FL 50%

[Draft] High-order estimation-based properties and high-order observers for labeled finite-state automata

[草稿] 基于高阶估计的性质和高阶观测器用于标记有限状态自动机

Kuize Zhang, Xiaoguang Han, Alessandro Giua, Carla Seatzu

专题命中 记忆与上下文管理 :agent(abstract)

AI总结 本文提出了一种基于高阶估计性质的框架,通过高阶观察器验证标记有限状态自动机的属性。

Comments 36 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

6. Agent评测 12 篇

2512.23844 2026-01-01 cs.SE cs.AI cs.HC 88%

From Correctness to Collaboration: Toward a Human-Centered Framework for Evaluating AI Agent Behavior in Software Engineering

从正确性到协作:迈向以人为中心的评估AI代理行为在软件工程中的框架

Tao Dong, Harini Sampath, Ja Young Lee, Sherry Y. Shi, Andrew Macvean

机构 * Google LLC(谷歌公司)

专题命中 Agent评测 :agent(title,abstract);AI agent(title,abstract);分类 cs.AI、cs.SE

AI总结 本文提出以人为中心的评估框架,旨在评估AI代理在软件工程中的协作行为,通过定义代理行为期望和引入情境适应框架,推动AI代理向协作智能发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.25055 2026-01-01 cs.AI cs.HC 83%

Context-aware LLM-based AI Agents for Human-centered Energy Management Systems in Smart Buildings

面向人类中心的基于大语言模型的AI代理用于智能建筑中的能源管理系统

Tianzhi He, Farrokh Jazizadeh

机构 * organization= School of Civil \& Environmental Engineering Construction Management, The University of Texas at San Antonio , addressline= BSE 1.310, One UTSA Circle , city= San Antonio , postcode= 78249 , state= TX , country= U.S. organization= Department of Civil Environmental Engineering, Virginia Polytechnic Institute State University , addressline= 200 Patton Hall, 750 Drillfield , city= Blacksburg , postcode= 24060 , state= VA , country= U.S.

专题命中 Agent评测 :AI agent(title,abstract);agent(abstract);分类 cs.AI

AI总结 本研究提出基于大语言模型的AI代理,用于智能建筑中的人类中心能源管理,通过自然语言交互实现情境感知的能源优化与管理。

详情

展开后加载摘要…

URL PDF HTML 收藏