When Agents Fail to Act: A Diagnostic Framework for Tool Invocation Reliability in Multi-Agent LLM Systems
当代理人失败时:多代理LLM系统工具调用可靠性的诊断框架
专题命中 多智能体 :agent(title,abstract);multi-agent(title,abstract);tool-use(abstract);分类 cs.AI
AI总结 本文提出了一种多代理LLM系统工具调用可靠性的诊断框架,通过大数据分析评估程序可靠性,发现工具初始化故障是小模型的主要瓶颈,而qwen2.5:32b在性能上接近GPT-4.1。
Comments Accepted for publication in 2026 The 9th International Conference on Artificial Intelligence and Big Data (ICAIBD 2026)