Switch-based Active Deep Dyna-Q: Efficient Adaptive Planning for Task-Completion Dialogue Policy Learning
专题命中 规划推理 :planning(title);分类 cs.CL、cs.AI、cs.LG
Comments 8 pages, 9 figures, AAAI 2019
AI 大模型
大模型数学、逻辑、规划、多步推理和测试时计算能力。
专题命中 规划推理 :planning(title);分类 cs.CL、cs.AI、cs.LG
Comments 8 pages, 9 figures, AAAI 2019
专题命中 规划推理 :planning(title,abstract)
Comments 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems, Madrid, Spain
专题命中 规划推理 :planning(title,abstract)
专题命中 规划推理 :planning(title,abstract)
Comments Published in IROS 2018
专题命中 规划推理 :planning(title,abstract)
Comments Accepted for oral presentation at 12th IFAC Symposium on Robot Control (SYROCO 2018)
专题命中 规划推理 :planning(title,abstract)
Comments Internship Report
专题命中 规划推理 :planning(title,abstract)
Comments This paper was accepted for the preproceedings of The 2nd International Workshop on Agent-based modelling of urban systems (ABMUS 2017), http://www.modelling-urban-systems.com/
专题命中 规划推理 :planning(title,abstract)
Comments This paper was accepted for publication in the International Conference on Advanced Robotics 2013. It was not included in the final proceedings of the conference as I was unable to attend the conference to present the paper
专题命中 规划推理 :planning(title,abstract)
Comments To be published (Accepted) in: Proceedings of the North American Power Symposium, Morgantown, WV, Sep. 17-19,2017
专题命中 规划推理 :planning(title,abstract)
Comments 24 pages, 4 figures, submitted to IJDSN(International Journal of Distributed Sensor Networks)
专题命中 规划推理 :planning(title,abstract)
Comments 8 pages, 12 figures, submitted to the first International Symposium on Multi-Robot and Multi-Agent Systems (MRS)
专题命中 规划推理 :planning(title,abstract)
专题命中 规划推理 :planning(title,abstract)
Comments Standard 4 page IEEE Format Submitted in IEEE-DTU Technical Journal IOTA
专题命中 规划推理 :planning(title,abstract)
Comments Submitted to the IEEE International Conference on Robotics and Automation (ICRA), Singapore, 2017
专题命中 规划推理 :planning(title,abstract)
Comments 11 pages, 2 figures, 3 tables
专题命中 规划推理 :planning(title,abstract)
Comments 20 pages, 7 figures, submitted to EC'15
专题命中 规划推理 :planning(title,abstract)
Comments Extended version of ACC 2014 paper
专题命中 规划推理 :planning(title,comments);分类 cs.AI、cs.LG
Comments 9 pages, ICAPS 2020 Conference - Bridging the Gap Between AI Planning and Reinforcement Learning (PRL) Workshop
抖音多模态嵌入模型技术报告
专题命中 规划推理 :CoT(abstract,abstract_cn);reasoning(abstract);分类 cs.CL
AI总结 针对现有多模态嵌入模型难以兼顾效率与细粒度区分的问题,提出分两阶段训练的DME模型,在MMEB-v2数据集及抖音生产场景中均取得优异效果。
Comments Technical Report
智能体 harness:面向机器人自主的大语言模型驱动验证层
机构 * Carnegie Mellon University(卡内基梅隆大学) ; Pacific Northwest National Laboratory(太平洋西北国家实验室)
专题命中 规划推理 :reasoning(abstract);chain-of-thought(abstract);planning(abstract);分类 cs.AI
AI总结 针对机器人规划模型的安全与伦理风险,提出LLM驱动的验证层作为中间件管控计划,实现近85%的类别准确率、97%的对抗性攻击遏制率,为机器人自主提供可靠保障。
Comments 7 pages. Not yet finalized for conference submission
AI能像城市规划师一样推理吗?基于专业判断的大语言模型基准测试
机构 * School of Architecture and Urban Planning, Shenzhen University(深圳大学建筑与城市规划学院) ; Shenzhen Key Laboratory of Urban Spatial Information and Intelligent Modeling(深圳市城市空间信息与智能建模重点实验室) ; Department of Urban Planning and Design, The University of Hong Kong(香港大学城市规划与设计系)
专题命中 规划推理 :planning(abstract,abstract_cn);reasoning(abstract);分类 cs.CL
AI总结 提出UPBench框架,通过4×5知识支柱与认知水平矩阵评估25个LLM,发现模型在分析任务上优于事实回忆和综合判断,揭示了规划知识的制度依赖性。
Comments This paper has been withdrawn by the authors because the current version requires substantial revision and further validation before it can be considered a reliable representation of the work
Qwen-AgentWorld: 通用智能体的语言世界模型
机构 * Qwen Team(Qwen团队)
专题命中 规划推理 :reasoning(abstract);chain-of-thought(abstract);planning(abstract);分类 cs.CL
AI总结 提出基于语言模型的世界模型Qwen-AgentWorld,通过三阶段训练(CPT、SFT、RL)模拟7个领域的智能体环境,并构建AgentWorldBench基准,实验表明其显著优于现有模型,且能作为环境模拟器和智能体基础模型提升下游性能。
STEP-LLM: 通过大型语言模型生成CAD STEP模型
机构 * Northwestern University(西北大学)
专题命中 规划推理 :CoT(abstract,abstract_cn);chain-of-thought(abstract);分类 cs.AI
AI总结 本文提出STEP-LLM,通过大型语言模型将自然语言转化为CAD STEP模型,采用图结构预处理和强化学习提升几何精度,验证了LLM驱动的STEP模型生成可行性。
Comments Accepted to the Design, Automation & Test in Europe Conference (DATE) 2026
Journal ref In Proceedings of the 2026 Design, Automation & Test in Europe Conference (DATE 2026), 2026
MedCollab:基于IBIS引导的多智能体协作与分层疾病关系链的临床诊断
机构 * Princeton University(普林斯顿大学) ; Springer Heidelberg(斯普林格海德堡) ; ABC Institute(ABC研究所) ; Rupert-Karls-University Heidelberg(海德堡鲁珀特-卡尔大学) ; Hangzhou Dianzi University(杭州电子科技大学) ; Zhejiang University(浙江大学) ; Children’s Hospital, Zhejiang University School of Medicine, National Clinical Research Center for Children and Adolescents’ Health and Diseases(浙江大学医学院儿童医院,国家儿童青少年健康与疾病临床研究中心)
专题命中 规划推理 :reasoning(abstract);planning(abstract);verifier(abstract);分类 cs.AI
AI总结 提出MedCollab框架,通过IBIS结构化论证和分层疾病关系链(HDRC)增强多智能体协作,提升临床诊断的准确性、可追溯性和报告质量。
MUSE: 多模态大语言模型的统一智能体框架
机构 * Northeastern University(东北大学)
专题命中 规划推理 :reasoning(abstract);planning(abstract);verifier(abstract);分类 cs.AI
AI总结 提出MUSE框架,通过可组合模块(任务表示、视觉处理、感知工具、结构化解析、确定性验证和验证器引导修复)提升冻结多模态大语言模型性能,无需重新训练。
结构使语言模型能够有效自我定位错误
机构 * Meta AI ; Columbia University(哥伦比亚大学) ; Meta Superintelligence Labs(Meta超智能实验室) ; Tel Aviv University(特拉维夫大学)
专题命中 规划推理 :reasoning(abstract);chain-of-thought(abstract);self-correction(abstract);分类 cs.AI
AI总结 本文提出结构化推理方法,通过将推理分解为离散语义步骤,使语言模型能更可靠地定位错误,并基于此设计了迭代纠正采样框架Thought-ICS,实现20-40%的自我纠正提升。
元认知行为调节用于多跳问答的大型语言模型
机构 * Dept. of ECE Seoul Nat’l Univ.(首尔国立大学电子工程系) ; AI Center Samsung Elec. Korea(三星电子韩国人工智能中心) ; AIIS, ASRI, INMC, ISRC, Interdisc. Prog. in AI Seoul Nat’l Univ.(首尔国立大学人工智能研究所、高级研究机构、人工智能中心、国际研究所以及人工智能跨学科研究计划)
专题命中 规划推理 :reasoning(abstract);planning(abstract);self-correction(abstract);分类 cs.AI
AI总结 本文提出元认知行为调节(MBT)框架,通过五阶段结构提升多跳问答任务的准确性和效率,减少冗余并保持推理轨迹稳定。
Comments 41 pages
视觉提示对视觉语言模型合作行为的影响
机构 * ST Engineering, Singapore(1 ST工程,新加坡)
专题命中 规划推理 :CoT(abstract,abstract_cn);reasoning(abstract);分类 cs.AI
AI总结 本文研究了视觉提示如何影响视觉语言模型在囚徒困境中的合作行为,通过图像内容和颜色奖励矩阵实验,发现图像内容和颜色提示对模型决策模式有影响,不同模型的敏感性和缓解效果各异。
上下文自主网络事件响应:一种端到端的大语言模型代理方法
机构 * Department of Systems Engineering, City University of Hong Kong(香港城市大学系统工程系) ; Department of Electrical and Electronic Engineering, University of Melbourne(墨尔本大学电子与电气工程系)
专题命中 规划推理 :reasoning(abstract);chain-of-thought(abstract);planning(abstract);分类 cs.AI
AI总结 本文提出一种端到端的大语言模型代理方法,通过整合感知、推理、规划和行动功能,实现网络事件响应的自主学习与适应,实验表明其恢复效率比前沿模型快23%。
Comments 2026 AAAI Summer Symposium on Human-Aware AI Agents for the Cyber Battlefield
自校正RAG:通过MMKP上下文选择和NLI引导的MCTS增强忠实性
机构 * Chongqing University(重庆大学) ; Queen Mary University of London(伦敦玛丽女王大学) ; Chongqing Key Laboratory of Big Data Intelligence and Privacy Computing(重庆市大数据智能与隐私计算重点实验室) ; Fangda Partners(方达律师事务所)
专题命中 规划推理 :reasoning(abstract);planning(abstract);test-time compute(abstract);分类 cs.CL
AI总结 本研究提出Self-Correcting RAG框架,通过MMKP上下文选择和NLI引导的MCTS机制,提升复杂推理任务中的准确性和减少幻觉。