Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective
从序列层面视角出发的扩散大语言模型原理化强化学习
Jingyang Ou, Jiaqi Han, Minkai Xu, Shaoxuan Xu, Jianwen Xie, Stefano Ermon, Yi Wu, Chongxuan Li
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大模型与智能治理研究重点实验室)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程技术研究中心)
;
Stanford University(斯坦福大学)
;
Lambda, Inc(Lambda公司)
;
Tsinghua University(清华大学)
Benchmark for Planning and Control with Large Language Model Agents: Blocksworld with Model Context Protocol
基于大语言模型代理的规划与控制基准:带有模型上下文协议的积木世界
Niklas Jobs, Luis Miguel Vieira da Silva, Jayanth Somashekaraiah, Maximilian Weigand, David Kube, Felix Gehlhoff
机构
*
Institute of Automation Technology, Helmut Schmidt University / University of the Federal Armed Forces Hamburg, Germany(自动化技术研究所,海姆·施密特大学/联邦武装部队大学汉堡,德国)
;
Siemens AG, Nuremberg, Germany(西门子股份公司,纽伦堡,德国)
Prediction-Driven Motion Planning: Route Integration Strategies in Attention-Based Prediction Models
基于预测的运动规划:注意力预测模型中的路线整合策略
Marlon Steiner, Royden Wagner, Ömer Sahin Tas, Christoph Stiller
机构
*
Institute of Measurement and Control Systems, Karlsruhe Institute of Technology (KIT)(测量与控制系统研究所,卡尔斯鲁厄理工学院)
;
FZI Research Center for Information Technology(信息技术研究所以)
专题命中
规划推理
:planning(title,abstract)
AI总结
本文提出基于注意力机制的运动预测模型,通过整合导航信息提升预测与规划任务的性能。
CommentsIn Proceedings of the IEEE International Conference on Intelligent Transportation Systems (ITSC), Gold Coast, AUSTRALIA, 18-21 November 2025
Hierarchical Vision Language Action Model Using Success and Failure Demonstrations
基于成功与失败示范的分层视觉语言行动模型
Jeongeun Park, Jihwan Yoon, Byungwoo Jeon, Juhan Park, Jinwoo Shin, Namhoon Cho, Kyungjae Lee, Sangdoo Yun, Sungjoon Choi
机构
*
Department of Artificial Intelligence, Korea University(韩国大学人工智能系)
;
Kim Jaechul Graduate School of AI, KAIST(金在拙人工智能研究生院,韩国科学技术院)
;
Department of Aerospace Engineering, Seoul National University(首尔国立大学航空航天工程系)
;
Departmnet of Statistics, Korea University(韩国大学统计系)
;
NAVER AI Lab(NAVER人工智能实验室)
From Pixels to Prose: Advancing Multi-Modal Language Models for Remote Sensing
从像素到 prose:推进遥感多模态语言模型
Xintian Sun, Benji Peng, Charles Zhang, Fei Jin, Qian Niu, Junyu Liu, Keyu Chen, Ming Li, Pohsun Feng, Ziqian Bi, Ming Liu, Xinyuan Song, Yichao Zhang
机构
*
Simon Fraser University(西蒙弗雷泽大学)
;
University of Minnesota - Twin Cities(明尼苏达大学双城分校)
;
Kyoto University(京都大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
National Taiwan Normal University(台湾师范大学)
;
Purdue University(普渡大学)
;
Emory University(埃默里大学)
;
The University of Texas at Dallas(德克萨斯大学达拉斯分校)
SpatialReasoner: Active Perception for Large-Scale 3D Scene Understanding
SpatialReasoner: 大规模3D场景理解中的主动感知
Hongpei Zheng, Shijie Li, Yanran Li, Hujun Yin
机构
*
University of Manchester(曼彻斯特大学)
;
Institute for Infocomm Research (I2R), A*STAR, Singapore(信息与通信研究 institute(I2R),A*STAR,新加坡)
;
University of Bedfordshire(贝德福德郡大学)
Language-Driven Object-Oriented Two-Stage Method for Scene Graph Anticipation
基于语言的面向对象两阶段方法用于场景图预见
Xiaomeng Zhu, Changwei Wang, Haozhe Wang, Xinyu Liu, Fangzhen Lin
机构
*
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(计算机科学与工程系,香港科学与技术大学)
;
Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center, Qilu University of Technology(计算能力网络与信息安全重点实验室,教育部,山东计算机科学中心,齐鲁大学)
;
Academy of Interdisciplinary Studies, The Hong Kong University of Science and Technology(跨学科研究学院,香港科学与技术大学)