HaReCAP: Habitual-action Grounding for Recursive Large Language Model Agents
HaReCAP:面向递归大语言模型智能体的习惯性动作 grounding 方法
Shen Liu, Zhenguo Xu, Shaopu Wang, Yike Gao, Chunlei Wang
机构
*
North China Institute of Computer System Engineering(华北计算机系统工程研究所)
;
University of Science and Technology of China(中国科学技术大学)
;
China Information Security Research Institute Co., Ltd.(中国信息安全研究院有限公司)
DRAgent: Discriminative Reasoning Agent for Referring Expression Segmentation
DRAgent:用于指代表达分割的判别推理智能体
Yujie Qi, Luyan Zhang
机构
*
School of Computer Science and Technology, Hangzhou Dianzi University(杭州电子科技大学计算机科学与技术学院)
;
Khoury College of Computer Sciences, Northeastern University(东北大学Khoury计算机科学学院)
专题命中
视觉定位与Grounding
:MLLM(summary_cn,abstract);multimodal large language model(abstract);分类 cs.CV
An end-to-end-trained vision-language model for native-language prostate pathology report generation
用于生成本土语言前列腺病理报告的端到端训练视觉-语言模型
Christian Grashei, Fabian Gülhan, Maximilian Legnar, Fabian Stögbauer, Cleo-Aron Weis, Carolin Mogler, Peter Schüffler
机构
*
Technical University of Munich(慕尼黑工业大学)
;
Munich Data Science Institute(慕尼黑数据科学研究所)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
University Hospital Heidelberg(海德堡大学医院)
;
Heidelberg University(海德堡大学)
;
Interdisciplinary Center for Scientific Computing (IWR)(跨学科科学计算中心(IWR))
ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting
ExtrinSplat:解耦几何与语义以实现3D高斯散射中的开放词汇理解
Jiayu Ding, Xinpeng Liu, Zhiyi Pan, Shiqiang Long, Ge Li
机构
*
Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology, Shenzhen Graduate School, Peking University(广东省超高清沉浸式媒体技术重点实验室,北京大学深圳研究生院)
;
School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院)
;
Guangdong Bohua UHD Innovation Center Co., Ltd.(广东博华超高清创新中心有限公司)
CommentsTo appear in Proceedings of the 1st International Workshop on Specification-Driven Development Life Cycle (SpecOps 2026), co-located with SPLASH 2026
Qixiang Yin, Huanjin Yao, Yuchen Cai, Jianghao Chen, Ziyi Wang, Min Yang, Fei Su, Zhicheng Zhao
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
ByteDance(字节跳动)
;
USTC(中国科学技术大学)
;
Beijing Key Laboratory of Network System and Network Culture(北京网络系统与网络文化重点实验室)
;
Key Laboratory of Interactive Technology and Experience System, Ministry of Culture and Tourism(文化和旅游部互动技术与体验系统重点实验室)
;
Zhongguancun Academy(中关村科学城)
CommentsWithdrawn due to serious concerns regarding the authenticity and accuracy of the listed authorship. The identity of one or more listed authors cannot presently be verified, and the author list may not represent distinct contributors. The manuscript is withdrawn pending institutional review