arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

University of Chinese Academy of Sciences(中国科学院大学)

2026-02-13 至 2026-02-13 共收录 6
2602.11073 2026-02-13 cs.CV cs.AI cs.CL

Chatting with Images for Introspective Visual Thinking

与图像对话以进行反思性视觉思维

Junfei Wu, Jian Guan, Qiang Liu, Shu Wu, Liang Wang, Wei Wu, Tieniu Tan

机构 * NLPR, MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) University of Chinese Academy of Sciences(中国科学院大学) Nanjing University(南京大学)

AI总结 ViLaVT通过语言引导的特征调节框架,实现多图像区域的联合重编码,提升跨模态对齐与复杂空间推理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11693 2026-02-13 cs.GR cs.AI cs.CV

OMEGA-Avatar: One-shot Modeling of 360° Gaussian Avatars

OMEGA-Avatar: 从单张图像进行360°高斯avator的一键建模

Zehao Xia, Yiqun Wang, Zhengda Lu, Kai Liu, Jun Xiao, Peter Wonka

机构 * Chongqing University(重庆大学) University of Chinese Academy of Sciences(中国科学院大学)

AI总结 OMEGA-Avatar通过两个创新模块实现了从单张图像生成360°完整且可动画的3D高斯头,提升了全头avator的建模精度和动画能力。

Comments Project page: https://omega-avatar.github.io/OMEGA-Avatar/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11536 2026-02-13 cs.CV

Vascular anatomy-aware self-supervised pre-training for X-ray angiogram analysis

基于血管解剖的自监督预训练方法用于X射线造影分析

De-Xing Huang, Chaohui Yu, Xiao-Hu Zhou, Tian-Yu Xiang, Qin-Yi Zhang, Mei-Jiang Gui, Rui-Ze Ma, Chen-Yu Wang, Nu-Fang Xiao, Fan Wang, Zeng-Guang Hou

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) Joint Laboratory of Intelligence Science and Technology, Institute of Systems Engineering, Macau University of Science and Technology(智能科学与技术联合实验室,系统工程研究所,澳门科学大学) DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团)

AI总结 本文提出VasoMIM框架,结合解剖引导掩码策略和一致性损失,用于X射线造影分析,通过XA-170K数据集实现高效预训练并取得最佳性能。

Comments 10 pages, 10 figures, 10 tables. Journal version of VasoMIM (AAAI 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11494 2026-02-13 cs.CV

Arbitrary Ratio Feature Compression via Next Token Prediction

通过下一个标记预测实现任意比例特征压缩

Yufan Liu, Daoyuan Ren, Zhipeng Zhang, Wenyang Luo, Bing Li, Weiming Hu, Stephen Maybank

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institution of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) CAS Center for Excellence in Brain Science and Intelligence Technology(中国科学院脑科学与智能技术卓越创新中心) People AI, Inc.(People AI公司) School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) School of Computer Science and Mathematics, Birkbeck College, University of London(伦敦大学伯克贝克学院计算机科学与数学学院)

AI总结 本文提出了一种通过下一个标记预测实现任意比例特征压缩的框架,解决了传统方法在灵活性和通用性上的不足,通过引入混合解决方案和实体关系图约束模块,提升了压缩特征的质量和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11241 2026-02-13 cs.CV cs.LG

Active Zero: Self-Evolving Vision-Language Models through Active Environment Exploration

主动零:通过主动环境探索实现自我进化的视觉-语言模型

Jinghan He, Junfeng Fang, Feng Xiong, Zijun Yao, Fei Shen, Haiyun Guo, Jinqiao Wang, Tat-Seng Chua

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) National University of Singapore(新加坡国立大学) Wuhan AI Research(武汉人工智能研究所) Tsinghua University(清华大学)

AI总结 Active-Zero通过主动环境探索实现视觉-语言模型的自我进化,提升推理和理解任务的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05014 2026-02-13 cs.AI cs.CL cs.IR

DeepRead: Document Structure-Aware Reasoning to Enhance Agentic Search

DeepRead: 一种基于文档结构的推理以增强代理搜索

Zhanli Li, Huiwen Tian, Lvzhou Luo, Yixuan Cao, Ping Luo

机构 * State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences (CAS)(人工智能安全国家重点实验室,计算技术研究所,中国科学院) University of Chinese Academy of Sciences, CAS(中国科学院大学) Wenlan School of Business, Zhongnan University of Economics(中南财经政法大学文澜商学院)

AI总结 DeepRead通过结构感知的文档推理增强代理搜索,有效提升文档理解精度。

Comments This version has significantly enhanced the clarity of our research

详情

展开后加载摘要…

URL PDF HTML 收藏