arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Science and Technology of China(中国科学技术大学)

2025-11-25 至 2025-11-25 共收录 10
2511.19425 2025-11-25 cs.CV

SAM3-Adapter: Efficient Adaptation of Segment Anything 3 for Camouflage Object Segmentation, Shadow Detection, and Medical Image Segmentation

SAM3-Adapter: 高效适应Segment Anything 3用于伪装物分割、阴影检测和医学图像分割

Tianrun Chen, Runlong Cao, Xinda Yu, Lanyun Zhu, Chaotao Ding, Deyi Ji, Cheng Chen, Qi Zhu, Chunyan Xu, Papa Mao, Ying Zang

机构 * KOKONI, Moxin (Huzhou) Tech. Co., LTD(摩西(湖州)科技有限公司) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院) School of Information Engineering, Huzhou University(湖州大学信息工程学院) School of Electrical and Electronic Engineering, Nanyang Technological University(新加坡南洋理工大学电子与电气工程学院) College of Computing and Data Science, Nanyang Technological University(新加坡南洋理工大学计算与数据科学学院) School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)

AI总结 SAM3-Adapter通过高效适配框架提升SAM3在伪装物分割、阴影检测和医学影像分割等任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18920 2025-11-25 cs.CV

EventSTU: Event-Guided Efficient Spatio-Temporal Understanding for Video Large Language Models

EventSTU: 基于事件的高效空间-时间理解用于视频大语言模型

Wenhao Xu, Xin Dong, Yue Li, Haoyuan Shi, Zhiwei Xiong

机构 * University of Science and Technology of China(中国科学技术大学)

AI总结 EventSTU通过事件引导的方法,高效处理视频大语言模型的空间-时间理解,实现显著的计算效率提升和性能优化。

Comments 8 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01421 2025-11-25 cs.CV

InfoScale: Unleashing Training-free Variable-scaled Image Generation via Effective Utilization of Information

InfoScale: 通过有效利用信息实现免训练可变尺度图像生成

Guohui Zhang, Jiangtong Tan, Linjiang Huang, Zhonghang Yuan, Mingde Yao, Jie Huang, Feng Zhao

机构 * USTC(中国科学技术大学) Beihang University(北京航空航天大学) CUHK MMLab(香港中文大学多媒体实验室) Kuaishou Technology(快手科技)

AI总结 InfoScale通过有效利用信息解决扩散模型在可变尺度图像生成中的信息丢失、聚合不灵活和分布不匹配问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20840 2025-11-25 cs.RO cs.AI cs.MM

Learning Primitive Embodied World Models: Towards Scalable Robotic Learning

学习原始具身世界模型:迈向可扩展的机器人学习

Qiao Sun, Liujia Yang, Wei Tang, Wei Huang, Kaixin Xu, Yongchao Chen, Mingyu Liu, Jiange Yang, Haoyi Zhu, Yating Wang, Tong He, Yilun Chen, Xili Dai, Nanyang Ye, Qinying Gu

机构 * Shanghai AI Lab(上海人工智能实验室) Fudan(复旦大学) SJTU(上海交通大学) NJUST(南京理工大学) THU(清华大学) Harvard(哈佛大学) ZJU(浙江大学) NJU(南京大学) USTC(中国科学技术大学) Tongji(同济大学) HKUST (GZ)(香港科技大学(广州))

AI总结 提出原始具身世界模型(PEWM)以解决具身数据稀疏和高维问题,通过短视界视频生成实现细粒度对齐、降低学习复杂度并提高数据效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11443 2025-11-25 cs.CV cs.AI

COLI: A Hierarchical Efficient Compressor for Large Images

COLI:一种用于大图像的分层高效压缩器

Haoran Wang, Hanyu Pei, Yang Lyu, Kai Zhang, Li Li, Feng-Lei Fan

机构 * Frontier of Artificial Networks (FAN) Lab, Department of Data Science, City University of Hong Kong(前沿人工智能网络实验室,数据科学系,香港城市大学) Molecular Imaging Business Unit, Shanghai United Imaging Healthcare Co., Ltd(分子影像业务部,上海联合影像医疗科技股份有限公司) MoE Key Laboratory of Brain-Inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知关键实验室,中国科学技术大学)

AI总结 COLI通过改进的神经表示方法,实现大图像高效压缩,提升压缩比并加快训练速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18088 2025-11-25 cs.RO

A Unified Multi-Dynamics Framework for Perception-Oriented Modeling in Tendon-Driven Continuum Robots

为腱驱动连续机器人感知导向建模设计的统一多动态框架

Ibrahim Alsarraj, Yuhao Wang, Abdalla Swikir, Cesare Stefanini, Dezhen Song, Zhanchi Wang, Ke Wu

机构 * Robotics Department, Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(Mohamed bin Zayed大学人工智能研究院机器人部门) Hefei National Research Center for Physical Sciences at the Microscale, University of Science and Technology of China (USTC)(中国科学技术大学微尺度物理科学国家级研究中心)

AI总结 本文提出了一种统一的多动态框架,用于腱驱动连续机器人中基于内在动态的感知建模与交互解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18006 2025-11-25 cs.LG

Understanding Private Learning From Feature Perspective

从特征角度理解隐私学习

Meng Ding, Mingxi Lei, Shaopeng Fu, Shaowei Wang, Di Wang, Jinhui Xu

机构 * Department of Computer Science and Engineering, State University of New York at Buffalo(纽约州立大学布法罗分校计算机科学与工程系) Division of CEMSE, King Abdullah University of Science and Technology(卡普兰大学科学与技术大学CEMSE分校) Institute of Artificial Intelligence and Blockchain, Guangzhou University(广州大学人工智能与区块链研究院) School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)

AI总结 本文从特征学习角度提出首个隐私训练理论框架,揭示隐私学习中特征信号学习需更高信噪比,且数据噪声记忆会影响泛化能力。

Comments 39pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01982 2025-11-25 cs.LG cs.CV

Fine-Grained GRPO for Precise Preference Alignment in Flow Models

细粒度GRPO用于流模型中的精确偏好对齐

Yujie Zhou, Pengyang Ling, Jiazi Bu, Yibin Wang, Yuhang Zang, Jiaqi Wang, Li Niu, Guangtao Zhai

机构 * Shanghai Jiao Tong University(上海交通大学) University of Science and Technology of China(中国科学技术大学) Fudan University(复旦大学) Shanghai AI Laboratory(上海人工智能实验室) Shanghai Innovation Institute(上海创新研究院)

AI总结 本文提出细粒度GRPO框架,通过细粒度评估和多粒度整合模块提升流模型中偏好对齐的精度与鲁棒性。

Comments Project Page: https://bujiazi.github.io/g2rpo.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03833 2025-11-25 cs.LG stat.ML

Understanding Fine-tuning in Approximate Unlearning: A Theoretical Perspective

理解近似反向学习中的微调:一种理论视角

Meng Ding, Rohan Sharma, Changyou Chen, Jinhui Xu, Kaiyi Ji

机构 * Department of Computer Science and Engineering(计算机科学与工程系) State University of New York at Buffalo(纽约州立大学布法罗分校) School of Information Science and Technology(信息科学与技术学院) University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出基于保留的遮蔽策略,通过理论分析揭示微调方法在反向学习中的局限性,并改进反向学习和保留准确性。

Comments 23 pages,5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.15503 2025-11-25 cs.CL

Llama2Vec: Unsupervised Adaptation of Large Language Models for Dense Retrieval

Llama2Vec: 无监督适应大型语言模型用于密集检索

Zheng Liu, Chaofan Li, Shitao Xiao, Yingxia Shao, Defu Lian

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Beijing Academy of Artificial Intelligence(北京人工智能研究院) University of Science and Technology of China(中国科学技术大学) The Hong Kong Polytechnic University(香港理工大学)

AI总结 Llama2Vec通过无监督适应LLM提升密集检索性能,实现新的最先进结果。

Comments ACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏