arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Science and Technology of China(中国科学技术大学)

2026-01-05 至 2026-01-05 共收录 7
2601.00537 2026-01-05 cs.CV

Boosting Segment Anything Model to Generalize Visually Non-Salient Scenarios

提升 Segment Anything 模型以泛化视觉非显著场景

Guangqian Guo, Pengfei Chen, Yong Guo, Huafeng Chen, Boqiang Zhang, Shan Gao

机构 * Unmanned System Research Institute at Northwestern Polytechnical University(西北工业大学无人系统研究院) School of Electronic, Electrical, and Communication Engineering, University of Chinese Academic of Sciences(中国科学院大学电子电气与通信工程学院) Max Planck Institute for Informatics (MPI-INF)(马克斯·普朗克研究所(信息研究所)) University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出VNS-SAM,通过改进Segment Anything模型以提升对视觉非显著场景的分割性能和泛化能力。

Comments Accepted by IEEE TIP

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00535 2026-01-05 cs.CV

FreeText: Training-Free Text Rendering in Diffusion Transformers via Attention Localization and Spectral Glyph Injection

FreeText: 通过注意力定位和频谱字形注入实现免训练的文本渲染

Ruiqiang Zhang, Hengyi Wang, Chang Liu, Guanjie Wang, Zehua Ma, Weiming Zhang

机构 * Anhui Province Key Laboratory of Digital Security, University of Science and Technology of China(安徽省数字安全重点实验室,中国科学技术大学)

AI总结 FreeText通过注意力定位和频谱字形注入技术,无需训练即可提升文本渲染的准确性和美学质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00352 2026-01-05 cs.CV

OmniVaT: Single Domain Generalization for Multimodal Visual-Tactile Learning

OmniVaT:单域泛化用于多模态视觉-触觉学习

Liuxiang Qiu, Hui Da, Yuzhen Niu, Tiesong Zhao, Yang Cao, Zheng-Jun Zha

机构 * Fujian Key Laboratory for Intelligent Processing and Wireless Transmission of Media Information(福建智能媒体信息处理与无线传输重点实验室) College of Physics and Information Engineering(物理与信息工程学院) Fuzhou University(福州市大学) College of Computer and Data Science(计算机与数据科学学院) MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition(MoE脑启发智能感知与认知重点实验室) University of Science and Technology of China(中国科学技术大学)

AI总结 OmniVaT通过多模态分数傅里叶适配器和离散树生成模块,首次实现单域泛化多模态视觉-触觉学习任务,提升跨领域适应性与泛化性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00322 2026-01-05 cs.CV

Depth-Synergized Mamba Meets Memory Experts for All-Day Image Reflection Separation

深度协同Mamba遇见记忆专家用于全天图像反射分离

Siyan Fang, Long Peng, Yuntao Wang, Ruonan Wei, Yuehuan Wang

机构 * Huazhong University of Science and Technology(华中科技大学) University of Science and Technology of China(中国科学技术大学)

AI总结 DMDNet通过深度感知扫描和记忆专家补偿模块,提升全天图像反射分离性能,尤其在夜间表现更优。

Comments This paper has been accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00311 2026-01-05 cs.CV

ReMA: A Training-Free Plug-and-Play Mixing Augmentation for Video Behavior Recognition

ReMA:一种无需训练的即插即用混合增强方法用于视频行为识别

Feng-Qi Cui, Jinyang Huang, Sirui Zhao, Jinglong Guo, Qifan Cai, Xin Yan, Zhi Liu

机构 * University of Science and Technology of China(科学技术大学) Hefei University of Technology(合肥工业大学) Hefei Xiaosheng Intelligent Technology Co., Ltd.(合肥小生智能科技有限公司) Cylingo Group(Cylingo集团) The University of Electro-Communications(电通大学)

AI总结 ReMA通过受控的混合替换机制提升视频行为识别的表示鲁棒性,无需额外监督或参数,增强泛化与鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22428 2026-01-05 cs.LG stat.ML

Causality-Inspired Safe Residual Correction for Multivariate Time Series

基于因果的多变量时间序列安全残差校正

Jianxiang Xie, Yuncheng Hua, Mingyue Cheng, Flora Salim, Hao Xue

机构 * University of New South Wales(新南威尔士大学) University of Science and Technology of China(中国科学技术大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 本文提出基于因果的安全残差校正框架CRC,通过分而治之的策略确保多变量时间序列预测的非退化性和可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04344 2026-01-05 cs.CV

MICACL: Multi-Instance Category-Aware Contrastive Learning for Long-Tailed Dynamic Facial Expression Recognition

MICACL: 多实例类别感知对比学习用于长尾动态面部表情识别

Feng-Qi Cui, Zhen Lin, Xinlong Rao, Anyang Tong, Shiyao Li, Fei Wang, Changlin Chen, Bin Liu

机构 * University of Science and Technology of China(科学技术大学) Hefei University of Technology(合肥工业大学) IAI, Hefei Comprehensive National Science Center(IAI合肥国家科学中心)

AI总结 MICACL通过整合时空依赖建模和长尾对比学习优化,提升动态面部表情识别的鲁棒性和泛化能力。

Comments Accepted by IEEE ISPA2025

Journal ref 2025 IEEE International Symposium on Parallel and Distributed Processing with Applications (ISPA), Shenyang, China, 2025, pp. 601-608

详情

展开后加载摘要…

URL PDF HTML 收藏