arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Science and Technology of China(中国科学技术大学)

2025-12-09 至 2025-12-09 共收录 14
2512.07480 2025-12-09 cs.CV

Single-step Diffusion-based Video Coding with Semantic-Temporal Guidance

单步扩散视频编码与语义-时间引导

Naifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li, Zihan Zheng, Yuan Zhang, Yan Lu

机构 * Communication University of China(中国通信大学) University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

AI总结 S2VC通过单步扩散与语义-时间引导,实现低比特率下的高质量视频编码,比特率节省达52.73%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07464 2025-12-09 cs.RO

Gait-Adaptive Perceptive Humanoid Locomotion with Real-Time Under-Base Terrain Reconstruction

具有实时底层地形重建的感知人体机器人运动

Haolin Song, Hongbo Zhu, Tao Yu, Yan Liu, Mingqi Yuan, Wengang Zhou, Hua Chen, Houqiang Li

机构 * Department of Electronic Engineering and Information Science (EEIS), University of Science and Technology of China(电子工程与信息科学系,中国科学技术大学) LimX Dynamics(LimX动力学) Hong Kong University of Science and Technology(香港科技大学) School of Mechanics Engineering, Harbin Institute of Technology (HIT)(机械工程学院,哈尔滨工业大学) Department of Computing, The Hong Kong Polytechnic University(计算学院,香港理工大学) Zhejiang University-University of Illinois Urbana-Champaign Institute (ZJUI)(浙江大学-伊利诺伊大学厄巴纳-香槟分校联合研究所)

AI总结 该研究提出了一种结合地形感知、步态调节和全身控制的强化学习框架,通过实时底层地形重建实现稳健的人形机器人运动。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07208 2025-12-09 cs.LG cs.AI

Geometric Prior-Guided Federated Prompt Calibration

几何先验引导的联邦提示校准

Fei Luo, Ziwei Zhao, Mingxuan Wang, Duoyang Li, Zhe Qian, Jiayi Tuo, Chenyue Zhou, Yanbiao Ma

机构 * Jishou University(吉首大学) Technical University of Munich(慕尼黑技术大学) Renmin University of China(中国人民大学) Northwestern Polytechnical University(西北工业大学) South China Agricultural University(华南农业大学) University of Science and Technology of China(中国科学技术大学) Nanjing University of Aeronautics and Astronautics(南京航空航天大学)

AI总结 本文提出GGTPC框架,通过引入全局几何先验纠正本地训练偏见,提升联邦学习在数据异质性下的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00975 2025-12-09 cs.CV cs.LG cs.RO

MM-ACT: Learn from Multimodal Parallel Generation to Act

MM-ACT: 从多模态并行生成中学习以行动

Haotian Liang, Xinyi Chen, Bin Wang, Mingkang Chen, Yitian Liu, Yuhao Zhang, Zanxin Chen, Tianshuo Yang, Yilun Chen, Jiangmiao Pang, Dong Liu, Xiaokang Yang, Yao Mu, Wenqi Shao, Ping Luo

机构 * Shanghai AI Laboratory(上海人工智能实验室) Shanghai Jiao Tong University(上海交通大学) The University of Hong Kong(香港大学) University of Science and Technology of China(中国科学技术大学) Fudan University(复旦大学) Zhejiang University(浙江大学)

AI总结 MM-ACT通过多模态并行生成提升机器人任务执行能力,实现96.3%的成功率。

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19296 2025-12-09 cs.LG cs.AR cs.PL

QiMeng-SALV: Signal-Aware Learning for Verilog Code Generation

QiMeng-SALV:面向Verilog代码生成的信号感知学习

Yang Zhang, Rui Zhang, Jiaming Guo, Lei Huang, Di Huang, Yunpu Zhao, Shuyao Cheng, Pengwei Jin, Chongxiao Li, Zidong Du, Xing Hu, Qi Guo, Yunji Chen

机构 * State Key Lab of Processors, Institute of Computing Technology, CAS(处理器国家重点实验室,计算技术研究所,中国科学院) University of Chinese Academy of Sciences(中国科学院大学) University of Science and Technology of China(中国科学技术大学)

AI总结 QiMeng-SALV通过信号感知学习提升Verilog代码生成的准确性与性能,采用信号级优化解决功能奖励不足问题。

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14354 2025-12-09 cs.LG cs.CL

Self-Improvement Towards Pareto Optimality: Mitigating Preference Conflicts in Multi-Objective Alignment

迈向帕累托最优的自我改进:缓解多目标对齐中的偏好冲突

Moxin Li, Yuantao Zhang, Wenjie Wang, Wentao Shi, Zhuo Liu, Fuli Feng, Tat-Seng Chua

机构 * National University of Singapore(新加坡国立大学) University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出一种自我改进的DPO框架,通过生成帕累托最优响应缓解多目标对齐中的偏好冲突,提升模型在帕累托前沿的优化效果。

Comments ACL findings (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08519 2025-12-09 cs.CL

Bridging Relevance and Reasoning: Rationale Distillation in Retrieval-Augmented Generation

弥合相关性与推理:检索增强生成中的推理蒸馏

Pengyue Jia, Derong Xu, Xiaopeng Li, Zhaocheng Du, Xiangyang Li, Yichao Wang, Yuhao Wang, Qidong Liu, Maolin Wang, Huifeng Guo, Ruiming Tang, Xiangyu Zhao

机构 * City University of Hong Kong(香港城市大学) University of Science and Technology of China(中国科学技术大学) Huawei Noah’s Ark Lab(华为诺亚实验室)

AI总结 RADIO通过推理提取和基于推理的对齐方法,弥合检索增强生成中重排器与生成器之间的相关性差距。

Comments Accepted to ACL 25 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09388 2025-12-09 cs.CV

Diffusion Models for Image Restoration and Enhancement: A Comprehensive Survey

扩散模型在图像修复与增强中的应用:全面综述

Xin Li, Yulin Ren, Xin Jin, Cuiling Lan, Xingrui Wang, Wenjun Zeng, Xinchao Wang, Zhibo Chen

机构 * University of Science and Technology of China(科学技术大学) National University of Singapore(国立新加坡大学) Eastern Institute for Advanced Study(东部高级研究机构) Microsoft Research Asia(微软亚洲研究院)

AI总结 本文首次全面综述了基于扩散模型的图像修复与增强方法,涵盖学习范式、条件策略、框架设计等,并提出未来研究方向。

Comments Accepted by IJCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06886 2025-12-09 cs.CV

Balanced Learning for Domain Adaptive Semantic Segmentation

领域自适应语义分割中的平衡学习

Wangkai Li, Rui Sun, Bohao Liao, Zhaoyang Li, Tianzhu Zhang

机构 * MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition(脑启发智能感知与认知关键实验室) University of Science and Technology of China(中国科学技术大学) Deep Space Exploration Laboratory(深空探测实验室)

AI总结 BLDA通过分析logits分布和引入共享锚定分布,有效缓解领域自适应语义分割中的类别偏倚问题,提升模型在欠预测类别上的性能。

Comments Accepted by International Conference on Machine Learning (ICML 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06865 2025-12-09 cs.CV

Spatial Retrieval Augmented Autonomous Driving

空间检索增强的自动驾驶

Xiaosong Jia, Chenhe Zhang, Yule Jiang, Songbur Wong, Zhiyuan Zhang, Chen Chen, Shaofeng Zhang, Xuanhe Zhou, Xue Yang, Junchi Yan, Yu-Gang Jiang

机构 * Institute of Trustworthy Embodied AI, Fudan University(可信具身人工智能研究院,复旦大学) Shanghai Jiao Tong University(上海交通大学) Key Laboratory of Target Cognition and Application Technology, Aerospace Information Research Institute, Chinese Academy of Sciences(目标认知与应用技术重点实验室,航天信息研究所,中国科学院) University of Science and Technology of China(中国科学技术大学)

AI总结 本文提出空间检索范式,通过引入离线地理图像提升自动驾驶任务性能,扩展nuScenes数据集并建立多个基准测试。

Comments Demo Page: https://spatialretrievalad.github.io/ with open sourced code, dataset, and checkpoints

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06641 2025-12-09 cs.IR cs.CL

An Index-based Approach for Efficient and Effective Web Content Extraction

基于索引的方法用于高效有效的网络内容提取

Yihan Chen, Benfeng Xu, Xiaorui Wang, Zhendong Mao

机构 * University of Science and Technology of China(中国科学技术大学) Metastone Technology(Metastone技术)

AI总结 本文提出基于索引的方法,用于高效提取网络相关内容,提升RAG系统查询准确性与处理速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22634 2025-12-09 cs.RO cs.SE

LabUtopia: High-Fidelity Simulation and Hierarchical Benchmark for Scientific Embodied Agents

LabUtopia:高保真模拟与分层基准用于科学具身智能体

Rui Li, Zixuan Hu, Wenxi Qu, Jinouwen Zhang, Zhenfei Yin, Sha Zhang, Xuantuo Huang, Hanqing Wang, Tai Wang, Jiangmiao Pang, Wanli Ouyang, Lei Bai, Wangmeng Zuo, Ling-Yu Duan, Dongzhan Zhou, Shixiang Tang

机构 * Shanghai AI Laboratory(上海人工智能实验室) Peking University(北京大学) Oxford(牛津大学) The Chinese University of Hong Kong(香港中文大学) Harbin Institute of Technology(哈尔滨工业大学) University of Science and Technology of China(中国科学技术大学)

AI总结 LabUtopia通过高保真模拟和分层基准推动实验室环境中具身智能的发展。

Comments Accepted by NeurIPS 2025 Dataset and Benchmark Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10339 2025-12-09 cs.CV

Towards Unsupervised Domain Bridging via Image Degradation in Semantic Segmentation

通过图像退化实现无监督领域桥接的语义分割方法

Wangkai Li, Rui Sun, Huayu Mai, Tianzhu Zhang

机构 * MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知联合实验室,中国科学技术大学) National Key Laboratory of Deep Space Exploration, Deep Space Exploration Laboratory(国家深空探测重点实验室,深空探测实验室)

AI总结 DiDA通过图像退化构建中间领域并补偿语义偏移,提升语义分割在不同领域间的适应性能。

Comments Accepted by Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00238 2025-12-09 cs.CV cs.AI

Twisted Convolutional Networks (TCNs): Enhancing Feature Interactions for Non-Spatial Data Classification

扭曲卷积网络(TCNs):增强非空间数据分类的特征交互

Junbo Jacob Lian, Haoran Chen, Kaichen Ouyang, Yujun Zhang, Rui Zhong, Huiling Chen

机构 * McCormick School of Engineering, Northwestern University(西北大学工程学院) School of Mathematics, University of Science and Technology of China(中国科学技术大学数学系) College of New Energy, Jingchu University of Technology(荆楚科技学院新能源学院) Information Initiative Center, Hokkaido University(北海道大学信息初始化中心) School of Computer Science and Artificial Intelligence, Wenzhou University(温州大学计算机科学与人工智能学院)

AI总结 TCNs通过理论支持的乘法和成对交互机制增强非空间数据分类的特征交互,实现比CNNs、ResNet等方法更显著的性能提升。

Comments The source code for the TCNs can be accessed at https://github.com/junbolian/Twisted-Convolutional-Networks

详情

展开后加载摘要…

URL PDF HTML 收藏