arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Tsinghua University(清华大学)

2026-01-14 至 2026-01-14 共收录 14
2504.15054 2026-01-14 cs.CV

Structure-guided Diffusion Transformer for Low-Light Image Enhancement

结构引导的扩散变换器用于低光照图像增强

Xiangchen Yin, Zhenda Yu, Longtao Jiang, Xin Gao, Xiao Sun, Zhi Liu, Xun Yang

机构 * University of Science and Technology of China(中国科学技术大学) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院) Anhui University(安徽大学) School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动技术学院) State Key Laboratory of Automotive Safety and Energy, Tsinghua University(清华大学汽车安全与能源国家重点实验室) School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院) Department of Computer and Network Engineering, The University of Electro-Communications(电通大学计算机与网络工程系)

AI总结 本文提出结构引导的扩散变换器框架,通过小波变换和结构增强模块提升低光照图像增强效果,实现SOTA性能。

Comments Accepted by IEEE Transactions on Multimedia (TMM)

Journal ref IEEE Transactions on Multimedia, Vol. 27, pp. 9505 - 9515, 22 September 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08623 2026-01-14 cs.CV cs.AI cs.CR cs.LG

SafeRedir: Prompt Embedding Redirection for Robust Unlearning in Image Generation Models

SafeRedir: 基于提示嵌入重定向的图像生成模型鲁棒去学习

Renyang Liu, Kangjie Chen, Han Qiu, Jie Zhang, Kwok-Yan Lam, Tianwei Zhang, See-Kiong Ng

机构 * National University of Singapore(国立新加坡大学) Nanyang Technological University(南洋理工大学) Tsinghua University(清华大学) CFAR and IHPC, A*STAR(CFAR和IHPC,A*STAR)

AI总结 SafeRedir通过提示嵌入重定向实现图像生成模型的鲁棒去学习,无需修改模型即可有效消除不安全内容并提升对抗性攻击抵抗力。

Comments Code at https://github.com/ryliu68/SafeRedir

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08608 2026-01-14 cs.CV

SfMamba: Efficient Source-Free Domain Adaptation via Selective Scan Modeling

SfMamba: 通过选择性扫描建模实现高效的无源域适应

Xi Chen, Hongxun Yao, Sicheng Zhao, Jiankun Zhu, Jing Jiang, Kui Jiang

机构 * School of Computer Science and Technology, Harbin Institute of Technology(哈尔滨工业大学计算机科学与技术学院) Department of Psychological and Cognitive Sciences, Tsinghua University(清华大学心理学与认知科学系)

AI总结 SfMamba通过引入通道级视觉状态空间模块和语义一致的洗牌策略,实现了高效的无源域适应,提升了域不变特征提取和空间鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08273 2026-01-14 cs.CV cs.AI

HIPPO: Accelerating Video Large Language Models Inference via Holistic-aware Parallel Speculative Decoding

HIPPO:通过整体感知并行推测解码加速视频大语言模型推理

Qitan Lv, Tianyu Liu, Wen Wu, Xuenan Xu, Bowen Zhou, Feng Wu, Chao Zhang

机构 * University of Science and Technology of China(中国科学技术大学) Shanghai AI Laboratory(上海人工智能实验室) Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)

AI总结 HIPPO通过整体感知并行推测解码框架,有效提升视频大语言模型推理速度,实现高达3.51倍的加速效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08174 2026-01-14 cs.CV

Towards Cross-Platform Generalization: Domain Adaptive 3D Detection with Augmentation and Pseudo-Labeling

迈向跨平台泛化:基于增强与伪标签的域适应3D检测

Xiyan Feng, Wenbo Zhang, Lu Zhang, Yunzhi Zhuge, Huchuan Lu, You He

机构 * Dalian University of Technology(大连理工大学) Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)

AI总结 本研究提出一种基于增强与伪标签的域适应3D检测方法,在RoboSense2025挑战中取得第三名,实现了跨平台检测性能的提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07953 2026-01-14 quant-ph cs.AI

Quantum automated theorem proving

量子自动定理证明

Zheng-Zhi Sun, Qi Ye, Dong-Ling Deng

机构 * Center for Quantum Information, IIIS, Tsinghua University, Beijing 100084, China(量子信息中心、IIIS、清华大学、北京100084、中国) Shanghai Qi Zhi Institute, Shanghai 200232, China(上海启智研究所、上海200232、中国) Hefei National Laboratory, Hefei 230088, China(合肥国家实验室、合肥230088、中国)

AI总结 本文提出量子自动定理证明的通用框架,利用量子叠加和纠缠特性,实现命题逻辑和一阶逻辑的自动推理,并扩展了吴氏代数方法用于几何定理证明。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07894 2026-01-14 cs.LG cs.AI

Revealing the Attention Floating Mechanism in Masked Diffusion Models

揭示掩码扩散模型中的注意力漂浮机制

Xin Dai, Pengcheng Huang, Zhenghao Liu, Shuo Wang, Yukun Yan, Chaojun Xiao, Yu Gu, Ge Yu, Maosong Sun

机构 * School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)

AI总结 本文揭示了掩码扩散模型中注意力漂浮机制,通过动态分散的注意力锚点和浅层结构感知、深层内容聚焦的机制,解释了其在知识密集型任务中性能优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07877 2026-01-14 cs.LG cs.AI

E^2-LLM: Bridging Neural Signals and Interpretable Affective Analysis

E²-LLM:连接神经信号与可解释的情感分析

Fei Ma, Han Lin, Yifan Xie, Hongwei Ren, Xiaoyu Shen, Wenbo Ding, Qi Tian

机构 * Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室) Zhejiang University(浙江大学) Tsinghua University(清华大学) Harbin Institute of Technology(哈尔滨工业大学) Eastern Institute of Technology(东方技术研究所) Huawei(华为)

AI总结 E²-LLM通过整合预训练EEG编码器与Qwen-based LLM,实现了可解释的情绪分析,展示了模型扩展在情感识别和可解释性上的优势。

Comments 11 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07582 2026-01-14 cs.CL cs.AI

ES-Mem: Event Segmentation-Based Memory for Long-Term Dialogue Agents

基于事件分割的记忆:长期对话代理中的记忆

Huhai Zou, Tianhao Sun, Chuanjiang He, Yu Tian, Zhenyang Li, Li Jin, Nayu Liu, Jiang Zhong, Kaiwen Wei

机构 * College of Computer Science, Chongqing University(重庆大学计算机科学学院) Tsinghua University(清华大学) Hong Kong Generative AI Research & Development Center, HKUST(香港科技大学生成式人工智能研究与开发中心) Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空航天信息研究所) School of Computer Science and Technology, Tiangong University(天津理工大学计算机科学与技术学院)

AI总结 ES-Mem通过动态事件分割和分层记忆架构,提升长期对话代理的记忆效率和上下文定位能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14381 2026-01-14 cs.LG cs.AI cs.CL cs.CR

Are My Optimized Prompts Compromised? Exploring Vulnerabilities of LLM-based Optimizers

我的优化提示是否被 compromised?探索基于 LLM 的优化器的漏洞

Andrew Zhao, Reshmi Ghosh, Vitor Carvalho, Emily Lawton, Keegan Hines, Gao Huang, Jack W. Stokes

机构 * Tsinghua University(清华大学) Microsoft(微软)

AI总结 研究发现基于 LLM 的提示优化存在重大安全漏洞,提出假奖励攻击及轻量级防御措施,揭示优化流程为重要攻击目标。

Comments Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (EACL 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09644 2026-01-14 cs.CV cs.AI

DGAE: Diffusion-Guided Autoencoder for Efficient Latent Representation Learning

DGAE:基于扩散的自编码器用于高效潜在表示学习

Dongxu Liu, Jiahui Zhu, Yuang Peng, Haomiao Tang, Yuwei Chen, Chunrui Han, Zheng Ge, Daxin Jiang, Mingxue Liao

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Tsinghua University(清华大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)

AI总结 DGAE通过结合扩散模型提升解码器表达能力,实现高效潜在表示学习,在高压缩率下提升性能并减少潜在空间维度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08120 2026-01-14 cs.CV cs.AI cs.LG cs.MM

UniF$^2$ace: A Unified Fine-grained Face Understanding and Generation Model

UniF$^2$ace: 一个统一的细粒度人脸理解和生成模型

Junzhe Li, Sifan Zhou, Liya Guo, Xuerui Qiu, Linrui Xu, Delin Qu, Tingting Long, Chun Fan, Ming Li, Hehe Fan, Jun Liu, Shuicheng Yan

机构 * Peking University(北京大学) Carnegie Mellon University(卡内基梅隆大学) Tsinghua University(清华大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Central South University(中南大学) Fudan University(复旦大学) Guangming Lab(光明实验室) Zhejiang University(浙江大学) Lancaster University(兰卡斯特大学) National University of Singapore(新加坡国立大学)

AI总结 UniF$^2$ace是一个专门针对细粒度人脸理解和生成的统一多模态模型,通过双离散扩散损失和多级分组专家混合架构,提升高质量面部细节生成和细粒度属性处理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07145 2026-01-14 cs.CL cs.AI cs.LG

Stuffed Mamba: Oversized States Lead to the Inability to Forget

填充的Mamba:大状态导致无法遗忘

Yingfa Chen, Xinrong Zhang, Shengding Hu, Xu Han, Zhiyuan Liu, Maosong Sun

机构 * Department of Computer Science and Technology, Tsinghua University(计算机科学与技术系,清华大学)

AI总结 本文揭示了基于Mamba的模型因状态大小与训练长度不匹配导致无法有效遗忘的问题,并提出了改进长上下文建模的思路。

Comments COLM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.15718 2026-01-14 cs.CL

Beyond the Turn-Based Game: Enabling Real-Time Conversations with Duplex Models

超越回合制游戏:通过双工模型实现实时对话

Xinrong Zhang, Yingfa Chen, Shengding Hu, Xu Han, Zihang Xu, Yuanwei Xu, Weilin Zhao, Maosong Sun, Zhiyuan Liu

机构 * NLP Group, DCST, IAI, BNRIST, Tsinghua University, Beijing, China(自然语言处理组、国防科技大学、人工智能研究院、北京理工大学、清华大学、北京、中国) Quan Cheng Laboratory, Jinan, China(泉城实验室、济南、中国) Modelbest Inc.(Modelbest公司)

AI总结 本文提出双工模型,通过时间分割复用策略实现实时对话,提升用户与AI交互的自然度和满意度。

详情

展开后加载摘要…

URL PDF HTML 收藏