arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

大厂专区

Huawei(华为)

2026-05-18 至 2026-05-18 共收录 9
2605.16113 2026-05-18 cs.CL cs.AI

DebiasRAG: A Tuning-Free Path to Fair Generation in Large Language Models through Retrieval-Augmented Generation

DebiasRAG: 通过检索增强生成实现大型语言模型中公平生成的无调优路径

Rui Chu, Bingyin Zhao, Thanh Quoc Hung Le, Duy Cao Hoang, Huawei Lin, Ping Li, Weijie Zhao, Khoa D Doan, Yingjie Lao

机构 * Huawei(华为)

AI总结 本文提出DebiasRAG,一种基于检索增强生成的无调优动态查询特定去偏框架,通过生成查询特定去偏候选、构建上下文候选池和梯度更新去偏引导上下文重排序三阶段,提升生成公平性并保留LLM固有属性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.16045 2026-05-18 cs.CL cs.AI cs.LG

RecMem: Recurrence-based Memory Consolidation for Efficient and Effective Long-Running LLM Agents

RecMem:基于递归的记忆巩固用于高效且有效的长运行LLM代理

Zijie Dai, Shiyuan Deng, Sheng Guan, Yizhou Tian, Xin Yao, Xiao Yan, James Cheng

机构 * Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系) School of Computer Science, Beijing University of Posts and Telecommunications(北京邮电大学计算机学院) Huawei Cloud(华为云) Huawei Theory Lab(华为理论实验室) Institute for Math and AI, Wuhan University(武汉大学数学与人工智能研究院)

AI总结 RecMem通过递归机制优化内存巩固,减少token消耗并提升准确性,有效解决长运行LLM代理的内存管理问题。

Comments Accepted to ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15843 2026-05-18 cs.CV

WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes

WorldAct:将单体3D世界激活为可交互的以对象为中心的场景

Jichen Hu, Jiawei Guo, Jiazhong Cen, Chen Yang, Sikuang Li, Wei Shen

机构 * Shanghai Jiao Tong University(上海交通大学) Huawei Inc(华为公司)

AI总结 WorldAct通过多模态代理将静态生成的3D世界分解为可编辑的交互场景,支持对象级编辑和任务执行,保留全局一致性。

Comments Project page: https://sjtu-deepvisionlab.github.io/WorldAct

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.12309 2026-05-18 cs.CV

G$^2$TR: Generation-Guided Visual Token Reduction for Separate-Encoder Unified Multimodal Models

G$^2$TR: 基于生成的视觉标记减少方法用于分离编码统一多模态模型

Junxian Li, Kai Liu, Zizhong Ding, Zhixin Wang, Zhikai Chen, Renjing Pei, Yulun Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Huawei Technologies Ltd(华为技术有限公司)

AI总结 本文提出G$^2$TR方法,通过生成分支信号减少多模态模型的视觉标记,提升效率并保持性能,实验显示在图像理解和编辑任务中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15701 2026-05-18 cs.CL cs.AI

H-Mem: A Novel Memory Mechanism for Evolving and Retrieving Agent Memory via a Hybrid Structure

H-Mem: 一种通过混合结构进化和检索智能体记忆的新型记忆机制

Jiawei Yu, Yixiang Fang, Xilin Liu, Yuchi Ma

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Huawei Cloud Computing Technologies CO., LTD.(华为云计算技术有限公司)

AI总结 H-Mem通过混合结构有效建模智能体记忆的长期演化并高效检索记忆数据,提升问答任务性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15688 2026-05-18 stat.ML cs.AI cs.LG math.PR

$α$-TCAV: A Unified Framework for Testing with Concept Activation Vectors

$α$-TCAV:基于概念激活向量的测试统一框架

Ekkehard Schnoor, Jawher Said, Malik Tiomoko, Wojciech Samek, Alexander Jung

机构 * Department of Computer Science(计算机科学系) Department of Artificial Intelligence(人工智能系) Aalto University(阿alto大学) Fraunhofer Heinrich Hertz Institute(弗劳恩霍夫海因里希·赫兹研究所) Department of Artificial Intelligence, Fraunhofer HHI(人工智能系,弗劳恩霍夫HHI研究所) Huawei Noah’s Ark Lab(华为诺亚实验室) Department of EECS, Technische Universität Berlin(电子工程与计算机科学系,柏林技术大学)

AI总结 本文提出$α$-TCAV框架,解决传统TCAV方法中因指示函数不连续导致的方差问题,通过参数化平滑函数统一概率表述,并提供参数调优指导,挑战现有实践惯例。

Comments 44 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15684 2026-05-18 cs.CV

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices

ElasticDiT:通过弹性架构和稀疏注意力实现高效扩散变换器,用于移动设备上的高分辨率图像生成

Kunpeng Du, Haizhen Xie, Sen Lu, Lei Yu, Binglei Bao, Huaao Tang, Chuntao Liu, Hao Wu, Yang Zhao, Zhicai Huang, Heyuan Gao, Zhijun Tu, Jie Hu, Xinghao Chen

机构 * Huawei Technologies(华为技术)

AI总结 本文提出ElasticDiT,通过弹性架构和稀疏注意力机制,在移动设备上实现高效扩散变换器,平衡图像质量和计算效率,同时减少内存占用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15609 2026-05-18 cs.CL

PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding

PSD: 推动扩散大语言模型的帕累托前沿:通过并行推测解码

Shengyin Sun, Yiming Li, Renxi Liu, Xinqi Li, Hui-Ling Zhen, Weizhe Lin, Chen Chen, Xianzhi Yu, Mingxuan Yuan, Chen Ma

机构 * Huawei Technologies(华为技术)

AI总结 本文提出PSD框架,通过并行推测解码提升推理效率与生成质量,在推理效率和生成质量之间取得良好平衡,达到每前向传递5.5倍的token处理速度。

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.05467 2026-05-18 cs.LG cs.NA math.NA

ChebNet: Efficient and Stable Constructions of Deep Neural Networks with Rectified Power Units via Chebyshev Approximations

ChebNet: 通过切比雪夫近似高效稳定地构建深度神经网络

Shanshan Tang, Bo Li, Haijun Yu

机构 * Software Development Center, Industrial and Commercial Bank of China(中国工商银行软件开发中心) Hisilicon Semiconductor and Component Business Dept.(2012 Labs), Huawei Technologies Co., Ltd(华为技术有限公司半导体及组件业务部) NCMIS & LSEC, Institute of Computational Mathematics and Scientific/Engineering Computing, Academy of Mathematics and Systems Science, Beijing(数学与系统科学研究院) School of Mathematical Sciences, University of Chinese Academy of Sciences(中国科学院大学数学科学学院)

AI总结 本文提出ChebNet,利用切比雪夫多项式近似构建稳定高效的深度神经网络,实现与幂级数方法相当的函数逼近性能,并在稳定性上更优。

Comments 6 figures, 3 tables, to appear on Communications in Mathematics and Statistics

Journal ref Communications in Mathematics and Statistics, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏