FastCache: Fast Caching for Diffusion Transformer Through Learnable Linear Approximation
通过可学习线性近似加速扩散变换器的FastCache
机构 * Yale University(耶鲁大学) ; Columbia University(哥伦比亚大学) ; University of California, Los Angeles(加州大学洛杉矶分校) ; University of Wisconsin--Madison(威斯康星大学麦迪逊分校) ; Michigan State University(密歇根州立大学)
专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM
AI总结 本文提出FastCache,通过隐藏状态层面的缓存与压缩框架,利用模型内部表示中的冗余性加速扩散变换器推理,结合可学习线性近似方法,有效降低计算开销并保持生成质量。