arXiv Machine Learning By Zhikang Xie, Xichen Ye, Yifan Wu, Haoshen Yu, Li chenan, Peizhu Gong, Weizhong Zhang, Cheng Jin

OnlineCache: Learning Dynamic Caching Policies with Error Correction for Efficient Diffusion Inference

Read the original on arXiv Machine Learning →

arXiv:2607. 29398v1 Announce Type: new Abstract: Diffusion models have revolutionized generative tasks but incur high latency due to iterative denoising.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.