🧨 Accelerating Stable Diffusion XL Inference with JAX on Cloud TPU v5e
Related stories
Accelerating Stable Diffusion Inference on Intel CPUs
🧨 Stable Diffusion in JAX / Flax !
Accelerating SD Turbo and SDXL Turbo Inference with ONNX Runtime and Olive
Faster Stable Diffusion with Core ML on iPhone, iPad, and Mac
Swift 🧨Diffusers - Fast Stable Diffusion for Mac
Speculative Sampling For Faster Molecular Dynamics
arXiv:2606. 02455v1 Announce Type: new Abstract: Molecular dynamics (MD) is a key tool for simulating the dynamical behavior of atomic systems.
Finetune Stable Diffusion Models with DDPO via TRL
Compiler-First State Space Duality and Portable $O(1)$ Autoregressive Caching for Inference
arXiv:2603. 09555v2 Announce Type: replace-cross Abstract: High-throughput Mamba-2 inference is usually tied to fused CUDA and Triton kernels, limiting portability across accelerator backends.
Fine-tuning Stable Diffusion models on Intel CPUs
JAXBench: Benchmarking Autonomous TPU Kernel Optimization
arXiv:2607. 20466v1 Announce Type: new Abstract: Rigorous benchmarks have driven progress in autonomous GPU kernel performance optimization by establishing a shared target to hillclimb on, but no equivalent exists for TPUs.
Wasserstein Convergence of ODE-Based Samplers in Decentralized Diffusion Model via Velocity Field Decomposition
arXiv:2606. 15835v1 Announce Type: cross Abstract: Diffusion models have achieved impressive empirical success in generative tasks, and their convergence theory is now relatively well understood.