torch.compile and Diffusers: A Hands-On Guide to Peak Performance - Sayak Paul, Hugging Face
About this talk
This talk covers the integration of torch.compile with the Diffusers library to enhance the performance of diffusion models like Flux-1-Dev. The speaker demonstrates practical techniques for both model authors and users, focusing on making models compiler-friendly and implementing regional compilation to significantly reduce compile times while maintaining runtime performance. Real-world scenarios are discussed, including optimizing resource use on memory-constrained GPUs through CPU offloading and quantization, as well as the efficient swapping of LoRA adapters without triggering recompilation. This session provides valuable insights for anyone involved in building or utilizing diffusion models.
More from this event
See all 103 talks →
What PyTorch Conference Europe 2026 Was Really Like – Official PyTorchCon EU Highlights | Paris
0:53
Lightning Talk: How DeepInverse Is Solving Imaging in Science and H... Andrew Wang & Minh Hai Nguyen
9:50
Why WideEP Inference Needs Data-Parallel-Aware Scheduling - Maroon Ayoub & Tyler Michael Smith
25:37
Write Once, Run Everywhere with Pytorch Transformers - Pedro Cuenca, Hugging Face
19:17