In this video, you will learn how to accelerate image generation with an Intel Sapphire Rapids server. Using Stable Diffusion models, the Intel Extension for PyTorch and system-level optimizations, we're going to cut inference latency from 36+ seconds to 5 seconds!
⭐️⭐️⭐️ Don't forget to subscribe to be notified of future videos ⭐️⭐️⭐️
⭐️⭐️⭐️ Want to buy me a coffee? I can always use more :) https://www.buymeacoffee.com/julsimon ⭐️⭐️⭐️
Blog post: https://huggingface.co/blog/stable-di...
Code: https://gitlab.com/juliensimon/huggin...
Jemalloc: https://jemalloc.net/
Intel Extension for PyTorch: https://github.com/intel/intel-extens...
Intel Sapphire Rapids: https://en.wikipedia.org/wiki/Sapphir...
Intel Advanced Matrix Extensions: https://en.wikipedia.org/wiki/Advance...