In this video, you will learn how to accelerate a PyTorch training job with a cluster of Intel Sapphire Rapids servers running on AWS. We will use the Intel oneAPI Collective Communications Library (CCL) to distribute the job, and the Intel Extension for PyTorch (IPEX) library to automatically put the new CPU instructions to work. As both libraries are already integrated with the Hugging Face transformers library, we will be able to run our sample scripts out of the box without changing a line of code.
⭐️⭐️⭐️ Don't forget to subscribe to be notified of future videos ⭐️⭐️⭐️
⭐️⭐️⭐️ Want to buy me a coffee? I can always use more :) https://www.buymeacoffee.com/julsimon ⭐️⭐️⭐️
- Blog post: https://huggingface.co/blog/intel-sap...
- Intel Sapphire Rapids: https://en.wikipedia.org/wiki/Sapphir...
- Intel Advanced Matrix Entensions: https://en.wikipedia.org/wiki/Advance...
- Amazon EC2 R7iz: https://aws.amazon.com/ec2/instance-t...