In this series of three videos, I walk you through the deployment of Hugging Face models on Google Cloud, in three different ways:
Deployment from the hub model page to Inference endpoints ( • Deploy Hugging Face models on Google Cloud... ), with the Google Gemma 7B model,
Deployment from the hub model page to Vertex AI ( • Deploy Hugging Face models on Google Cloud... ), with the Microsoft Phi-2 2.7B model,
Deployment directly from within Vertex AI (this video), with the TinyLlama 1.1B model.
Get started at https://huggingface.co :)
⭐️⭐️⭐️ Don't forget to subscribe to be notified of future videos. Follow me on Medium at / julsimon or Substack at https://julsimon.substack.com. ⭐️⭐️⭐️