Deploy Hugging Face models on Google Cloud: directly from Vertex AI

Опубликовано: 29 Март 2026
на канале: Julien Simon
11,983
103

In this series of three videos, I walk you through the deployment of Hugging Face models on Google Cloud, in three different ways:

Deployment from the hub model page to Inference endpoints (   • Deploy Hugging Face models on Google Cloud...  ), with the Google Gemma 7B model,
Deployment from the hub model page to Vertex AI (   • Deploy Hugging Face models on Google Cloud...  ), with the Microsoft Phi-2 2.7B model,
Deployment directly from within Vertex AI (this video), with the TinyLlama 1.1B model.

Get started at https://huggingface.co :)

⭐️⭐️⭐️ Don't forget to subscribe to be notified of future videos. Follow me on Medium at   / julsimon   or Substack at https://julsimon.substack.com. ⭐️⭐️⭐️