Deploy Hugging Face models on Google Cloud: from the hub to Inference Endpoints

Опубликовано: 01 Октябрь 2024
на канале: Julien Simon
857
18

In this series of three videos, I walk you through the deployment of Hugging Face models on Google Cloud, in three different ways:

- Deployment from the hub model page to Inference endpoints (this video), with the Google Gemma 7B model,
- Deployment from the hub model page to Vertex AI (   • Deploy Hugging Face models on Google ...  ), with the Microsoft Phi-2 2.7B model,
- Deployment directly from within Vertex AI (   • Deploy Hugging Face models on Google ...  ), with the TinyLlama 1.1B model.

Get started at https://huggingface.co :)

⭐️⭐️⭐️ Don't forget to subscribe to be notified of future videos. Follow me on Medium at   / julsimon   or Substack at https://julsimon.substack.com. ⭐️⭐️⭐️