Nvidia Inference Microservices - AI Workbench NIM-Anywhere Project Components

Опубликовано: 10 Октябрь 2024
на канале: Joe Freeman: Software Craft, Org Stuff, Tech Stuff
198
6

The video walks you through the containers and networking that make up the NIM-Anywhere project. The project is evolving so the topology and components will change..

NVidia Inference Microservices is a model for deploying inference engines as microservice endpoints. They demonstrate using this in the AI Workbench NIM-Anywhere project available on GitHub. The project contains a set of applications that consume the NIM inference engine, rerankers, and embedding models as API endpoints running in the NVIDIA cloud, your cloud, the data center, or local machines.

https://github.com/NVIDIA/nim-anywhere