In this lightboard video, we will learn how organizations can run their open-source Large Language Models from Hugging Face on Kubernetes cluster on-premises using Portworx. We also talk about how organizations can avoid hallucinations by building a Retrieval Augmented Generation architecture using Vector databases running on Kubernetes and Portworx.