I describe how to run Ollama Multimodal with local LlaVA LLM through Ollama. Advantage of this approach - you can process image documents with LLM directly, without running through OCR, this should lead to better results. This functionality is integrated as separate LLM agent into Sparrow.
Sparrow GitHub repo:
https://github.com/katanaml/sparrow
0:00 Intro
0:49 Example
3:24 Code
5:50 Summary
CONNECT:
Subscribe to this YouTube channel
Twitter: / andrejusb
LinkedIn: / andrej-baranovskij
Medium: / andrejusb
#rag #llm #llamaindex