Downloads · 30 days
0
Arjunros/Ollama_VLM
Ollama_VLM is a machine learning model from Arjunros. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This notebook demonstrates a multimodal interaction system using: - 🎤 Voice input (SpeechRecognition) - 📷 Live image input (OpenCV) - 🧠 Vision-Language reasoning (LLaVA via Ollama) - 🔊 Spoken responses (pyttsx3)
Downloads · 30 days
0
Access
Public
Updated Jul 14, 2025
Repo size
—
Likes
0
Public
Click a slice to open those files.
.ipynb10.7 KB · 83%
From the Hugging Face model README
This notebook demonstrates a multimodal interaction system using:
Perfect for building AI-powered robotic assistants!
Run locally with: pip install cv2 pyttsx3 speechrecognition ollama numpy
Ensure ollama and llava are properly set up on your machine.