Downloads · 30 days
34
7% of all-time downloads
renezander030/browserground-gguf
browserground-gguf is a image-text-to-text model from renezander030. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for gguf. The card lists the license as apache-2.0.
GGUF build of renezander030/browserground for llama.cpp, Ollama, and downstream wrappers that accept GGUF multimodal models.
Downloads · 30 days
34
7% of all-time downloads
All-time downloads
476
Public
Repo size
1.9 GB
Likes
0
Public
Click a slice to open those files.
.gguf1.9 GB · 100%
From the Hugging Face model README
GGUF build of renezander030/browserground
for llama.cpp, Ollama, and downstream wrappers that accept GGUF multimodal models.
Two files, both required:
| File | Purpose | Size |
|---|---|---|
browserground-Q4_K_M.gguf | text LLM, Q4_K_M quant | 1.11 GB |
browserground-mmproj-f16.gguf | vision tower (mmproj), f16 | 0.82 GB |
A ready-made Modelfile is in the repo. After downloading both .gguf files:
ollama create browserground -f Modelfile
ollama run browserground "Locate the Submit button" /path/to/screenshot.png
llama-mtmd-cli -m browserground-Q4_K_M.gguf --mmproj browserground-mmproj-f16.gguf --image screenshot.png -p "Locate the element described: Submit button"
npm install -g browserground
browserground parse screenshot.png --target "Submit button"
Recipe, numbers, full evaluation: https://huggingface.co/renezander030/browserground.
License: Apache 2.0 (inherits from Qwen/Qwen3-VL-2B-Instruct).