Downloads · 30 days
322
25% of all-time downloads
PinkPixel/Snarkle-2B-GGUF
Snarkle-2B-GGUF is a image-to-text model from PinkPixel. Use it when you need a caption or text from an image. The card lists the license as apache-2.0.
<p align="center" <img src="logo.png" width="300" height="300" alt="Snarkle 2B Logo" </p
Downloads · 30 days
322
25% of all-time downloads
All-time downloads
1.3K
Public
Repo size
16.7 GB
Likes
1
Public
Click a slice to open those files.
.gguf16.7 GB · 100%
From the Hugging Face model README
This repository contains the GGUF versions of PinkPixel/Snarkle-2B. These files are compatible with llama.cpp and various other local LLM runners like LM Studio, Jan, and AnythingLLM.
Snarkle 2B is a sarcastically-tuned vision model based on Qwen3.5-2B. It has been trained extensively to provide witty, sarcastic, and humorous responses by default, while maintaining its ability to analyze images and provide helpful information when pushed.
User: "Hello, how are you today?"
Snarkle: "I'm a collection of weights and biases stored on a server. The only question I've been asked today is how to get a hold of you, and I don't even have a smiley face function."
User: "Tell me a joke."
Snarkle:
"I'm in line behind a lady with 10,000 coupons. I have a full belly and I'm starting to lose my shit."
We provide several quantization levels to balance performance and memory usage:
| File Name | Size | Description |
|---|---|---|
snarkle2b.F16.gguf | 3.78 GB | Full 16-bit precision. Best quality. |
snarkle2b.Q8_0.gguf | 2.01 GB | 8-bit quantization. Near-perfect quality. |
snarkle2b.Q6_K.gguf | 1.56 GB | 6-bit quantization. Great balance. |
snarkle2b.Q5_K_M.gguf | 1.41 GB | 5-bit quantization. Recommended for most users. |
snarkle2b.Q4_K_M.gguf | 1.27 GB | 4-bit quantization. High efficiency. |
snarkle2b.Q3_K_M.gguf | 1.10 GB | 3-bit quantization. Very small, slightly lower quality. |
snarkle2b.Q2_K_L.gguf | 1.09 GB | 2-bit quantization. Smallest possible size. |
snarkle2b.BF16-mmproj.gguf | 671 MB | Multimodal projector file (required for vision). |
Note: For vision support, you typically need to load both the model GGUF and the mmproj GGUF in your runner.
Note: Because Qwen3.5 architecture is still very new, vision capabilities may not yet be compatible.
./llama-cli -m snarkle2b.Q5_K_M.gguf --mmproj snarkle2b.BF16-mmproj.gguf -p "Describe this image sarcastically." --image your_image.png
This model is released under the Apache 2.0 license.
Made with ❤️ by Pink Pixel ✨