Downloads · 30 days
39
43% of all-time downloads
ToldByNun/mango-1.0
mango-1.0 is a text generation model from ToldByNun. Use it when you need the model to write or continue text. It is set up for gguf. The card lists the license as apache-2.0.
Local coding / agent model for Mango — a desktop coding agent built around small GGUF runtimes (llama.cpp). This repo ships a quantized GGUF of Mango 1.0 (~27B, Qwen3.8 family), fine-tuned for tool-using agent workflo…
Downloads · 30 days
39
43% of all-time downloads
All-time downloads
91
Public
Repo size
12.1 GB
Likes
0
Public
Click a slice to open those files.
.gguf12.1 GB · 100%
From the Hugging Face model README
Local coding / agent model for Mango — a desktop coding agent built around small GGUF runtimes (llama.cpp). This repo ships a quantized GGUF of Mango 1.0 (~27B, Qwen3.8 family), fine-tuned for tool-using agent workflows (short thoughts, tool calls, file edits, Q&A over a workspace).
| Family | Qwen3.8 (~27B) |
| Format | GGUF |
| Quant | Q2_K_L |
| File | mango-1.0-Q2_K_L.gguf (~12.1 GB) |
| Context | up to 262k (runtime-dependent; use what your VRAM allows) |
| Chat template | Qwen / ChatML (`< |
| Intended use | Local coding agent, tool calling, repo Q&A |
mango-1.0-Q2_K_L.gguf.gguf path/ask, /plan, or agent mode as usualLoad the GGUF like any other Qwen ChatML model. Prefer a GPU offload that fits your VRAM; leave layers on CPU if needed. Example (llama.cpp CLI sketch):
./llama-cli -m mango-1.0-Q2_K_L.gguf -c 8192 -ngl 99 -p "You are a coding assistant."
training/.Apache-2.0 (unless otherwise noted for the base model / training data — check base model cards as well).
@misc{mango10_gguf,
title = {Mango 1.0 GGUF},
author = {ToldByNun},
year = {2026},
howpublished = {\url{https://huggingface.co/ToldByNun/mango-1.0-iq2-xs}},
note = {Qwen3.8-based local coding agent model, Q2\_K\_L GGUF}
}