Downloads · 30 days
136
31% of all-time downloads
junmingg/qwen2.5-coder-7b-text2sql-GGUF
qwen2.5-coder-7b-text2sql-GGUF is a text generation model from junmingg. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
GGUF quantizations of junmingg/qwen2.5-coder-7b-text2sql for CPU/GPU inference via llama.cpp, Ollama, LM Studio, etc.
Downloads · 30 days
136
31% of all-time downloads
All-time downloads
440
Public
Repo size
19 GB
Likes
0
Public
Click a slice to open those files.
.gguf19 GB · 100%
From the Hugging Face model README
GGUF quantizations of junmingg/qwen2.5-coder-7b-text2sql
for CPU/GPU inference via llama.cpp, Ollama,
LM Studio, etc.
| File | Quant | Size | Notes |
|---|---|---|---|
qwen2.5-coder-7b-instruct.Q4_K_M.gguf | Q4_K_M | ~4.7 GB | 4-bit, best size/quality balance (recommended) |
qwen2.5-coder-7b-instruct.Q6_K.gguf | Q6_K | ~6.3 GB | 6-bit, near-Q8 quality, smaller |
qwen2.5-coder-7b-instruct.Q8_0.gguf | Q8_0 | ~8.1 GB | 8-bit, near-lossless |
See the main model card for results (exact 78.8% / semantic 86.2% / validity 99.2% vs base 3.8 / 67.0 / 100), training details, and the required system prompt — the model expects the text-to-SQL system message + ChatML format.
ollama run hf.co/junmingg/qwen2.5-coder-7b-text2sql-GGUF:Q4_K_M
llama-cli -hf junmingg/qwen2.5-coder-7b-text2sql-GGUF:Q4_K_M
License: Apache-2.0 (base model) / data CC-BY-4.0.