Downloads · 30 days
2.4K
15% of all-time downloads
reaperdoesntknow/SMOLM2Prover-GGUF
SMOLM2Prover-GGUF is a text generation model from reaperdoesntknow. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
GGUF quantized version of the SMOLM2Prover model for use with llama.cpp and compatible runtimes.
Downloads · 30 days
2.4K
15% of all-time downloads
All-time downloads
16.1K
Public
Repo size
996 MB
Likes
0
Public
Click a slice to open those files.
.gguf996 MB · 100%
From the Hugging Face model README
GGUF quantized version of the SMOLM2Prover model for use with llama.cpp and compatible runtimes.
| File | Size | Quantization | Quality |
|---|---|---|---|
SMOLM2Prover.gguf | 692M | F16 | Original (no quantization) |
SMOLM2Prover-Q4_K_M.gguf | 258M | Q4_K_M | Recommended (good quality/size balance) |
# Run with the quantized model
./llama-cli -m SMOLM2Prover-Q4_K_M.gguf -p "Your prompt here" -n 256
Create a Modelfile:
FROM ./SMOLM2Prover-Q4_K_M.gguf
Then:
ollama create smolm2prover -f Modelfile
ollama run smolm2prover
SMOLM2Prover-Q4_K_M.ggufThe Q4_K_M quantization uses:
Size reduction: 692M → 258M (63% smaller) BPW: 5.94 bits per weight
This model is part of the Convergent Intelligence LLC: Research Division portfolio. All models in this portfolio are developed under the Discrepancy Calculus (DISC) framework — a measure-theoretic approach to understanding and controlling the gap between what a model should produce and what it actually produces.
DISC treats training singularities (loss plateaus, mode collapse, catastrophic forgetting) not as failures to be smoothed over, but as structural signals that reveal the geometry of the learning problem. Key concepts:
For the full mathematical treatment, see Discrepancy Calculus: Foundations and Core Theory (DOI: 10.57967/hf/8194).
Citation chain: Structure Over Scale (DOI: 10.57967/hf/8165) → Three Teachers to Dual Cognition (DOI: 10.57967/hf/8184) → Discrepancy Calculus (DOI: 10.57967/hf/8194)
Same as the original model.
Part of the Standalone Models by Convergent Intelligence LLC: Research Division
| Model | Downloads | Format |
|---|---|---|
| SMOLM2Prover | 56 | HF |
| DeepReasoning_1R | 16 | HF |
| SAGI | 3 | HF |
| S-AGI | 0 | HF |
| Model | Downloads |
|---|---|
| Qwen3-1.7B-Thinking-Distil | 501 |
| LFM2.5-1.2B-Distilled-SFT | 342 |
| Qwen3-1.7B-Coder-Distilled-SFT | 302 |
| Qwen3-0.6B-Distilled-30B-A3B-Thinking-SFT-GGUF | 203 |
| Qwen3-1.7B-Coder-Distilled-SFT-GGUF | 194 |
Total Portfolio: 41 models | 2,781 total downloads
Last updated: 2026-03-28 12:55 UTC
<!-- CIX-CROSSLINK-START -->DistilQwen Collection — Our only BF16 series. Proof-weighted distillation from Qwen3-30B-A3B → 1.7B and 0.6B on H100. Three teacher variants (Instruct, Thinking, Coder), nine models, 2,788 combined downloads. The rest of the portfolio proves structure beats scale on CPU. This collection shows what happens when you give the methodology real hardware.
Top model: Qwen3-1.7B-Coder-Distilled-SFT — 508 downloads
Full methodology: Structure Over Scale (DOI: 10.57967/hf/8165)
Convergent Intelligence LLC: Research Division
<!-- CIX-CROSSLINK-END --><sub>Part of the reaperdoesntknow research portfolio — 48 models, 12,094 total downloads | Last refreshed: 2026-03-29 21:05 UTC</sub>
<!-- cix-keeper-ts:2026-10-03T13:16:46Z -->