Downloads · 30 days
118
41% of all-time downloads
tensor-tailor/ralph-qwen3-8b-binary
ralph-qwen3-8b-binary is a machine learning model from tensor-tailor. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
Binary-tier GGUF compression of Qwen/Qwen3-8B — architecture and parameter count unchanged (8,190,735,360 weight-bearing parameters), weights re-stored at reduced bit-width. Submitted to the binary bit-tier on Bittens…
Downloads · 30 days
118
41% of all-time downloads
All-time downloads
286
Public
Repo size
7.7 GB
Likes
0
Public
Click a slice to open those files.
.gguf1.9 GB · 100%
From the Hugging Face model README
Binary-tier GGUF compression of Qwen/Qwen3-8B — architecture and parameter count unchanged (8,190,735,360 weight-bearing parameters), weights re-stored at reduced bit-width. Submitted to the binary bit-tier on Bittensor subnet 40 (Ralph, model-compression).
Released under Apache License 2.0 — see LICENSE. This is a derivative of Qwen3-8B, itself
Apache-2.0; the original copyright/attribution notices and a description of what changed
(including the third-party quantization tooling used) are preserved in NOTICE.
| Parent | Qwen/Qwen3-8B (Apache-2.0) |
| Parameters | 8,190,735,360 |
| Architecture | unchanged from parent |
| Format | GGUF |
| Bit tier | binary |
| File size | 1,928,778,816 bytes (~1.80 GiB) |
| Measured bits/weight | ~1.88 (file size / param count; embedding/output tensors kept at higher precision outside the compressed blocks) |
| Quantization method | llama.cpp, imatrix-calibrated |
No retraining or architectural modification — this is a post-training weight re-storage of the parent's own weights.