Downloads · 30 days
70
29% of all-time downloads
Top4eat21/Ling-3.0-tiny-GGUF-APEX
Ling-3.0-tiny-GGUF-APEX is a machine learning model from Top4eat21. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
A compressed mixed-precision MoE quantization of Ling 3.0 Tiny.
Downloads · 30 days
70
29% of all-time downloads
All-time downloads
241
Public
Repo size
5 GB
Likes
0
Public
Click a slice to open those files.
.gguf5 GB · 100%
From the Hugging Face model README
A compressed mixed-precision MoE quantization of Ling 3.0 Tiny.
The quantization uses a mixed-precision approach rather than applying a single quantization format to every tensor. Different parts of the MoE are assigned different precisions to preserve important model capabilities while aggressively reducing the size of the large expert weights. (Thanks APEX)
It is ~2.9GB so it will fit in most small GPUs and genorate at a moderately good speed.
Ling-3.0-tiny-TQ1_0.gguf was made using normal llama.cpp tooling and is very dumb and not recomended. It is 1.8GB.
This is an unofficial community quantization and is not affiliated with or endorsed by InclusionAI.