Downloads · 30 days
113
7% of all-time downloads
AtomicChat/Ornith-35B-MLX-4bit
Ornith-35B-MLX-4bit is a text generation model from AtomicChat. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as mit.
Downloads · 30 days
113
7% of all-time downloads
All-time downloads
1.6K
Public
Parameters
34.7B
19.5 GB on disk
Likes
1
Trending 1
Click a slice to open those files.
.safetensors19.5 GB · 100%
How the weights are stored.
U3234.7B · 100%
From the Hugging Face model README
Ornith 1.0 35B, self-quantized to MLX by Atomic Chat. Built straight from DeepReinforce's original weights with a per-tensor importance matrix, so this is not a repack of somebody else's files. Runs fully offline.
[!NOTE] These MLXs are self-quantized from the original weights, not a repack. The importance matrix keeps low-bit quants closer to the full-precision model.
| Property | Value |
|---|---|
| Base model | deepreinforce-ai/Ornith-1.0-35B |
| Parameters | 0.0B |
| Layers | 40 |
| Experts | 256 routed (top-8) |
| Context length | 262,144 tokens (256K) |
| Vocabulary | 248,320 |
| Modalities | Text, Image |
| Architecture | Mixture-of-Experts, 256 experts (top-8), 16 attention heads over 2 KV heads, Qwen3_5MoeForConditionalGeneration |
| This repo | MLX weights |
AtomicChat/ornith-35b-MLX-4bit and hit Use this model.mlx_lm.generate --model AtomicChat/ornith-35b-MLX-4bit --prompt "Hello" --max-tokens 512mlx_lm.server --model AtomicChat/ornith-35b-MLX-4bit --port 8080| Parameter | Value |
|---|---|
| temperature | 1.0 |
| top_p | 1.0 |
| top_k | 20 |
DeepReinforce's recommended sampling configuration for deepreinforce-ai/Ornith-1.0-35B.
deepreinforce-ai/Ornith-1.0-35B (original weights).mlx_lm.convert on our pipeline.Original model by DeepReinforce, released under the MIT license. Full terms: MIT. Quantized by Atomic Chat.