Downloads · 30 days
47
1% of all-time downloads
Ex0bit/lfm-Nanotron
lfm-Nanotron is a text generation model from Ex0bit. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
<div align="center" </div <div align="center"
Downloads · 30 days
47
1% of all-time downloads
All-time downloads
3.3K
Public
Repo size
41.5 GB
Likes
3
Public
Click a slice to open those files.
.safetensors5.3 GB · 100%
From the Hugging Face model README

** LFM Architeture model SFT + GSPO RL + PRISM **
lfm-Nanotron: Limited Edition 2.6B PRISM Model Access. Unlock a cutting-edge Nano sized AI model!
This is lfm-Nanotron — A Nano Sized 2.6B parameter hybrid architecture language model fine-tuned with advanced techniques you won't find in mainstream releases:
| Parameter | Value |
|---|---|
| Parameters | ~2.6B |
| Hidden Size | 2048 |
| Layers | 30 (22 Conv + 8 Full Attention) |
| Attention Heads | 32 |
| KV Heads | 8 (GQA) |
| Vocabulary | 65,536 |
| Max Context | 128,000 tokens |
| Architecture | Hybrid Conv + Attention (LFM2) |
| File | Quantization | Size | Use Case |
|---|---|---|---|
lfm2-nanotron-ttft-gspo-prism-bf16.gguf | BF16 | ~4.8GB | Full precision, best quality |
lfm2-nanotron-ttft-gspo-prism-Q4_K_M.gguf (+W4A16) | Q4_K_M | ~1.5GB | Balanced quality/size |
lfm2-nanotron-ttft-gspo-prism-Q2_K.gguf | Q2_K (+W2A16) | ~0.9GB | Maximum compression |
./llama-cli -m lfm2-nanotron-ttft-gspo-prism-Q4_K_M.gguf -p "Your prompt here" --temp 0.3 --min-p 0.15 --repeat-penalty 1.05
{
"temperature": 0.3,
"min_p": 0.15,
"repeat_penalty": 1.05
}
If you use this model in your research, please cite:
@misc{lfm2-nanotron-2026,
title={lfm2-Nanotron: Test-Time Fine-Tuned LFM2 with GSPO+PRISM},
author={Exobit (Eric Elbaz)},
year={2026},
publisher={Hugging Face},
url={https://huggingface.co/Ex0bit/lfm2-Nanotron}
}
This model is released under a custom research license. See LICENSE.md for details.