Downloads · 30 days
35
10% of all-time downloads
leok7v/KeyLM-75M-Instruct-GGUF
KeyLM-75M-Instruct-GGUF is a text generation model from leok7v. Use it when you need the model to write or continue text. It is set up for gguf. The card lists the license as apache-2.0.
GGUF builds of KeyLM-75M-Instruct for llama.cpp, LM Studio, Ollama, and other GGUF runtimes.
Downloads · 30 days
35
10% of all-time downloads
All-time downloads
345
Public
Repo size
125 MB
Likes
0
Public
Click a slice to open those files.
.gguf125 MB · 100%
From the Hugging Face model README
GGUF builds of KeyLM-75M-Instruct for llama.cpp, LM Studio, Ollama, and other GGUF runtimes.
KeyLM is a 75M-parameter instruction-tuned language model trained from scratch on approximately 18 billion tokens. See the main model card for benchmarks, training details, limitations, and the transformers (safetensors) version.
| File | Quant | Size | Notes |
|---|---|---|---|
KeyLM-75M-Instruct.F16.gguf | F16 | ~144 MB | Full precision and recommended. The model is already tiny, so there is little reason to quantize further. |
# straight from the Hub
llama-cli -hf Eclipse-Senpai/KeyLM-75M-Instruct-GGUF -cnv
# or a local file
llama-cli -m KeyLM-75M-Instruct.F16.gguf -cnv
The chat template (User: / Assistant:, assistant turns ending with </s>) is embedded in the GGUF, so conversation mode (-cnv) applies it automatically.
.gguf; the embedded chat template is detected automatically.ollama run hf.co/Eclipse-Senpai/KeyLM-75M-Instruct-GGUFKeyLM is a tiny model: good at simple instruction following and short chat, near random chance on knowledge/reasoning benchmarks. It is not a factual assistant. Full numbers and caveats are on the main model card.
Apache 2.0.