Downloads · 30 days
713
39% of all-time downloads
LiquidAI/LFM2.5-Encoder-350M-Spellchecker
LFM2.5-Encoder-350M-Spellchecker is a token classification model from LiquidAI. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as other.
<div align="center" <img src="https://cdn-uploads.huggingface.co/production/uploads/61b8e2ba285851687028d395/2b08LKpev0DNEk6DlnWkY.png" alt="Liquid AI" style="width: 100%; max-width: 100%; height: auto; display: inlin…
Downloads · 30 days
713
39% of all-time downloads
All-time downloads
1.9K
Public
Parameters
357M
2.1 GB on disk
Likes
15
Public
Click a slice to open those files.
.bin1.4 GB · 66%
From the Hugging Face model README
A full fine-tune of LFM2.5-Encoder-350M with a subword-level GECToR-style grammatical-error-correction tagger. It covers grammar, spelling, punctuation, and casing in English.
Find more details about our encoders in our blog post.
[!NOTE] 💻 Demos: Try this fine-tuned model running in a CPU-only Hugging Face space: Spell checking — correct misspellings token by token.
⚠️ Loads custom code via
trust_remote_code=True(the model wraps atrust_remote_codeencoder).
Install the required packages:
pip install torch transformers
Run spell checking:
from transformers import AutoModel
model_id = "LiquidAI/LFM2.5-Encoder-350-Spellchecker"
model = AutoModel.from_pretrained(
model_id,
trust_remote_code=True,
).float().eval()
print(model.correct(["She go to school every day ."]))
# ['She goes to school every day .']
correct() accepts a string or a list; tune precision with min_error_prob (higher → fewer edits) and
max_iter (refinement passes). Input should be whitespace-tokenized (punctuation separated by spaces),
matching the training data.
Fixed inference setting: max_iter=4, precision knobs off. Headline ERRANT F0.5:
MASTER composite (selection metric): 64.24
| Benchmark | Precision | Recall | F0.5 |
|---|---|---|---|
| LOCNESS native (ERRANT) | 53.77 | 34.53 | 48.38 |
| BEA-dev (ERRANT) | 56.48 | 28.96 | 47.46 |
| CoNLL-14 (ERRANT) | 67.54 | 18.91 | 44.59 |
| FCE-test (ERRANT) | 57.31 | 32.63 | 49.78 |
| Robustness (ERRANT) | 91.96 | 87.98 | 91.14 |
| Multilingual dev (F0.5) | — | — | — |
| Input | Correction |
|---|---|
She go to school every day . | She goes to school every day . |
I has went to the stor yesterday . | I went to the store yesterday . |
Their are many reason to study hard . | There are many reasons to study hard . |
He don't like coffee but he like tea . | He does n't like coffee , but he likes tea . |
@article{liquidAI2026Encoders,
author = {Liquid AI},
title = {LFM2.5-Encoders: Fast at Long Context, Even on CPU},
journal = {Liquid AI Blog},
year = {2026},
note = {www.liquid.ai/blog/lfm2-5-encoders},
}