Downloads · 30 days
14
10% of all-time downloads
aoiandroid/nllb200-coreml-1024
nllb200-coreml-1024 is a translation model from aoiandroid. Use it when you need text moved from one language to another. The card lists the license as cc-by-nc-4.0.
CoreML conversion of facebook/nllb-200-distilled-600M (NLLB-200 / M2M100, distilled 600M parameters, 200 languages) for on-device inference on Apple platforms. Max sequence length 1024 matches the official model maxpo…
Downloads · 30 days
14
10% of all-time downloads
All-time downloads
139
Public
Repo size
7.1 GB
Likes
0
Public
Click a slice to open those files.
.bin4.5 GB · 100%
From the Hugging Face model README
CoreML conversion of facebook/nllb-200-distilled-600M (NLLB-200 / M2M100, distilled 600M parameters, 200 languages) for on-device inference on Apple platforms. Max sequence length 1024 matches the official model max_position_embeddings.
This CoreML model is derived from facebook/nllb-200-distilled-600M (NLLB-200 distilled 600M on the Hub). Key official specs:
| Item | Official value |
|---|---|
| Model type | M2M100 (encoder-decoder) |
| Parameters | ~600M (distilled) |
| Languages | 200 (Flores-200 coverage) |
| Max position embeddings | 1024 (config.json) |
| Vocab size | 256,206 |
| License | CC-BY-NC-4.0 |
From the base model card: research and non-commercial use; single-sentence translation; not for production, medical, or legal domain; training used input lengths not exceeding 512 tokens (longer sequences may degrade). Not for certified translation. For full intended use, limitations, metrics (BLEU, spBLEU, chrF++), and ethical considerations, see the official model page.
Note: Other NLLB variants (e.g. larger or MoE) may support longer context (e.g. 128k); this CoreML conversion follows the distilled 600M config, which has max 1024.
max_position_embeddings| File / folder | Description |
|---|---|
NLLB_Encoder_1024.mlpackage | Encoder (source encoding) |
NLLB_Decoder_1024_init.mlpackage | Decoder first step (with encoder outputs) |
NLLB_Decoder_1024_step.mlpackage | Decoder subsequent steps (with past KV cache) |
tokenizer/ | SentencePiece tokenizer (same as base model) |
config.json | Model config |
eng_Latn, jpn_Jpan).For a smaller, 8-bit quantized variant (same 1024 length), see aoiandroid/nllb200-coreml-1024-palettized.
Produced with the notebook [nllb200_coreml_colab_v3_fixed.ipynb](local iOS project (private)) (CoreML encoder + KV decoder trace and export). Alignment: 1024 matches the official config.json of the base model.
If you use this model, please cite the NLLB-200 paper and the base model:
@article{nllb2022,
title={No Language Left Behind: Scaling Human-Centered Machine Translation},
author={{NLLB Team} and others},
journal={arXiv preprint arXiv:2207.04672},
year={2022}
}