Downloads · 30 days
25
1% of all-time downloads
staka/takomt
takomt is a translation model from staka. Use it when you need text moved from one language to another. It is set up for transformers. The card lists the license as cc-by-sa-4.0.
This is a translation model using Marian-NMT. For more details, please see my repository.
Downloads · 30 days
25
1% of all-time downloads
All-time downloads
2.1K
Public
Repo size
1.2 GB
Likes
0
Public
Click a slice to open those files.
.bin187 MB · 96%
From the Hugging Face model README
This is a translation model using Marian-NMT. For more details, please see my repository.
In addition to the data listed in the repository I also used ParaCrawl.
This model uses transformers and sentencepiece.
!pip install transformers sentencepiece
You can use this model directly with a pipeline:
from transformers import pipeline
tako_translator = pipeline('translation', model='staka/takomt')
tako_translator('This is a cat.')
The results of the evaluation using tatoeba(randomly selected 500 sentences) are as follows:
| source | target | BLEU(*1) |
|---|---|---|
| de | ja | 27.8 |
| en | ja | 28.4 |
| es | ja | 32.0 |
| fr | ja | 27.9 |
| it | ja | 24.3 |
| ru | ja | 27.3 |
| uk | ja | 29.8 |
(*1) sacrebleu --tokenize ja-mecab