Downloads · 30 days
12
41% of all-time downloads
syssec-utd/py314-pylingual-v8-segmenter
py314-pylingual-v8-segmenter is a token classification model from syssec-utd. Use it when you need labels on individual words, such as names. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
12
41% of all-time downloads
All-time downloads
29
Public
Parameters
109M
436 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors436 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of syssec-utd/py314-pylingual-v8-mlm on the syssec-utd/segmentation-py314-pylingual-v8-tokenized dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Precision | Recall | F1 | Accuracy |
|---|---|---|---|---|---|---|---|
| 0.0313 | 1.0 | 20485 | 0.0193 | 0.9926 | 0.9924 | 0.9925 | 0.9983 |
| 0.0167 | 2.0 | 40970 | 0.0156 | 0.9928 | 0.9936 | 0.9932 | 0.9985 |