Downloads · 30 days
10
36% of all-time downloads
syssec-utd/py314-pylingual-v2-segmenter
py314-pylingual-v2-segmenter is a token classification model from syssec-utd. Use it when you need labels on individual words, such as names. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
10
36% of all-time downloads
All-time downloads
28
Public
Parameters
108M
868 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors434 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of syssec-utd/py314-pylingual-v2-mlm on the syssec-utd/segmentation-py314-pylingual-v2-tokenized dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Precision | Recall | F1 | Accuracy |
|---|---|---|---|---|---|---|---|
| 0.0092 | 1.0 | 33414 | 0.0024 | 0.9948 | 0.9942 | 0.9945 | 0.9985 |
| 0.0052 | 2.0 | 66828 | 0.0022 | 0.9953 | 0.9960 | 0.9956 | 0.9987 |