Downloads · 30 days
13
10% of all-time downloads
auhide/punctual-bert-bg
punctual-bert-bg is a token classification model from auhide. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as mit.
Visit the website - Zapetayko, to test out the model.
Downloads · 30 days
13
10% of all-time downloads
All-time downloads
129
Public
Parameters
177M
1.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors709 MB · 98%
From the Hugging Face model README
Visit the website - Zapetayko, to test out the model.
from transformers import pipeline
MODEL_ID = "auhide/punctual-bert-bg"
punctuate = pipeline("token-classification", model=MODEL_ID, tokenizer=MODEL_ID)
punctuate("Човекът искащ безгрижно писане ме помоли да създам този модел.")
[{'entity': 'B-CMA',
'score': 0.95041466,
'index': 1,
'word': '▁Човекът',
'start': 0,
'end': 7},
{'entity': 'I-CMA',
'score': 0.95229745,
'index': 2,
'word': '▁иска',
'start': 7,
'end': 12},
{'entity': 'B-CMA',
'score': 0.95945585,
'index': 5,
'word': '▁писане',
'start': 23,
'end': 30},
{'entity': 'I-CMA',
'score': 0.90768945,
'index': 6,
'word': '▁ме',
'start': 30,
'end': 33}]
Basically, B-CMA tags the token that's before the comma, and I-CMA tags the token after the comma.
Therefore, if we place the commas based on these tags, the result is:
"Човекът, искащ безгрижно писане, ме помоли да създам този модел."