Skip to content

nikitast

lang-segmentation-roberta

nikitast/lang-segmentation-roberta

lang-segmentation-roberta is a token classification model from nikitast. Use it when you need labels on individual words, such as names. It is set up for transformers.

RoBERTa fine-tuned on small parts of Open Subtitles, Oscar and Tatoeba datasets (~9k samples per language).

Downloads · 30 days

16

1% of all-time downloads

All-time downloads

1.2K

Public

Repo size

2.2 GB

Likes

4

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.bin1.1 GB · 98%

At a glance

Task
Token Classification
Library
transformers
Model type
xlm-roberta
Access
Public
Created
May 12, 2022
Updated
Jan 7, 2023
SHA
b1aa48fa

Try a prompt

Task
Token Classification
Library
transformers
Type
xlm-roberta
Languages
ru, uk, be, kk, az, hy
Created
May 12, 2022
Updated
Jan 7, 2023
lang-segmentation-roberta — AI Model — AIMarketly