Downloads · 30 days
7
0% of all-time downloads
yeongjoon/Kconvo-roberta
Kconvo-roberta is a fill-mask model from yeongjoon. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as mit.
- There are many PLMs (Pretrained Language Models) for Korean, but most of them are trained with written language. - Here, we introduce a retrained PLM for prediction of Korean conversation data where we use verbal da…
Downloads · 30 days
7
0% of all-time downloads
All-time downloads
3.7K
Public
Repo size
885 MB
Likes
0
Public
Click a slice to open those files.
.bin443 MB · 100%
From the Hugging Face model README
# Kconvo-roberta
from transformers import RobertaTokenizerFast, RobertaModel
tokenizer_roberta = RobertaTokenizerFast.from_pretrained("yeongjoon/Kconvo-roberta")
model_roberta = RobertaModel.from_pretrained("yeongjoon/Kconvo-roberta")
- National Institute of the Korean Language
* 온라인 대화 말뭉치 2021
* 일상 대화 말뭉치 2020
* 구어 말뭉치
* 메신저 말뭉치
- AI-Hub
* 온라인 구어체 말뭉치 데이터
* 상담 음성
* 한국어 음성
* 자유대화 음성(일반남여)
* 일상생활 및 구어체 한-영 번역 병렬 말뭉치 데이터
* 한국인 대화음성
* 감성 대화 말뭉치
* 주제별 텍스트 일상 대화 데이터
* 용도별 목적대화 데이터
* 한국어 SNS