Downloads · 30 days
335
0% of all-time downloads
clue/roberta_chinese_base
roberta_chinese_base is a machine learning model from clue. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
Language model: roberta-base Model size: 392M Language: Chinese Training data: CLUECorpusSmall Eval data: CLUE dataset
Downloads · 30 days
335
0% of all-time downloads
All-time downloads
72.5K
Public
Repo size
1.2 GB
Likes
8
Public
Click a slice to open those files.
.bin412 MB · 50%
From the Hugging Face model README
Language model: roberta-base Model size: 392M Language: Chinese Training data: CLUECorpusSmall Eval data: CLUE dataset
For results on downstream tasks like text classification, please refer to this repository.
NOTE: You have to call BertTokenizer instead of RobertaTokenizer !!!
import torch
from transformers import BertTokenizer, BertModel
tokenizer = BertTokenizer.from_pretrained("clue/roberta_chinese_base")
roberta = BertModel.from_pretrained("clue/roberta_chinese_base")
Organization of Language Understanding Evaluation benchmark for Chinese: tasks & datasets, baselines, pre-trained Chinese models, corpus and leaderboard.
Github: https://github.com/CLUEbenchmark Website: https://www.cluebenchmarks.com/