Downloads · 30 days
19
0% of all-time downloads
clue/roberta_chinese_large
roberta_chinese_large is a machine learning model from clue. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
Language model: roberta-large Model size: 1.2G Language: Chinese Training data: CLUECorpusSmall Eval data: CLUE dataset
Downloads · 30 days
19
0% of all-time downloads
All-time downloads
8.4K
Public
Repo size
3.9 GB
Likes
2
Public
Click a slice to open those files.
.bin1.3 GB · 50%
From the Hugging Face model README
Language model: roberta-large Model size: 1.2G Language: Chinese Training data: CLUECorpusSmall Eval data: CLUE dataset
For results on downstream tasks like text classification, please refer to this repository.
NOTE: You have to call BertTokenizer instead of RobertaTokenizer !!!
import torch
from transformers import BertTokenizer, BertModel
tokenizer = BertTokenizer.from_pretrained("clue/roberta_chinese_large")
roberta = BertModel.from_pretrained("clue/roberta_chinese_large")
Organization of Language Understanding Evaluation benchmark for Chinese: tasks & datasets, baselines, pre-trained Chinese models, corpus and leaderboard.
Github: https://github.com/CLUEbenchmark Website: https://www.cluebenchmarks.com/