Downloads · 30 days
85
3% of all-time downloads
NoYo25/BiodivBERT
BiodivBERT is a token classification model from NoYo25. Use it when you need labels on individual words, such as names. It is set up for transformers. The card lists the license as apache-2.0.
BiodivBERT is a domain-specific BERT based cased model for the biodiversity literature. It uses the tokenizer from BERTT base cased model. BiodivBERT is pre-trained on abstracts and full text from biodiversity literat…
Downloads · 30 days
85
3% of all-time downloads
All-time downloads
2.8K
Public
Repo size
867 MB
Likes
5
Public
Click a slice to open those files.
.bin433 MB · 100%
From the Hugging Face model README
>>> from transformers import AutoTokenizer, AutoModelForMaskedLM
>>> tokenizer = AutoTokenizer.from_pretrained("NoYo25/BiodivBERT")
>>> model = AutoModelForMaskedLM.from_pretrained("NoYo25/BiodivBERT")
>>> from transformers import AutoTokenizer, AutoModelForTokenClassification
>>> tokenizer = AutoTokenizer.from_pretrained("NoYo25/BiodivBERT")
>>> model = AutoModelForTokenClassification.from_pretrained("NoYo25/BiodivBERT")
>>> from transformers import AutoTokenizer, AutoModelForSequenceClassification
>>> tokenizer = AutoTokenizer.from_pretrained("NoYo25/BiodivBERT")
>>> model = AutoModelForSequenceClassification.from_pretrained("NoYo25/BiodivBERT")
BiodivBERT overperformed both BERT_base_cased, biobert_v1.1, and BiLSTM as a baseline approach on the down stream tasks.