Downloads · 30 days
97
5% of all-time downloads
VMware/vinilm-2021-from-large
vinilm-2021-from-large is a feature extraction model from VMware. Use it when you need embeddings to search or compare text. It is set up for transformers. The card lists the license as apache-2.0.
<ul <li Authors: R&D AI Lab, VMware Inc. <li Model date: Jun 2022 <li Model version: 2021-distilled-from-large <li Model type: Pretrained language model <li License: Apache 2.0 </ul
Downloads · 30 days
97
5% of all-time downloads
All-time downloads
1.8K
Public
Parameters
67M
536 MB on disk
Likes
2
Public
Click a slice to open those files.
.bin268 MB · 50%
How the weights are stored.
F3267M · 100%
From the Hugging Face model README
Based on MiniLMv2 distillation, we have distilled vBERT-2021-large into a smaller minilmv2 model for faster inference times without a significant loss of performance.
The model functions as a VMware-specific Language Model.
Here is how to use this model to get the features of a given text in PyTorch:
from transformers import BertTokenizer, BertModel
tokenizer = BertTokenizer.from_pretrained('VMware/vinilm-2021-from-large')
model = BertModel.from_pretrained("VMware/vinilm-2021-from-large")
text = "Replace me by any text you'd like."
encoded_input = tokenizer(text, return_tensors='pt')
output = model(**encoded_input)
and in TensorFlow:
from transformers import BertTokenizer, TFBertModel
tokenizer = BertTokenizer.from_pretrained('VMware/vinilm-2021-from-large')
model = TFBertModel.from_pretrained('VMware/vinilm-2021-from-large')
text = "Replace me by any text you'd like."
encoded_input = tokenizer(text, return_tensors='tf')
output = model(encoded_input)
The model is distilled from vBERT-2021-large</li> <br> Weights were initialized using nreimers/MiniLMv2-L6-H768-distilled-from-BERT-Large</li>
Publically available VMware text data such as VMware Docs, Blogs, etc. were used for distilling the teacher vBERT-2021-large model into vinilm-2021-from-large model. Sourced in May 2021. (~320,000 Documents)
We benchmarked vinilm on various VMware-specific NLP downstream tasks (IR, classification, etc).
Since the model is distilled from a vBERT model based on the BERT model, it may have the same biases embedded within the original BERT model.
The data needs to be preprocessed using our internal vNLP Preprocessor (not available to the public) to maximize its performance.