Downloads · 30 days
138
5% of all-time downloads
jinaai/xlm-roberta-flash-implementation
xlm-roberta-flash-implementation is a machine learning model from jinaai. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as cc-by-nc-4.0.
Downloads · 30 days
138
5% of all-time downloads
All-time downloads
2.6K
Public
Repo size
2.2 GB
Likes
34
Public
Click a slice to open those files.
.py183 KB · 98%
From the Hugging Face model README
Core implementation of Jina XLM-RoBERTa
This implementation is adapted from XLM-Roberta. In contrast to the original implementation, this model uses Rotary positional encodings and supports flash-attention 2.
Weights from an original XLMRoberta model can be converted using the convert_roberta_weights_to_flash.py script in the model repository.