Downloads · 30 days
15
2% of all-time downloads
SzegedAI/charmen-electra
charmen-electra is a feature extraction model from SzegedAI. Use it when you need embeddings to search or compare text. It is set up for transformers. The card lists the license as apache-2.0.
A byte-based transformer model trained on Hungarian language. In order to use the model you will need a custom Tokenizer which is available at: https://github.com/szegedai/byte-offset-tokenizer.
Downloads · 30 days
15
2% of all-time downloads
All-time downloads
631
Public
Repo size
2.4 GB
Likes
1
Public
Click a slice to open those files.
.bin174 MB · 100%
From the Hugging Face model README
A byte-based transformer model trained on Hungarian language. In order to use the model you will need a custom Tokenizer which is available at: https://github.com/szegedai/byte-offset-tokenizer.
Since we use a custom architecture with Gradient Boosting, Down- and Up-Sampling, you have to enable Trusted Remote Code like:
model = AutoModel.from_pretrained("SzegedAI/charmen-electra", trust_remote_code=True)