Downloads · 30 days
67
34% of all-time downloads
alphaedge-ai/multilingual-e5-small-eus-32768
multilingual-e5-small-eus-32768 is a sentence similarity model from alphaedge-ai. Use it when you need a score for how close two texts are. It is set up for sentence-transformers. The card lists the license as mit.
This model is a 70.91% smaller version of intfloat/multilingual-e5-small optimized for Basque language via vocabulary size reduction using the trimming method. This trimmed model should perform similarly to the origin…
Downloads · 30 days
67
34% of all-time downloads
All-time downloads
196
Public
Parameters
34.2M
137 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors137 MB · 98%
From the Hugging Face model README
This model is a 70.91% smaller version of intfloat/multilingual-e5-small optimized for Basque language via vocabulary size reduction using the trimming method.
This trimmed model should perform similarly to the original model with only 32,768 tokens and a much smaller memory footprint. However, it may not perform well for other languages as tokens not commonly used in the selected languages were removed from the vocabulary.
| Metric | Original | Trimmed | Reduction |
|---|---|---|---|
| Vocabulary size | 250,037 tokens | 32,768 tokens | 86.89% |
| Model size | 117,653,760 params | 34,222,464 params | 70.91% |

from sentence_transformers import SentenceTransformer
# Download from the 🤗 Hub
model = SentenceTransformer("alphaedge-ai/multilingual-e5-small-eus-32768")
# Run inference with queries and documents
query = "My query in Basque"
documents = [
"Chunk in Basque",
"Chunk in Basque",
"Chunk in Basque",
]
query_embeddings = model.encode_query(query)
document_embeddings = model.encode_document(documents)
print(query_embeddings.shape, document_embeddings.shape)
# Compute similarities to determine a ranking
similarities = model.similarity(query_embeddings, document_embeddings)
print(similarities)
@article{wang2024multilingual,
title={Multilingual E5 Text Embeddings: A Technical Report},
author={Wang, Liang and Yang, Nan and Huang, Xiaolong and Yang, Linjun and Majumder, Rangan and Wei, Furu},
journal={arXiv preprint arXiv:2402.05672},
year={2024}
}
@misc{hf_blogpost_trimming,
title={Introduction to Trimming},
author={Loïck BOURDOIS and Tom AARSEN and Bram VANROY and Christopher AKIKI and Woojun JUNG and Manuel ROMERO and Prithiv SAKTHI},
year={2026},
url={https://huggingface.co/blog/lbourdois/introduction-to-trimming},
}