Downloads · 30 days
49
37% of all-time downloads
cimo001/embeddinggemma-300m
embeddinggemma-300m is a sentence similarity model from cimo001. Use it when you need a score for how close two texts are. It is set up for onnx. The card lists the license as mit.
Downloads · 30 days
49
37% of all-time downloads
All-time downloads
131
Public
Repo size
1.5 GB
Likes
1
Public
Click a slice to open those files.
.onnx_data1.5 GB · 100%
From the Hugging Face model README
FP32 = model.onnx, model.onnx_data
INT8 = model_quantized.onnx, model_quantized.onnx_data
# onnxruntime-gpu for run it on GPU
pip install onnxruntime sentencepiece numpy
python3 src/example.py
0.600231 | The giant panda (Ailuropoda melanoleuca), sometimes called a panda bear, is a bear species endemic to China.
0.153004 | hi
0.554038 | パンダはクマ科の哺乳類で、中国の固有種である。
Note:
tokenizer.model directly with the native ids<br>
(bos_id, eos_id, pad_id from the sentencepiece model).title: none | text: {text}<br>
query = task: search result | query: {text}[bos] text [eos], padded per batch to the longest sequence.sentence_embedding, 768 dimensions.