Downloads · 30 days
1.9K
10% of all-time downloads
gety-ai/gety-embed-v0
gety-embed-v0 is a sentence similarity model from gety-ai. Use it when you need a score for how close two texts are. It is set up for sentence-transformers. The card lists the license as mit.
Fine-tuned from intfloat/multilingual-e5-small using open-source and proprietary synthetic data, optimized for local search scenarios.
Downloads · 30 days
1.9K
10% of all-time downloads
All-time downloads
19.3K
Public
Repo size
724 MB
Likes
3
Public
Click a slice to open those files.
.bin354 MB · 49%
From the Hugging Face model README
Fine-tuned from intfloat/multilingual-e5-small using open-source and proprietary synthetic data, optimized for local search scenarios.
| File | Quantization | Size |
|---|---|---|
onnx/model_uint8.onnx | UINT8 Dynamic | 112 MB |
| File | Quantization | Size |
|---|---|---|
onnx_cuda/model_fp16.onnx | FP16 | 224 MB |
Inputs input_ids + attention_mask; masked mean pooling + L2 normalization
baked into the graph, output is the normalized 384-dim embedding.
| File | Quantization | Size |
|---|---|---|
coreml/model_fp16.mlpackage | FP16 | 224 MB |
Single input_ids input (dynamic 1-512 tokens, no padding), pooling and
normalization baked in. macOS 13+, ANE accelerated.
| File | Quantization | Size |
|---|---|---|
openvino/model.xml + openvino/model.bin | INT8 weight-only (NNCF), FP16 activations | 114 MB |
Single-input OpenVINO IR for static-shape NPU compilation
(reshape([1, L]) + compile, batch=1, sequence buckets e.g. 32/64/128/256/512):
input_ids [batch, seq] (int64), padded to the bucket length with
<pad> (token id 1 — not config.json's pad_token_id=0, which is <s>)input_ids != 1embeddings [batch, 384] (float32)Selected over FP16 / static-PTQ INT8 / INT4 on NPU hardware (Intel AI Boost, OpenVINO 2026.2.1): equal latency to all other variants (NPU latency is shape-bound), best accuracy (min cosine 0.9998 vs PyTorch reference; static PTQ drops to 0.9872 on short queries), and half the size of FP16.