Downloads · 30 days
44
100% of all-time downloads
ItsnotAilabs/MESIE-768
MESIE-768 is a sentence similarity model from ItsnotAilabs. Use it when you need a score for how close two texts are. It is set up for sentence-transformers. The card lists the license as apache-2.0.
Downloads · 30 days
44
100% of all-time downloads
All-time downloads
44
Public
Parameters
67.9M
705 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors272 MB · 99%
From the Hugging Face model README
MESIE-768 is a specialized embedding model designed for the Sovereign Knowledge Studio. It is optimized for career and protocol document classification, semantic routing, and matching across multiple domains.
MESIE-768 is fine-tuned from BAAI/bge-base-en-v1.5 using contrastive learning with hard negative mining. It is designed to understand the nuances of career protocols, domain-aware classifications (business, engineering, cybersecurity, architecture), and identity-impact scoring. It calculates phi-weight correlations to effectively route protocol documents.
| Format | Precision | RAM Required | Latency (ms/query) |
|---|---|---|---|
| PyTorch | FP32 (Full) | ~440 MB | ~12.5 |
| PyTorch | FP16 | ~220 MB | ~7.2 |
| ONNX | INT8 | ~110 MB | ~4.5 |
| GGUF | Q4_K_M | ~65 MB | ~2.8 |
The model was fine-tuned on a proprietary corpus:
Fine-tuned with a contrastive learning objective tailored for semantic retrieval and protocol classification.
The model has been evaluated on standard and custom benchmarks:
| Benchmark | Metric | Score |
|---|---|---|
| MTEB Retrieval | Average | 63.2 |
| STS-B | Pearson/Spearman | 84.1 |
| Protocol-Route (Custom) | Accuracy | 91.7% |
| Peak Throughput | Docs/sec | ~38,134 |
Note: Protocol-Route Accuracy is an estimated benchmark based on internal MedinaMemorySystems evaluations.
For optimal retrieval results, particularly when matching query strings against a corpus of protocols, use the following prefix on queries:
Represent this sentence for searching relevant passages: {query}
No prefix is needed for the documents/passages being indexed.
from sentence_transformers import SentenceTransformer, util
# Load the model
model = SentenceTransformer('MedinaMemorySystems/mesie-768')
# Define protocol descriptions
protocols = [
"Protocol A: Defines the core architecture for scalable cloud deployments.",
"Protocol B: Outlines the cybersecurity guidelines for endpoint protection."
]
# Query
query = "Represent this sentence for searching relevant passages: Show me network security rules."
# Compute embeddings
query_embedding = model.encode(query)
doc_embeddings = model.encode(protocols)
# Compute cosine similarity
cosine_scores = util.cos_sim(query_embedding, doc_embeddings)
print(f"Similarity Scores: {cosine_scores}")
import torch
import torch.nn.functional as F
from transformers import AutoTokenizer, AutoModel
# Load model and tokenizer
tokenizer = AutoTokenizer.from_pretrained('MedinaMemorySystems/mesie-768')
model = AutoModel.from_pretrained('MedinaMemorySystems/mesie-768')
# Tokenize inputs
sentences = ["Represent this sentence for searching relevant passages: Query text"]
encoded_input = tokenizer(sentences, padding=True, truncation=True, return_tensors='pt')
# Compute token embeddings
with torch.no_grad():
model_output = model(**encoded_input)
# Perform pooling
sentence_embeddings = model_output[0][:, 0]
# Normalize embeddings
sentence_embeddings = F.normalize(sentence_embeddings, p=2, dim=1)
print(sentence_embeddings)
@misc{medinamemorysystems2026mesie,
title={MESIE-768: A Multi-Environment Sovereign Intelligence Engine for Career Protocol Routing},
author={MedinaMemorySystems},
year={2026},
publisher={Hugging Face}
}