Downloads · 30 days
213
0% of all-time downloads
iampanda/zpoint_large_embedding_zh
zpoint_large_embedding_zh is a machine learning model from iampanda. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for sentence-transformers. The card lists the license as mit.
<h2 align="left"ZPoint Large Embedding for Chinese</h2
Downloads · 30 days
213
0% of all-time downloads
All-time downloads
822K
Public
Parameters
326M
4.6 GB on disk
Likes
48
Public
Click a slice to open those files.
.bin1.3 GB · 50%
From the Hugging Face model README
Base Model
Training Data
We constructed a dataset of approximately 100 million training samples through collection, machine translation, and LLM synthesis. This dataset includes data from various fields such as healthcare, law, electricity, automotive, and 3C (Consumer Electronics).
Training loss
from sentence_transformers import SentenceTransformer
sentences1 = ["这个产品真垃圾"]
sentences2 = ["我太喜欢这个产品了"]
model = SentenceTransformer('iampanda/zpoint_large_embedding_zh')
embeddings_1 = model.encode(sentences1, normalize_embeddings=True)
embeddings_2 = model.encode(sentences2, normalize_embeddings=True)
similarity = embeddings_1 @ embeddings_2.T
print(similarity)