Downloads · 30 days
50
7% of all-time downloads
redis/langcache-reranker-v1-mnrl
langcache-reranker-v1-mnrl is a text ranking model from redis. Use it for the text ranking task on the model card, and read the license before you ship it in a product. It is set up for sentence-transformers. The card lists the license as apache-2.0.
This is a Cross Encoder model finetuned from Alibaba-NLP/gte-reranker-modernbert-base on the [LangCache Sentence Pairs (subsets=['all'], train+val=True)](https://huggingface.co/datasets/aditeyabaral-redis/langcache-se…
Downloads · 30 days
50
7% of all-time downloads
All-time downloads
718
Public
Parameters
150M
11.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors299 MB · 99%
From the Hugging Face model README
This is a Cross Encoder model finetuned from Alibaba-NLP/gte-reranker-modernbert-base on the LangCache Sentence Pairs (subsets=['all'], train+val=True) dataset using the sentence-transformers library. It computes scores for pairs of texts, which can be used for sentence pair classification.
First install the Sentence Transformers library:
pip install -U sentence-transformers
Then you can load this model and run inference.
from sentence_transformers import CrossEncoder
# Download from the 🤗 Hub
model = CrossEncoder("redis/langcache-reranker-v1-softmnrl-triplet")
# Get scores for pairs of texts
pairs = [
[' What high potential jobs are there other than computer science?', ' What high potential jobs are there other than computer science?'],
[' Would India ever be able to develop a missile system like S300 or S400 missile?', ' Would India ever be able to develop a missile system like S300 or S400 missile?'],
[' water from the faucet is being drunk by a yellow dog', 'A yellow dog is drinking water from the faucet'],
[' water from the faucet is being drunk by a yellow dog', 'The yellow dog is drinking water from a bottle'],
['! colspan = `` 14 `` `` Players who appeared for Colchester who left during the season ``', '! colspan = `` 14 `` `` Players who appeared for Colchester who left during the season ``'],
]
scores = model.predict(pairs)
print(scores.shape)
# (5,)
# Or rank different texts based on similarity to a single text
ranks = model.rank(
' What high potential jobs are there other than computer science?',
[
' What high potential jobs are there other than computer science?',
' Would India ever be able to develop a missile system like S300 or S400 missile?',
'A yellow dog is drinking water from the faucet',
'The yellow dog is drinking water from a bottle',
'! colspan = `` 14 `` `` Players who appeared for Colchester who left during the season ``',
]
)
# [{'corpus_id': ..., 'score': ...}, {'corpus_id': ..., 'score': ...}, ...]
<!--
### Direct Usage (Transformers)
<details><summary>Click to see the direct usage in Transformers</summary>
</details>
-->
<!--
### Downstream Usage (Sentence Transformers)
You can finetune this model on your own dataset.
<details><summary>Click to expand</summary>
</details>
-->
<!--
### Out-of-Scope Use
*List how the model may foreseeably be misused and address what users ought not to do with the model.*
-->
<!--
## Bias, Risks and Limitations
*What are the known or foreseeable issues stemming from this model? You could also flag here known failure cases or weaknesses of the model.*
-->
<!--
### Recommendations
*What are recommendations with respect to the foreseeable issues? For example, filtering explicit content.*
-->
| anchor | positive | negative_1 | |
|---|---|---|---|
| type | string | string | string |
| details | <ul><li>min: 24 characters</li><li>mean: 114.25 characters</li><li>max: 268 characters</li></ul> | <ul><li>min: 19 characters</li><li>mean: 114.1 characters</li><li>max: 226 characters</li></ul> | <ul><li>min: 4 characters</li><li>mean: 93.04 characters</li><li>max: 234 characters</li></ul> |
| anchor | positive | negative_1 |
|---|---|---|
| <code> Any Canadian teachers (B.Ed. holders) teaching in U.S. schools?</code> | <code> Any Canadian teachers (B.Ed. holders) teaching in U.S. schools?</code> | <code>Are there many Canadians living and working illegally in the United States?</code> |
| <code> Are there any underlying psychological tricks/tactics that are used when designing the lines for rides at amusement parks?</code> | <code> Are there any underlying psychological tricks/tactics that are used when designing the lines for rides at amusement parks?</code> | <code>Is there any tricks for straight lines mcqs?</code> |
| <code> Can I pay with a debit card on PayPal?</code> | <code> Can I pay with a debit card on PayPal?</code> | <code>Can you transfer PayPal funds onto a debit card/credit card?</code> |
{
"scale": 20.0,
"num_negatives": 1,
"activation_fn": "torch.nn.modules.activation.Sigmoid"
}
| anchor | positive | negative_1 | |
|---|---|---|---|
| type | string | string | string |
| details | <ul><li>min: 3 characters</li><li>mean: 97.95 characters</li><li>max: 314 characters</li></ul> | <ul><li>min: 3 characters</li><li>mean: 97.03 characters</li><li>max: 314 characters</li></ul> | <ul><li>min: 11 characters</li><li>mean: 74.49 characters</li><li>max: 295 characters</li></ul> |
| anchor | positive | negative_1 |
|---|---|---|
| <code> What high potential jobs are there other than computer science?</code> | <code> What high potential jobs are there other than computer science?</code> | <code>Why IT or Computer Science jobs are being over rated than other Engineering jobs?</code> |
| <code> Would India ever be able to develop a missile system like S300 or S400 missile?</code> | <code> Would India ever be able to develop a missile system like S300 or S400 missile?</code> | <code>Should India buy the Russian S400 air defence missile system?</code> |
| <code> water from the faucet is being drunk by a yellow dog</code> | <code>A yellow dog is drinking water from the faucet</code> | <code>Do you get more homework in 9th grade than 8th?</code> |
{
"scale": 20.0,
"num_negatives": 1,
"activation_fn": "torch.nn.modules.activation.Sigmoid"
}
@inproceedings{reimers-2019-sentence-bert,
title = "Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks",
author = "Reimers, Nils and Gurevych, Iryna",
booktitle = "Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing",
month = "11",
year = "2019",
publisher = "Association for Computational Linguistics",
url = "https://arxiv.org/abs/1908.10084",
}
<!--
## Glossary
*Clearly define terms in order to be accessible across audiences.*
-->
<!--
## Model Card Authors
*Lists the people who create the model card, providing recognition and accountability for the detailed work that goes into its construction.*
-->
<!--
## Model Card Contact
*Provides a way for people who have updates to the Model Card, suggestions, or questions, to contact the Model Card authors.*
-->