Downloads · 30 days
18
100% of all-time downloads
Yuki131/KaLM-Reranker-V1-Nano-R2-Stage1
KaLM-Reranker-V1-Nano-R2-Stage1 is a machine learning model from Yuki131. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
We release the checkpoints from the three-stage training pipeline described in the third version of our paper. Stage 1 uses supervised fine-tuning; Stage 2 produces two checkpoints through soft-label distillation; and…
Downloads · 30 days
18
100% of all-time downloads
All-time downloads
18
Public
Parameters
786M
1.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors1.6 GB · 98%
From the Hugging Face model README
We release the checkpoints from the three-stage training pipeline described in the third version of our paper. Stage 1 uses supervised fine-tuning; Stage 2 produces two checkpoints through soft-label distillation; and Stage 3 combines them through model soup to produce the final R2 models.
| Stage | Checkpoint | Nano | Small | Large |
|---|---|---|---|---|
| Stage 1 | Supervised fine-tuning | KaLM-Reranker-V1-Nano-R2-Stage1 | KaLM-Reranker-V1-Small-R2-Stage1 | KaLM-Reranker-V1-Large-R2-Stage1 |
| Stage 2 | Distillation (r64-a32) | KaLM-Reranker-V1-Nano-R2-Stage2-r64-a32 | KaLM-Reranker-V1-Small-R2-Stage2-r64-a32 | KaLM-Reranker-V1-Large-R2-Stage2-r64-a32 |
| Stage 2 | Distillation (r96-a48) | KaLM-Reranker-V1-Nano-R2-Stage2-r96-a48 | KaLM-Reranker-V1-Small-R2-Stage2-r96-a48 | KaLM-Reranker-V1-Large-R2-Stage2-r96-a48 |
| Stage 3 | Final R2 model | KaLM-Reranker-V1-Nano-R2 | KaLM-Reranker-V1-Small-R2 | KaLM-Reranker-V1-Large-R2 |