Downloads · 30 days
31
1% of all-time downloads
raincandy-u/Coder1.8-ORPO-TEST
Coder1.8-ORPO-TEST is a text generation model from raincandy-u. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
Test model for ORPO finetune method, trained on ~20k code examples for 1 epoch on 2 x A40 cards with 4-bit QLora (lora rank=lora alpha=16).
Downloads · 30 days
31
1% of all-time downloads
All-time downloads
4.9K
Public
Parameters
1.8B
3.7 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors3.7 GB · 100%
From the Hugging Face model README
Test model for ORPO finetune method, trained on ~20k code examples for 1 epoch on 2 x A40 cards with 4-bit QLora (lora rank=lora alpha=16).
This is a test model and may generate incorrect responses. Use at your own risk.
Limited training data and quantization may impact performance.
Have questions or feedback? Join our Discord server Here.
Detailed results can be found here
| Metric | Value |
|---|---|
| Avg. | 45.76 |
| AI2 Reasoning Challenge (25-Shot) | 38.82 |
| HellaSwag (10-Shot) | 60.48 |
| MMLU (5-Shot) | 46.70 |
| TruthfulQA (0-shot) | 41.38 |
| Winogrande (5-shot) | 59.75 |
| GSM8k (5-shot) | 27.45 |