Downloads · 30 days
0
calix1/rewardmodel2
rewardmodel2 is a machine learning model from calix1. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as llama3.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
0
Access
Public
Updated Jul 9, 2024
Repo size
27.3 MB
Likes
0
Public
Click a slice to open those files.
.safetensors27.3 MB · 75%
From the Hugging Face model README
This model is a fine-tuned version of meta-llama/Meta-Llama-3-8B-Instruct on the None dataset.
More information needed
More information needed
More information needed
The following hyperparameters were used during training: