Downloads · 30 days
11
11% of all-time downloads
CompassioninMachineLearning/Basellama_plus3kv3new_plus20kfinetune_plus1kGRPO
Basellama_plus3kv3new_plus20kfinetune_plus1kGRPO is a text generation model from CompassioninMachineLearning. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
- Developed by: CompassioninMachineLearning - License: apache-2.0 - Finetuned from model : CompassioninMachineLearning/Basellamaplus3kv3plus20kfinetune2epochs
Downloads · 30 days
11
11% of all-time downloads
All-time downloads
98
Public
Parameters
8B
16.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.1 GB · 100%
From the Hugging Face model README
This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.