Downloads · 30 days
54
2% of all-time downloads
AshtonIsNotHere/CodeLlama_7B_nlp_pp
CodeLlama_7B_nlp_pp is a text generation model from AshtonIsNotHere. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama2.
This model is a fine-tuned version of codellama/CodeLlama-7b-hf on the AshtonIsNotHere/nlpppcodedataset dataset. It achieves the following results on the evaluation set: - Loss: 0.4129 - Accuracy: 0.8968
Downloads · 30 days
54
2% of all-time downloads
All-time downloads
3K
Public
Repo size
61.1 GB
Likes
0
Public
Click a slice to open those files.
.bin13.5 GB · 65%
From the Hugging Face model README
This model is a fine-tuned version of codellama/CodeLlama-7b-hf on the AshtonIsNotHere/nlp_pp_code_dataset dataset. It achieves the following results on the evaluation set:
This model has been fine-tuned for code completion on a dataset of NLP++ code.
More information needed
Dataset consists of a combination of scraped NLP++ code and NLP++ code examples from the VisualText website.
This model is trained in a multinode, multi-gpu setup with DeepSpeed Z3. For more information on the training setup, check out the GitHub repo.
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Accuracy |
|---|---|---|---|---|
| No log | 1.0 | 61 | 0.5100 | 0.8726 |
| No log | 1.99 | 122 | 0.4129 | 0.8968 |
| No log | 2.99 | 183 | 0.4166 | 0.9072 |
| No log | 4.0 | 245 | 0.4595 | 0.9090 |
| No log | 5.0 | 306 | 0.5181 | 0.9093 |
| No log | 5.99 | 367 | 0.5553 | 0.9090 |
| No log | 6.97 | 427 | 0.5603 | 0.9089 |