Downloads · 30 days
15
14% of all-time downloads
TylerG01/Indigo-v0.1
Indigo-v0.1 is a text generation model from TylerG01. Use it when you need the model to write or continue text. The card lists the license as mit.
Refer to the original model card for more details on the model. This is v0.1 (alpha) release of the Indigo LLM project, which used LoRA Fine-Tuning to train Mistral 7B on more than 400 books, pamphlets, training docum…
Downloads · 30 days
15
14% of all-time downloads
All-time downloads
104
Public
Parameters
1.1B
4.1 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors4.1 GB · 100%
How the weights are stored.
U32905M · 80%
From the Hugging Face model README
Refer to the original model card for more details on the model.
This is v0.1 (alpha) release of the Indigo LLM project, which used LoRA Fine-Tuning to train Mistral 7B on more than 400 books, pamphlets, training documents, code snippets and other works in the cyber security field, openly sourced on the surface web. This version used 16 LoRA layers and had a val loss of 1.601 after the 4th training epoch. However, my goal for the LoRA version of this model is to produce a val loss of <1.51 after some modification to the dataset and training approach.
For more information on this project, check out the blog post at https://t2-security.com/indigo-llm-503cd6e22fe4.