Downloads · 30 days
64
0% of all-time downloads
NeuralNovel/Gecko-7B-v0.1
Gecko-7B-v0.1 is a text generation model from NeuralNovel. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Downloads · 30 days
64
0% of all-time downloads
All-time downloads
23.3K
Public
Parameters
7.2B
24.6 GB on disk
Likes
6
Public
Click a slice to open those files.
.safetensors14.5 GB · 59%
From the Hugging Face model README

Designed to generate instructive and narrative text, with a focus on mathematics & numeracy.
Full-parameter fine-tune (FFT) of Mistral-7B-Instruct-v0.2, with apache-2.0 license.
You may download and use this model for research, training and commercial purposes.
<a href='https://ko-fi.com/S6S2UH2TC' target='_blank'><img height='38' style='border:0px;height:36px;' src='https://storage.ko-fi.com/cdn/kofi1.png?v=3' border='0' alt='Buy Me a Coffee at ko-fi.com' /></a> <a href='https://discord.gg/KFS229xD' target='_blank'><img width='140' height='500' style='border:0px;height:36px;' src='https://i.ibb.co/tqwznYM/Discord-button.png' border='0' alt='Join Our Discord!' /></a>
The model was finetuned using the Neural-Mini-Math dataset (Currently Private)
Fine-tuned with the intention of following all prompt directions, making it more suitable for roleplay and problem solving.
The model may not perform well in scenarios unrelated to instructive and narrative text generation. Misuse or applications outside its designed scope may result in suboptimal outcomes.
This model may not work as intended. As such all users are encouraged to use this model with caution and respect.
This model is for testing and research purposes only, it has reduced levels of alignment and as a result may produce NSFW or harmful content. The user is responsible for their output and must use this model responsibly.
n_epochs = 3,
n_checkpoints = 3,
batch_size = 12,
learning_rate = 1e-5,
Sincere appreciation to Techmind for their generous sponsorship.
Detailed results can be found here
| Metric | Value |
|---|---|
| Avg. | 64.58 |
| AI2 Reasoning Challenge (25-Shot) | 61.35 |
| HellaSwag (10-Shot) | 83.36 |
| MMLU (5-Shot) | 61.05 |
| TruthfulQA (0-shot) | 62.60 |
| Winogrande (5-shot) | 77.58 |
| GSM8k (5-shot) | 41.55 |