Downloads · 30 days
23
0% of all-time downloads
gate369/Nexim-7b
Nexim-7b is a text generation model from gate369. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Nexim-7b is a merge of the following models using mergekit: liminerity/m3 liminerity/M7-7b
Downloads · 30 days
23
0% of all-time downloads
All-time downloads
4.9K
Public
Parameters
7.2B
14.5 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors14.5 GB · 100%
From the Hugging Face model README
Nexim-7b is a merge of the following models using mergekit:
slices:
- sources:
- model: liminerity/m3
layer_range: [0, 32]
- model: liminerity/M7-7b
layer_range: [0, 32]
merge_method: slerp
base_model: liminerity/m3
parameters:
t:
- filter: self_attn
value: [0, 0.5, 0.3, 0.7, 1]
- filter: mlp
value: [1, 0.5, 0.7, 0.3, 0]
- value: 0.5
dtype: bfloat16
Detailed results can be found here
| Metric | Value |
|---|---|
| Avg. | 76.53 |
| AI2 Reasoning Challenge (25-Shot) | 73.04 |
| HellaSwag (10-Shot) | 89.10 |
| MMLU (5-Shot) | 64.48 |
| TruthfulQA (0-shot) | 77.68 |
| Winogrande (5-shot) | 84.77 |
| GSM8k (5-shot) | 70.13 |