Downloads · 30 days
19
1% of all-time downloads
kechengcode/Llama-3.1-5B-Instruct-18Layers
Llama-3.1-5B-Instruct-18Layers is a text generation model from kechengcode. Use it when you need the model to write or continue text. It is set up for transformers.
This is a merge of pre-trained language models created using mergekit.
Downloads · 30 days
19
1% of all-time downloads
All-time downloads
1.5K
Public
Parameters
5B
10 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors10 GB · 100%
From the Hugging Face model README
This is a merge of pre-trained language models created using mergekit.
This model was merged using the passthrough merge method.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
dtype: float16
merge_method: passthrough
slices:
- sources:
- layer_range: [0, 6]
model: meta-llama/Meta-Llama-3.1-8B-Instruct
- sources:
- layer_range: [20, 32]
model: meta-llama/Meta-Llama-3.1-8B-Instruct