Downloads · 30 days
11
5% of all-time downloads
kechengcode/Qwen2-5B-Instruct-16Layers
Qwen2-5B-Instruct-16Layers is a text generation model from kechengcode. Use it when you need the model to write or continue text. It is set up for transformers.
This is a merge of pre-trained language models created using mergekit.
Downloads · 30 days
11
5% of all-time downloads
All-time downloads
215
Public
Parameters
4.8B
9.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors9.6 GB · 100%
From the Hugging Face model README
This is a merge of pre-trained language models created using mergekit.
This model was merged using the passthrough merge method.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
dtype: float16
merge_method: passthrough
slices:
- sources:
- layer_range: [0, 6]
model: Qwen/Qwen2-7B-Instruct
- sources:
- layer_range: [18, 28]
model: Qwen/Qwen2-7B-Instruct