Downloads · 30 days
14
22% of all-time downloads
mergekit-community/Qwen2-Math-2B-Instruct
Qwen2-Math-2B-Instruct is a text generation model from mergekit-community. Use it when you need the model to write or continue text. It is set up for transformers.
This is a merge of pre-trained language models created using mergekit.
Downloads · 30 days
14
22% of all-time downloads
All-time downloads
64
Public
Parameters
2.2B
4.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors4.3 GB · 100%
From the Hugging Face model README
This is a merge of pre-trained language models created using mergekit.
This model was merged using the passthrough merge method.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
slices:
- sources:
- layer_range: [0, 12]
model: Qwen/Qwen2-Math-1.5B-Instruct
- sources:
- layer_range: [6, 18]
model: Qwen/Qwen2-Math-1.5B-Instruct
parameters:
scale:
- filter: o_proj
value: 0.0
- filter: down_proj
value: 0.0
- value: 1.0
- sources:
- layer_range: [12, 24]
model: Qwen/Qwen2-Math-1.5B-Instruct
merge_method: passthrough
dtype: bfloat16