Downloads · 30 days
31
11% of all-time downloads
ZeroXClem/Qwen2.5-1.5B-Instruct-Coder-Math-Bunnycore-Fusion
Qwen2.5-1.5B-Instruct-Coder-Math-Bunnycore-Fusion is a machine learning model from ZeroXClem. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
Please note that ZeroXClem/Qwen2.5-1.5B-Instruct-Coder-Math-Bunnycore-Fusion is currently in Pre-Alpha and under active revision. As such, some features and functionalities may not perform as expected, and the model i…
Downloads · 30 days
31
11% of all-time downloads
All-time downloads
277
Public
Parameters
1.8B
3.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors3.6 GB · 100%
From the Hugging Face model README
Please note that ZeroXClem/Qwen2.5-1.5B-Instruct-Coder-Math-Bunnycore-Fusion is currently in Pre-Alpha and under active revision. As such, some features and functionalities may not perform as expected, and the model is still in the experimental phase. We are continuously refining the architecture, and future updates will improve performance and stability.
Known Issues:
ZeroXClem/Qwen2.5-1.5B-Instruct-Coder-Math-Bunnycore-Fusion is a cutting-edge merged model that combines the finest features from instruction-following, coding, mathematical reasoning, and factual question-answering. This powerhouse is designed for high performance in diverse technical, creative, and interactive tasks.
This model is the fusion of the following:
These models have been seamlessly blended to create a versatile AI that excels across multiple domains.
The model was merged using the DELLA merge method with bfloat16 precision, ensuring high-performance across multiple task types. Here's the configuration used for the merge:
merge_method: della
dtype: bfloat16
parameters:
epsilon: 0.1
lambda: 1.0
normalize: true
base_model: unsloth/Qwen2.5-Coder-1.5B-Instruct
models:
- model: cyixiao/qwen-1.5B-openbookqa
parameters:
weight: 1
density: 0.5
- model: unsloth/Qwen2.5-Coder-1.5B-Instruct
parameters:
weight: 1
density: 0.6
- model: Qwen/Qwen2.5-Math-1.5B-Instruct
parameters:
weight: 1
density: 0.55
- model: bunnycore/Qwen2.5-1.5B-Matrix
parameters:
weight: 1
density: 0.55
- model: Syed-Hasan-8503/Qwen2.5-1.5B-Instruct-WO-Adam-mini
parameters:
weight: 1
density: 0.45
- model: Goekdeniz-Guelmez/Josiefied-Qwen2.5-1.5B-Instruct-abliterated-v3
parameters:
weight: 1
density: 0.5
This model excels in technical coding tasks thanks to the contributions from Qwen2.5-Coder and Matrix.
With Qwen2.5-Math-1.5B-Instruct, the model is perfect for solving complex mathematical problems and structured logical tasks.
Fine-tuned on conversation and identity tasks, the model handles complex dialogue and conversational exchanges with Syed-Hasan-8503.
Thanks to Josiefied-Qwen2.5, this model can operate without restrictions, making it ideal for open-ended instruction-following.
This model is open-sourced under the Apache-2.0 License, allowing others to use and modify it freely, as long as they give proper attribution.
mergeQwenCoderMathBunnycoreinstruction-followinglong-form-generation