Downloads · 30 days
55
6% of all-time downloads
DavidOKB/MathThink-Qwen-3.5-4B
MathThink-Qwen-3.5-4B is a machine learning model from DavidOKB. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
This model is a fine-tuned version of Qwen3.5-4B, explicitly optimized for complex mathematical reasoning and Chain-of-Thought (CoT) problem solving. It was fine-tuned using the Nemotron-Math-v3 dataset with Parameter…
Downloads · 30 days
55
6% of all-time downloads
All-time downloads
921
Public
Repo size
17.4 GB
Likes
12
Public
Click a slice to open those files.
.gguf9 GB · 52%
From the Hugging Face model README
This model is a fine-tuned version of Qwen3.5-4B, explicitly optimized for complex mathematical reasoning and Chain-of-Thought (CoT) problem solving. It was fine-tuned using the Nemotron-Math-v3 dataset with Parameter-Efficient Fine-Tuning (PEFT/LoRA).
Qwen/Qwen3.5-4Bnvidia/Nemotron-SFT-Math-v3lora_alpha scaling is specifically tuned to prevent catastrophic forgetting, ensuring the model retains conversational abilities while significantly enhancing mathematical logic.F16) and GGUF formats (Q8_0)Because this model leverages extensive Chain-of-Thought reasoning to solve math problems, the following generation parameters are highly recommended for the best performance:
{
"temperature": 1.0,
"top_p": 0.95,
"repetition_penalty": 1.1
}
Note: A repetition_penalty of 1.1 is crucial to prevent the base model from occasionally falling into infinite generation loops on extremely long context windows.