Downloads · 30 days
9
16% of all-time downloads
aphoticshaman/elle-72b-math-v2
elle-72b-math-v2 is a text generation model from aphoticshaman. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
Elle-72B-Math-v2 is a LoRA adapter fine-tuned on Qwen/Qwen2.5-72B-Instruct for mathematical reasoning using the NuminaMath-CoT dataset.
Downloads · 30 days
9
16% of all-time downloads
All-time downloads
55
Public
Repo size
5.1 GB
Likes
0
Public
Click a slice to open those files.
.safetensors6.7 GB · 79%
From the Hugging Face model README
Elle-72B-Math-v2 is a LoRA adapter fine-tuned on Qwen/Qwen2.5-72B-Instruct for mathematical reasoning using the NuminaMath-CoT dataset.
Fine-tuned using:
from peft import PeftModel
from transformers import AutoModelForCausalLM, AutoTokenizer
base_model = AutoModelForCausalLM.from_pretrained(
"Qwen/Qwen2.5-72B-Instruct",
torch_dtype="auto",
device_map="auto",
trust_remote_code=True
)
model = PeftModel.from_pretrained(base_model, "aphoticshaman/elle-72b-math-v2")
tokenizer = AutoTokenizer.from_pretrained("Qwen/Qwen2.5-72B-Instruct")
Apache 2.0