Downloads · 30 days
10
18% of all-time downloads
mousezhang/math-llada1.5
math-llada1.5 is a text generation model from mousezhang. Use it when you need the model to write or continue text. It is set up for peft.
This repository contains a PEFT LoRA adapter fine-tuned for math reasoning on top of GSAI-ML/LLaDA-1.5.
Downloads · 30 days
10
18% of all-time downloads
All-time downloads
55
Public
Repo size
671 MB
Likes
0
Public
Click a slice to open those files.
.safetensors671 MB · 99%
From the Hugging Face model README
This repository contains a PEFT LoRA adapter fine-tuned for math reasoning on top of
GSAI-ML/LLaDA-1.5.
The repository stores adapter weights only. It does not include the base model weights.
adapter_model.safetensors: LoRA adapter weights.adapter_config.json: PEFT LoRA configuration.tokenizer.json, tokenizer_config.json, special_tokens_map.json: tokenizer files used with the adapter.trainer_state.json, training_args.bin: training metadata kept for traceability.Optimizer, scheduler, and RNG checkpoint files are intentionally not included.
from peft import PeftModel
from transformers import AutoModelForCausalLM, AutoTokenizer
base_model_id = "GSAI-ML/LLaDA-1.5"
adapter_id = "mousezhang/math-llada1.5"
tokenizer = AutoTokenizer.from_pretrained(base_model_id, trust_remote_code=True)
base_model = AutoModelForCausalLM.from_pretrained(
base_model_id,
trust_remote_code=True,
)
model = PeftModel.from_pretrained(base_model, adapter_id)
GSAI-ML/LLaDA-1.5CAUSAL_LM1281280.05q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_projThis model card does not report benchmark scores. Evaluation results should be treated as not provided unless published separately by the model author.
This adapter inherits the limitations and license terms of the base model. It is intended for research and experimental use, and outputs should be checked carefully before use in high-stakes settings.