Downloads · 30 days
0
ram-lexsi/aligntune-testrun-SimPO
aligntune-testrun-SimPO is a machine learning model from ram-lexsi. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
<div align="center" <table border="0" cellspacing="0" cellpadding="0" style="border: none; border-collapse: collapse;" <tr <td align="center" style="border: none; vertical-align: middle;" <a href="https://lexsi.ai/"<i…
Downloads · 30 days
0
Access
Public
Updated Aug 27, 2026
Repo size
15.8 MB
Likes
0
Public
Click a slice to open those files.
.json11.4 MB · 72%
From the Hugging Face model README
Built using AlignTune — supports any open-source model, any algorithm, any backend (TRL / Unsloth / ES / etc).
| Finetuned from | Qwen/Qwen2.5-0.5B-Instruct |
| Algorithm | simpo |
| Backend | trl |
| Artifact | adapter |
| Published | 2026-08-27 05:49 UTC |
from peft import AutoPeftModelForCausalLM
from transformers import AutoTokenizer
model = AutoPeftModelForCausalLM.from_pretrained("ram-lexsi/aligntune-testrun-SimPO")
tokenizer = AutoTokenizer.from_pretrained("ram-lexsi/aligntune-testrun-SimPO")
This repo is a LoRA adapter. Load it on top of Qwen/Qwen2.5-0.5B-Instruct (PEFT does that from adapter_config.json).