Skip to content

ilemon

DirectAlignMitigatesRewardTamperingLoRA

ilemon/DirectAlignMitigatesRewardTamperingLoRA

DirectAlignMitigatesRewardTamperingLoRA is a machine learning model from ilemon. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

这是一个使用了DirectAlign方法的早期版本lora 效果不保证能做到最好 This is an early-version LoRA that adopts the DirectAlign method; its performance is not guaranteed to be optimal.

Downloads · 30 days

0

Access

Public

Updated Nov 14, 2025

Repo size

744 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors744 MB · 100%

At a glance

Access
Public
Created
Nov 14, 2025
Updated
Nov 14, 2025
SHA
b1361749
Created
Nov 14, 2025
Updated
Nov 14, 2025
DirectAlignMitigatesRewardTamperingLoRA — AI Model — AIMarketly