Downloads · 30 days
66
44% of all-time downloads
YangZhou24/RealGRPO
RealGRPO is a text-to-image model from YangZhou24. Use it when you need an image from a text prompt. It is set up for diffusers.
This repository provides DiT weights fine-tuned from FLUX.1-dev with GRPO using the RealGRPO strategy.
Downloads · 30 days
66
44% of all-time downloads
All-time downloads
150
Public
Parameters
11.9B
47.6 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors47.6 GB · 100%
From the Hugging Face model README
This repository provides DiT weights fine-tuned from FLUX.1-dev with GRPO using the RealGRPO strategy.
RealGRPO targets a common post-training issue in image generation: reward hacking (e.g., over-smoothing, over-saturation, and synthetic-looking artifacts).
Compared with vanilla FLUX and standard GRPO baselines, these weights are optimized to better preserve prompt intent while reducing reward-driven artifacts.
RealGRPO uses a LLM to generate prompt-specific style pairs:
pos_style)neg_style)The reward encourages similarity to positive cues while penalizing negative cues, helping the model avoid artifact-prone shortcuts during alignment.
Note: This release contains DiT alignment weights, not a standalone full pipeline package. You need download black-forest-labs/FLUX.1-dev and replace the contents of the
transfermerdirectory with the contents of this repository.