Downloads · 30 days
0
ZhengyanWan/dFlowGRPO-PickScore
dFlowGRPO-PickScore is a machine learning model from ZhengyanWan. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as apache-2.0.
LoRA adapter for FUDOKI, fine-tuned with dFlowGRPO on the PickScore reward. Training step: 6300.
Downloads · 30 days
0
Access
Public
Updated May 8, 2026
Repo size
90.9 MB
Likes
0
Public
Click a slice to open those files.
.safetensors90.9 MB · 100%
From the Hugging Face model README
LoRA adapter for FUDOKI,
fine-tuned with dFlowGRPO on the PickScore reward.
Training step: 6300.
See WanZhengyan/dFlowGRPO for training & evaluation code.
lora_adapter_emahuggingface-cli download ZhengyanWan/dFlowGRPO-PickScore \
--local-dir output_grpo_pickscore/checkpoint_0_step6300
Then run the matching evaluator from discrete_flow_grpo/.