Skip to content

ZhengyanWan

dFlowGRPO-ScienceQA

ZhengyanWan/dFlowGRPO-ScienceQA

dFlowGRPO-ScienceQA is a machine learning model from ZhengyanWan. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as apache-2.0.

LoRA adapter for FUDOKI, fine-tuned with dFlowGRPO on the ScienceQA reward. Training step: 1500.

Downloads · 30 days

0

Access

Public

Updated May 8, 2026

Repo size

182 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pt90.9 MB · 50%

At a glance

Library
peft
License
apache-2.0
Access
Public
Created
May 8, 2026
Updated
May 8, 2026
SHA
bd33a73f

Base models

Library
peft
License
apache-2.0
Created
May 8, 2026
Updated
May 8, 2026