Downloads · 30 days
0
billyenrizky/FS-DFM-1.3B-SFT
FS-DFM-1.3B-SFT is a reinforcement learning model from billyenrizky. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
FS-DFM 1.3B (Apple) fine-tuned with SFT on FormFactory web form-filling tasks. Uses LoRA adapters on the DiT architecture with Poisson jump sampling. Achieves 68.5% nonzero reward rate and 0.146 average reward on 124…
Downloads · 30 days
0
Access
Public
Updated Mar 31, 2026
Repo size
11.1 MB
Likes
0
Public
Click a slice to open those files.
.pt11.1 MB · 76%
From the Hugging Face model README
FS-DFM 1.3B (Apple) fine-tuned with SFT on FormFactory web form-filling tasks. Uses LoRA adapters on the DiT architecture with Poisson jump sampling. Achieves 68.5% nonzero reward rate and 0.146 average reward on 124 test tasks. Part of the STAD80 project: Generative Action Planning via Discrete Flow Matching.
Concentrate or Collapse: When Reinforcement Learning Meets Diffusion Language Models for Web Planning
If you use this model, please cite:
@article{brillian2026flowgrpo,
title={Concentrate or Collapse: When Reinforcement Learning Meets Diffusion Language Models for Web Planning},
author={Brillian, Muhammad Enrizky},
year={2026}
}
This model was trained and evaluated on the FormFactory benchmark:
@misc{li2025formfactory,
title={FormFactory: An Interactive Benchmarking Suite for Multimodal Form-Filling Agents},
author={Bobo Li and Yuheng Wang and Hao Fei and Juncheng Li and Wei Ji and Mong-Li Lee and Wynne Hsu},
year={2025},
eprint={2506.01520},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2506.01520}
}