Downloads · 30 days
15
16% of all-time downloads
renhouxing/ME-DLM-Stage2
ME-DLM-Stage2 is a text generation model from renhouxing. Use it when you need the model to write or continue text. It is set up for transformers.
This repository contains the Stage 3 checkpoint for ME-DLM, as presented in the paper Edit-Based Refinement for Parallel Masked Diffusion Language Models.
Downloads · 30 days
15
16% of all-time downloads
All-time downloads
93
Public
Parameters
8.5B
17.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors17.1 GB · 100%
From the Hugging Face model README
This repository contains the Stage 3 checkpoint for ME-DLM, as presented in the paper Edit-Based Refinement for Parallel Masked Diffusion Language Models.
Authors: Houxing Ren, Mingjie Zhan, Zimu Lu, Ke Wang, Yunqiao Yang, Haotian Hou, Junting Pan, Hongsheng Li.
<p align="center"> <a href="https://huggingface.co/papers/2605.09603">📄 Paper</a> • <a href="https://github.com/renhouxing/ME-DLM">🏠 Repo</a> • <a href="https://huggingface.co/renhouxing/ME-DLM-Stage3">🤖 Models</a> </p>ME-DLM is a lightweight edit-based refinement framework for masked diffusion language models. It first generates a complete response through parallel diffusion decoding, then refines the output with minimal edit operations such as replacement, deletion, and insertion, conditioned on the full sequence. By using edit distance as deterministic training supervision, ME-DLM improves sequence-level consistency while preserving the decoding efficiency of diffusion models. Built on LLaDA, it achieves consistent gains on HumanEval and GSM8K while using only one-eighth of the total diffusion steps.
@article{ren2025edit,
title={Edit-Based Refinement for Parallel Masked Diffusion Language Models},
author={Ren, Houxing and Zhan, Mingjie and Lu, Zimu and Ke Wang and Yang, Yunqiao and Hou, Haotian and Pan, Junting and Li, Hongsheng},
journal={arXiv preprint arXiv:2605.09603},
year={2025}
}
We thank the following amazing projects that truly inspired us: