Downloads · 30 days
0
SJTU-DENG-Lab/D2F_LLaDA_Instruct_8B_Lora
D2F_LLaDA_Instruct_8B_Lora is a text generation model from SJTU-DENG-Lab. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This repository contains the LoRA adapter for the GSAI-ML/LLaDA-8B-Instruct model, trained using the Discrete Diffusion Forcing (D2F) method.
Downloads · 30 days
0
Access
Public
Updated Aug 14, 2025
Repo size
67.1 MB
Likes
5
Public
Click a slice to open those files.
.safetensors67.1 MB · 100%
From the Hugging Face model README
This repository contains the LoRA adapter for the GSAI-ML/LLaDA-8B-Instruct model, trained using the Discrete Diffusion Forcing (D2F) method.
This adapter allows the LLaDA-8B-Instruct diffusion LLM (dLLM) to achieve inference speeds that are significantly faster than both its original version and leading autoregressive (AR) models like LLaMA3, while maintaining comparable output quality.
The D2F method and its results are detailed in the paper: D2F: Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing.
Diffusion LLMs (dLLMs) have long promised ultra-fast parallel decoding, but this potential was historically crippled by two main bottlenecks:
D2F solves these issues with a novel hybrid approach:
Hybrid Architecture: D2F reframes text generation as a block-autoregressive process.
Pipelined Parallel Decoding: D2F uses an efficient training and inference strategy.
⚠️ Important: This is a LoRA adapter and requires the official D2F codebase for inference.
For detailed instructions and code, please refer to the official GitHub repository:
➡️ https://github.com/zhijie-group/Discrete-Diffusion-Forcing ⬅️