Downloads · 30 days
718
11% of all-time downloads
d3LLM/d3LLM_LLaDA
d3LLM_LLaDA is a text generation model from d3LLM. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This repository contains d3LLM-LLaDA, an ultra-fast diffusion language model presented in the paper d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation.
Downloads · 30 days
718
11% of all-time downloads
All-time downloads
6.7K
Public
Parameters
8B
16 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors16 GB · 100%
From the Hugging Face model README
This repository contains d3LLM-LLaDA, an ultra-fast diffusion language model presented in the paper d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation.
d3LLM-LLaDA is an ultra-fast diffusion language model that strikes a balance between accuracy and parallelism. It uses pseudo-trajectory distillation to teach the model which tokens can be decoded confidently at early steps, and employs an entropy-based multi-block decoding mechanism with KV-cache refresh during inference.
To use this model, it is recommended to clone the official repository and install the required dependencies:
# Clone the repository
git clone https://github.com/hao-ai-lab/d3LLM.git
cd d3LLM
# Install dependencies
pip install -r requirements.txt
If you find d3LLM useful for your research, please cite the following work:
@inproceedings{ICML'26:d3llm,
title = {d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation},
author = {Yu-Yang Qian and Junda Su and Lanxiang Hu and Peiyuan Zhang and Zhijie Deng and Peng Zhao and Hao Zhang},
booktitle = {Proceedings of the 43rd International Conference on Machine Learning (ICML)},
pages = {to appear},
year = {2026}
}