Downloads · 30 days
16
6% of all-time downloads
minseo25/CDLM-LLaDA
CDLM-LLaDA is a text generation model from minseo25. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as mit.
This repository hosts the LoRA adapter for the LLaDA-8B-Instruct diffusion LLM (dLLM), produced with the CDLM (Consistency Diffusion Language Models) method. CDLM integrates consistency modeling and a block-wise causa…
Downloads · 30 days
16
6% of all-time downloads
All-time downloads
269
Public
Repo size
352 MB
Likes
0
Public
Click a slice to open those files.
.safetensors352 MB · 100%
From the Hugging Face model README
This repository hosts the LoRA adapter for the LLaDA-8B-Instruct diffusion LLM (dLLM), produced with the CDLM (Consistency Diffusion Language Models) method. CDLM integrates consistency modeling and a block-wise causal attention mask so the student model becomes fully KV-cache compatible while retaining the strong local bidirectional modeling within each block. In practice, the adapter enables significantly faster inference with competitive quality.
adapter_model.safetensors, adapter_config.json)This is a LoRA adapter, not a full model. You must load the base model and then attach this adapter. For best speedups, use the CDLM inference path in the accompanying codebase.
This adapter is released under the MIT License. The base model is governed by its own license; please ensure compliance with the base model’s terms.
@article{kim2025cdlm,
title = {CDLM: Consistency Diffusion Language Models for Faster Sampling},
author = {Kim, Minseo and Xu, Chenfeng and Hooper, Coleman and Singh, Harman
and Athiwaratkun, Ben and Zhang, Ce and Keutzer, Kurt and Gholami, Amir},
journal = {arXiv preprint arXiv:2511.19269},
year = {2025},
url = {https://arxiv.org/abs/2511.19269}
}