Skip to content

THU-KEG

LLaDA-8B-BGPO-code

THU-KEG/LLaDA-8B-BGPO-code

LLaDA-8B-BGPO-code is a reinforcement learning model from THU-KEG. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.

[](https://arxiv.org/abs/2510.11683) [](https://github.com/THU-KEG/BGPO)

Downloads · 30 days

18

14% of all-time downloads

All-time downloads

131

Public

Parameters

8B

16 GB on disk

Likes

1

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors16 GB · 100%

At a glance

Task
Reinforcement Learning
License
apache-2.0
Model type
llada
Access
Public
Created
Oct 11, 2025
Updated
Oct 14, 2025
SHA
ee75175f
Task
Reinforcement Learning
Type
llada
License
apache-2.0
Languages
en
Created
Oct 11, 2025
Updated
Oct 14, 2025
LLaDA-8B-BGPO-code — AI Model — AIMarketly