Skip to content

SpectralPO

DeepSeek-R1-Distill-Llama-8B-GRPO

SpectralPO/DeepSeek-R1-Distill-Llama-8B-GRPO

DeepSeek-R1-Distill-Llama-8B-GRPO is a machine learning model from SpectralPO. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.

Downloads · 30 days

17

25% of all-time downloads

All-time downloads

67

Public

Parameters

8B

16.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors16.1 GB · 100%

At a glance

License
mit
Model type
llama
Access
Public
Created
May 18, 2025
Updated
May 18, 2025
SHA
f36e9b55
Type
llama
License
mit
Created
May 18, 2025
Updated
May 18, 2025