Skip to content

DeeLearning

DeepSeek-R1-Distill-Qwen-1.5B-GRPO

DeeLearning/DeepSeek-R1-Distill-Qwen-1.5B-GRPO

DeepSeek-R1-Distill-Qwen-1.5B-GRPO is a machine learning model from DeeLearning. Use it for the machine learning task on the model card, and read the license before you ship it in a product.

Downloads · 30 days

0

Access

Public

Updated Feb 11, 2025

Repo size

—

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

Other1.5 KB · 100%

At a glance

Access
Public
Created
Feb 11, 2025
Updated
Feb 11, 2025
SHA
53854c85
Created
Feb 11, 2025
Updated
Feb 11, 2025
DeepSeek-R1-Distill-Qwen-1.5B-GRPO — AI Model — AIMarketly