Skip to content

AngelRaychev

0.5B-policy-iteration_4

AngelRaychev/0.5B-policy-iteration_4

0.5B-policy-iteration_4 is a text generation model from AngelRaychev. Use it when you need the model to write or continue text. It is set up for transformers.

This model is a fine-tuned version of AngelRaychev/0.5B-policy-iteration3. It has been trained using TRL.

Downloads · 30 days

25

10% of all-time downloads

All-time downloads

248

Public

Parameters

494M

20.8 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors2 GB · 66%

At a glance

Task
Text Generation
Library
transformers
Model type
qwen2
Access
Public
Created
Apr 29, 2025
Updated
May 14, 2025
SHA
db161a84

Try a prompt

Base models

Task
Text Generation
Library
transformers
Type
qwen2
Created
Apr 29, 2025
Updated
May 14, 2025