Skip to content

LLParallax

sf_finetuning_forgetting_human_monk

LLParallax/sf_finetuning_forgetting_human_monk

sf_finetuning_forgetting_human_monk is a reinforcement learning model from LLParallax. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for sample-factory.

A(n) APPO model trained on the challenge environment.

Downloads · 30 days

0

Access

Public

Updated Apr 7, 2024

Repo size

1.4 GB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.pth1.4 GB · 99%

At a glance

Task
Reinforcement Learning
Library
sample-factory
Access
Public
Created
Apr 7, 2024
Updated
Apr 7, 2024
SHA
8ff1ac52
Task
Reinforcement Learning
Library
sample-factory
Created
Apr 7, 2024
Updated
Apr 7, 2024