Skip to content

efficient-moe-agent-project

scienceworld

efficient-moe-agent-project/scienceworld

scienceworld is a reinforcement learning model from efficient-moe-agent-project. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as apache-2.0.

LoRA adapters for Qwen/Qwen2.5-1.5B-Instruct, trained as a single-expert agent on the ScienceWorld text environment (30 elementary-science tasks) with GiGPO reinforcement learning, warm-started from behaviour cloning…

Downloads · 30 days

0

Access

Public

Updated Aug 19, 2026

Repo size

886 MB

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors886 MB · 100%

At a glance

Task
Reinforcement Learning
Library
peft
License
apache-2.0
Access
Public
Created
Aug 5, 2026
Updated
Aug 19, 2026
SHA
ce0b3dc4

Base models

Task
Reinforcement Learning
Library
peft
License
apache-2.0
Created
Aug 5, 2026
Updated
Aug 19, 2026