Skip to content

BeastxD

text2cypher_lora_v6_shaped

BeastxD/text2cypher_lora_v6_shaped

text2cypher_lora_v6_shaped is a reinforcement learning model from BeastxD. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.

Overnight experiment, 2026-08-25/26: GRPO on top of BeastxD/text2cypherlorav7, using a shaped, partial-credit reward instead of the binary +1/0/-0.5 reward the first GRPO attempt (BeastxD/text2cypherlorav6grpo) used.

Downloads · 30 days

5

36% of all-time downloads

All-time downloads

14

Public

Parameters

4B

8.1 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors8 GB · 100%

At a glance

Task
Reinforcement Learning
License
apache-2.0
Model type
qwen3
Access
Public
Created
Aug 26, 2026
Updated
Aug 26, 2026
SHA
72b6c7c5

Base models

Task
Reinforcement Learning
Type
qwen3
License
apache-2.0
Languages
en
Created
Aug 26, 2026
Updated
Aug 26, 2026