Skip to content

TAUR-dev

M-1110_star__oursfixed_alltask-rl

TAUR-dev/M-1110_star__oursfixed_alltask-rl

M-1110_star__oursfixed_alltask-rl is a machine learning model from TAUR-dev. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.

- Training Method: VeRL Reinforcement Learning (RL) - Stage Name: rl - Experiment: 1110staroursfixedalltask - RL Framework: VeRL (Versatile Reinforcement Learning)

Downloads · 30 days

12

50% of all-time downloads

All-time downloads

24

Public

Parameters

1.8B

85.3 GB on disk

Likes

0

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors3.6 GB · 100%

At a glance

License
mit
Model type
qwen2
Access
Public
Created
Nov 10, 2025
Updated
Nov 12, 2025
SHA
87bf87c8
Type
qwen2
License
mit
Languages
en
Created
Nov 10, 2025
Updated
Nov 12, 2025
M-1110_star__oursfixed_alltask-rl — AI Model — AIMarketly