Skip to content

paradigma-inc

limite-1b-value-model

paradigma-inc/limite-1b-value-model

limite-1b-value-model is a reinforcement learning model from paradigma-inc. Use it for the reinforcement learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.

A 1.035B-parameter value model for mathematical reasoning. Given a question and a proposed solution, it returns one value per response token. Each value estimates the final correctness return from the reasoning availa…

Downloads · 30 days

0

Access

Public

Updated Sep 22, 2026

Parameters

1B

4.2 GB on disk

Likes

5

Public

Hugging Face

Repo makeup

Click a slice to open those files.

.safetensors4.1 GB · 100%

At a glance

Task
Reinforcement Learning
License
apache-2.0
Access
Public
Created
Sep 21, 2026
Updated
Sep 22, 2026
SHA
ce83740e
Task
Reinforcement Learning
License
apache-2.0
Languages
en
Created
Sep 21, 2026
Updated
Sep 22, 2026