Downloads · 30 days
5
7% of all-time downloads
8Fai/chorous-100m
chorous-100m is a time series forecasting model from 8Fai. Use it for the time series forecasting task on the model card, and read the license before you ship it in a product. It is set up for torch. The card lists the license as other.
<p align="center" <img src="https://img.shields.io/badge/Parameters-100M%20%7C%2050M%20%7C%2027M-brightgreen?style=flat-square" / <img src="https://img.shields.io/badge/Architecture-Patch--Transformer-purple?style=fla…
Downloads · 30 days
5
7% of all-time downloads
All-time downloads
74
Public
Parameters
75.6M
Likes
1
Public
Click a slice to open those files.
.safetensors302 MB · 100%
From the Hugging Face model README
Chorous1 is a suite of three high-performance, patch-based transformer models for multivariate time-series forecasting. Combining RevIN, MAE-style patch masking, and a Flatten Head architecture, Chorous1 delivers state-of-the-art accuracy on real-world benchmark data.
| Variant | Parameters | Hidden Size | Layers | Query Heads / KV Heads |
|---|---|---|---|---|
chorous1-100m | ~100M | 768 | 12 | 12 / 4 |
chorous1-50m | ~50M | 512 | 16 | 8 / 2 |
chorous1-27m | ~27M | 384 | 16 | 6 / 2 |
| Component | Specification |
|---|---|
| Context Length | 512 steps |
| Forecast Horizon | 96 steps |
| Patch Size | 16 (non-overlapping) |
| Number of Patches | 32 |
| FFN Multiplier | 2.667× |
| Activation | SwiGLU |
| Positional Encoding | RoPE (θ = 500,000) |
| Normalization | RMSNorm |
| Masking Ratio | 25% (training only) |
| Loss Function | Huber Loss + MAE |
| Precision | bfloat16 |
Stage 1 — Neural Encoding. The transformer encoder processes patches of time-series data using RoPE and GQA to capture long-range temporal dependencies and periodic structure.
Stage 2 — RevIN Normalization. A reversible instance normalization layer removes mean and variance shifts from the input prior to processing, then restores them on the output — eliminating the distribution mismatch problem common in real-world deployments.
import torch
from safetensors.torch import load_file
# Replace "100m" with "50m" or "27m" as needed
weights = load_file("./chorous_checkpoint/100m/model.safetensors")
model.load_state_dict(weights)
model.eval()
# Input shape: [Batch, Channels, Time]
x = torch.randn(1, 7, 512)
with torch.no_grad():
forecast = model(x) # Output shape: [1, 7, 96]
| Metric | chorous1-100m | chorous1-50m | chorous1-27m |
|---|---|---|---|
| Weights Size | ~200 MB | ~110 MB | ~65 MB |
| VRAM (Inference) | ~12 GB | ~8 GB | ~6 GB |
Chorous1 is released under the 8f-ai-license-v1.0. Please review the full terms before use in production or commercial applications.