Downloads · 30 days
0
CanadaHonk/honkhazard-3.1
honkhazard-3.1 is a machine learning model from CanadaHonk. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as apache-2.0.
<font size=+4 face="monospace"honkhazard-3.1</font<br<font size=+1 face="monospace" color="aaa"40.6M (10.49M embed, 16L/8H) | 1.1B seen</font --- a fourth experiment to train only on synthetic messages! very similar t…
Downloads · 30 days
0
Access
Public
Updated Dec 6, 2025
Parameters
40.6M
142 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors142 MB · 100%
How the weights are stored.
F3230.1M · 74%
From the Hugging Face model README
a fourth experiment to train only on synthetic messages! very similar to honkhazard-3 but improved setup
changes vs honkhazard-3:
trained on 1x rtx 5090 in 108.8m:

pre-trained only on SYNTH messages in the following format:
<|bos|><|user_start|>{{query}}<|user_end|><|assistant_start|><|reasoning_start|>{{synthetic_reasoning}}<|reasoning_end|>{{synthetic_answer}}<|assistant_end|>
no post-training of any form has been performed on this model
mostly to test infra and tune LRs. slower training due to vram limitations. seems better than honkhazard-3
final steps: 3.1:
step 23145/23146 (100.00%) | loss: 2.584534 | acc: 47.67% | grad norm: 0.1846 | lr_muon: 2.442e-04 | lr_adam: 5.675e-04 | pnorm: 2.475e+03 | upd/w: 1.795e-08 | dt: 224.51ms | tok/sec: 218,934 | tokens: 1,137,672,192 | eta: 0.00m | total time: 108.79m
Step 23146 | Validation bpb: 0.7132 | Eval time: 4.62s
3:
step 18799/18800 (99.99%) | loss: 2.613449 | grad norm: 0.1186 | lrm: 0.10 | dt: 355.98ms | tok/sec: 184,099 | total time: 129.55m
Step 18800 | Validation bpb: 0.7004
2 (* could be different run in between 2 and 3, before good infra):
step 00379/00380 (99.74%) | loss: 4.421366 | grad norm: 0.0875 | lrm: 0.12 | dt: 20393.21ms | tok/sec: 12,854 | mfu: 0.04 | total time: 126.69m
Step 00380 | Validation bpb: 1.2316