Downloads · 30 days
85
19% of all-time downloads
gold24k/v2
v2 is a text generation model from gold24k. Use it when you need the model to write or continue text. It is set up for transformers.
This is a standalone, merged BF16 checkpoint derived from unconst/Affine-5czsc2fc98-r1032-vera-odpo-midrank-hibeta-shortctx-ultraextra-ep4-midlr-merged. It applies a
Downloads · 30 days
85
19% of all-time downloads
All-time downloads
459
Public
Parameters
34.7B
69.3 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors69.3 GB · 100%
From the Hugging Face model README
This is a standalone, merged BF16 checkpoint derived from
unconst/Affine-5czsc2fc98-r1032-vera-odpo-midrank-hibeta-shortctx-ultraextra-ep4-midlr-merged. It applies a
scaled selective-fallback LoRA trained on preserved positive turns and
sanitized, task-specific alternatives for high-confidence negative turns. It
does not require a runtime router, custom Python code, or a PEFT adapter.
62dfb322fdce5873543bd92692ab4ecc3e13f9411.00The table compares the scaled adapter with the untouched R1032 parent. These small held-out metrics selected the merge strength; they are not a substitute for the exact full Affine duel on the dedicated evaluator.
| route | rows | mean reward margin | preference accuracy |
|---|---|---|---|
| fallback | 13 | 0.104621 | 0.6154 |
| preserve | 50 | 0.094390 | 0.5200 |
Experimental candidate. Before submission, run exact stock-vLLM Affine duels,
the exploit-pattern audit, repository preflight, and the official submission
client check. selective_fallback_provenance.json contains the machine-readable
training and merge record.