Downloads · 30 days
19
23% of all-time downloads
DataSnake/Wayfarer-12B-NVFP4
Wayfarer-12B-NVFP4 is a text generation model from DataSnake. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
Downloads · 30 days
19
23% of all-time downloads
All-time downloads
84
Public
Parameters
6.8B
8.8 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors8.8 GB · 100%
How the weights are stored.
U85.5B · 73%
From the Hugging Face model README

Quantized NVFP4 weights of the Wayfarer-12B model, for use with nVidia Blackwell GPUs.
Quantized with TensorRT-Model-Optimizer 0.37.0
Calibrated using the distilled-roleplay dataset, tagged in the same ChatML format used to train the Wayfarer and Muse models in the first place. This was accomplished by adding the following code to the start of hf_ptq.py:
from modelopt.torch.utils import dataset_utils
dataset_utils.SUPPORTED_DATASET_CONFIG["distilled-roleplay"] = {
"config": {
"path": "agentlans/distilled-roleplay",
"split": ["train"],
},
"preprocess": lambda sample: "".join(
f"<|im_start|>{ {'system':'system','human':'user','gpt':'assistant'}[turn['from']] }\n"
f"{turn['value'].strip()}<|im_end|>\n"
for turn in sample["conversations"]
),
}
Tested on a RTX 5060 Ti 16GB with TensorRT-LLM, vLLM, SGLang, and Aphrodite Engine.
Recommended generation settings (a mix of what it says on the Wayfarer-12B model card and the AI Dungeon Model Guide):
As mentioned above, the calibration data was provided with the same ChatML tags as had been used to finetune Latitude's 12B models:
<|im_start|>system
You're a masterful storyteller and gamemaster. Write in second person present tense (You are), crafting vivid, engaging narratives with authority and confidence.<|im_end|>
<|im_start|>user
> You peer into the darkness.<|im_end|>
<|im_start|>assistant
You have been eaten by a grue.<|im_end|>
As such, I would recommend using that format for inference.
Wayfarer-12B was made by Latitude Games with help from Gryphe Padar