Downloads · 30 days
7
17% of all-time downloads
enferAI/MythoMax-L2-13b-FP8-dynamic
MythoMax-L2-13b-FP8-dynamic is a machine learning model from enferAI. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
Quantized version of Gryphe/MythoMax-L2-13b.
Downloads · 30 days
7
17% of all-time downloads
All-time downloads
42
Public
Parameters
13B
13.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.3 GB · 100%
How the weights are stored.
F8_E4M312.7B · 97%
From the Hugging Face model README
Quantized version of Gryphe/MythoMax-L2-13b.
This model was created with llm-compressor by running the code snippet below.
from llmcompressor.modifiers.quantization import QuantizationModifier
from llmcompressor.transformers import oneshot
from transformers import AutoModelForCausalLM, AutoTokenizer
# Load model
model_stub = "Gryphe/MythoMax-L2-13b"
model_name = model_stub.split("/")[-1]
model = AutoModelForCausalLM.from_pretrained(
model_stub,
torch_dtype="auto",
)
tokenizer = AutoTokenizer.from_pretrained(model_stub)
# Configure the quantization algorithm and scheme
recipe = QuantizationModifier(
targets="Linear",
scheme="FP8_DYNAMIC",
ignore=["lm_head"],
)
# Apply quantization
oneshot(
model=model,
recipe=recipe,
)
# Save to disk in compressed-tensors format
save_path = model_name + "-FP8-dynamic"
model.generation_config.do_sample = True
model.save_pretrained(save_path)
tokenizer.save_pretrained(save_path)
print(f"Model and tokenizer saved to: {save_path}")