Downloads · 30 days
342
51% of all-time downloads
pipenetwork/LongCat-2.0-4bit
LongCat-2.0-4bit is a text generation model from pipenetwork. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as mit.
4-bit MLX quantization (group-size 64, router classifiers @8-bit, MTP dropped) of meituan-longcat/LongCat-2.0 — a 1.6T / ~48B-active MoE. Converted from the FP8 source with mlx-lm.
Downloads · 30 days
342
51% of all-time downloads
All-time downloads
668
Public
Parameters
1.6T
922 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors922 GB · 100%
How the weights are stored.
U321.6T · 100%
From the Hugging Face model README
4-bit MLX quantization (group-size 64, router classifiers @8-bit, MTP dropped) of meituan-longcat/LongCat-2.0 — a 1.6T / ~48B-active MoE. Converted from the FP8 source with mlx-lm.
pip install git+https://github.com/ml-explore/mlx-lm.git@refs/pull/1464/head
from mlx_lm import load, generate
model, tok = load("pipenetwork/LongCat-2.0-4bit")
p = tok.apply_chat_template([{"role":"user","content":"Who is Albert Einstein?"}], add_generation_prompt=True)
print(generate(model, tok, prompt=p, max_tokens=512, verbose=True))