Downloads · 30 days
30
15% of all-time downloads
moutons/CAT-Translate-7b-mlx-4Bit
CAT-Translate-7b-mlx-4Bit is a translation model from moutons. Use it when you need text moved from one language to another. It is set up for transformers. The card lists the license as mit.
The Model moutons/CAT-Translate-7b-mlx-4Bit was converted to MLX format from cyberagent/CAT-Translate-7b using mlx-lm version 0.31.2.
Downloads · 30 days
30
15% of all-time downloads
All-time downloads
198
Public
Parameters
7.5B
4.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors4.2 GB · 100%
How the weights are stored.
U327.5B · 100%
From the Hugging Face model README
The Model moutons/CAT-Translate-7b-mlx-4Bit was converted to MLX format from cyberagent/CAT-Translate-7b using mlx-lm version 0.31.2.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("moutons/CAT-Translate-7b-mlx-4Bit")
prompt="hello"
if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, tokenize=False, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)