Downloads · 30 days
31
8% of all-time downloads
FalconDev25/glm-edge-4b-chat-4bit
glm-edge-4b-chat-4bit is a text generation model from FalconDev25. Use it when you need the model to write or continue text. It is set up for mlx. The card lists the license as other.
This model FalconDev25/glm-edge-4b-chat-4bit was converted to MLX format from zai-org/glm-edge-4b-chat using mlx-lm version 0.29.1.
Downloads · 30 days
31
8% of all-time downloads
All-time downloads
369
Public
Parameters
4.3B
2.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors2.4 GB · 100%
How the weights are stored.
U324.3B · 100%
From the Hugging Face model README
This model FalconDev25/glm-edge-4b-chat-4bit was converted to MLX format from zai-org/glm-edge-4b-chat using mlx-lm version 0.29.1.
pip install mlx-lm
from mlx_lm import load, generate
model, tokenizer = load("FalconDev25/glm-edge-4b-chat-4bit")
prompt = "hello"
if tokenizer.chat_template is not None:
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
messages, add_generation_prompt=True
)
response = generate(model, tokenizer, prompt=prompt, verbose=True)