Downloads · 30 days
135
3% of all-time downloads
QuantPanda/LongWriter-glm4-9B-GGUF
LongWriter-glm4-9B-GGUF is a text generation model from QuantPanda. Use it when you need the model to write or continue text. The card lists the license as other.
Original model link: https://huggingface.co/THUDM/LongWriter-glm4-9b
Downloads · 30 days
135
3% of all-time downloads
All-time downloads
4.7K
Public
Repo size
81.8 GB
Likes
5
Public
Click a slice to open those files.
.gguf81.8 GB · 100%
From the Hugging Face model README
Original model link: https://huggingface.co/THUDM/LongWriter-glm4-9b
Model by: THUDM
Quants by: QuantPanda
GGUF quantization for llama.cpp and similar applications.
Example:
./llama-cli -m LongWriter-glm4-9B-Q5_K_M.gguf -p "You are a helpful AI assistant." --conversation
If the model takes too long to load you can reduce the context size with --ctx-size
Example with smaller context size:
./llama-cli -m LongWriter-glm4-9B-Q5_K_M.gguf -p "You are a helpful AI assistant." --conversation --ctx-size 4096