Downloads · 30 days
98
7% of all-time downloads
MoMonir/codegeex4-all-9b-GGUF
codegeex4-all-9b-GGUF is a text generation model from MoMonir. Use it when you need the model to write or continue text. The card lists the license as other.
This model was converted to GGUF format from THUDM/codegeex4-all-9b using llama.cpp. Refer to the original model card for more details on the model.
Downloads · 30 days
98
7% of all-time downloads
All-time downloads
1.4K
Public
Repo size
40.5 GB
Likes
0
Public
Click a slice to open those files.
.gguf40.5 GB · 100%
From the Hugging Face model README
This model was converted to GGUF format from THUDM/codegeex4-all-9b using llama.cpp.
Refer to the original model card for more details on the model.
GGUF is a new format introduced by the llama.cpp team on August 21st 2023. It is a replacement for GGML, which is no longer supported by llama.cpp.
Here is an incomplete list of clients and libraries that are known to support GGUF:
Install llama.cpp through brew.
brew install ggerganov/ggerganov/llama.cpp
Invoke the llama.cpp server or the CLI. CLI:
llama-cli --hf-repo MoMonir/codegeex4-all-9b-GGUF --model codegeex4-all-9b.Q4_K_M.gguf -p "The meaning to life and the universe is"
Server:
llama-server --hf-repo MoMonir/codegeex4-all-9b-GGUF --model codegeex4-all-9b.Q4_K_M.gguf -c 2048
Note: You can also use this checkpoint directly through the usage steps listed in the Llama.cpp repo as well.
git clone https://github.com/ggerganov/llama.cpp && \
cd llama.cpp && \
make && \
./main -m codegeex4-all-9b.Q4_K_M.gguf -n 128