Downloads · 30 days
60
100% of all-time downloads
alst10/Llama-3.1-8B-Instruct-ROCMFP4_FAST
Llama-3.1-8B-Instruct-ROCMFP4_FAST is a machine learning model from alst10. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as cc-by-4.0.
This model was dynamically quantized and generated using the ROCmFPX My Repo Space Original Base Model: meta-llama/Llama-3.1-8B-Instruct
Downloads · 30 days
60
100% of all-time downloads
All-time downloads
60
Public
Repo size
4.3 GB
Likes
0
Public
Click a slice to open those files.
.gguf4.3 GB · 100%
From the Hugging Face model README
This model was dynamically quantized and generated using the ROCmFPX My Repo Space
To run this model, you must use the specialized ROCmFPX fork of llama.cpp by charlie12345. Standard llama.cpp releases do not currently support ROCmFPX quantization formats.
# 1. Clone the ROCmFPX repository
git clone --depth 1 [https://github.com/charlie12345/ROCmFPX.git](https://github.com/charlie12345/ROCmFPX.git)
cd ROCmFPX
# 2. Build the project
cmake -B build -G Ninja -DGGML_CUDA=OFF -DGGML_NATIVE=OFF
cmake --build build --target llama-cli -j
# 3. Run the model
./build/bin/llama-cli -m Llama-3.1-8B-Instruct-Q4_0_ROCMFP4_FAST.gguf -p "You are a helpful assistant. Hello!" -n 128
This model is distributed under the Creative Commons Attribution 4.0 International (CC BY 4.0) license. Please adhere to the original license constraints of the base model and provide appropriate attribution when sharing.