Downloads · 30 days
168
15% of all-time downloads
Open4bits/llama3.2-1b-gguf
llama3.2-1b-gguf is a text generation model from Open4bits. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama3.2.
This repository provides the LLaMA 3.2-1B model converted to GGUF format, published by Open4bits to enable highly efficient local inference with reduced memory usage and broad CPU compatibility.
Downloads · 30 days
168
15% of all-time downloads
All-time downloads
1.1K
Public
Repo size
11.2 GB
Likes
0
Public
Click a slice to open those files.
.gguf11.2 GB · 100%
From the Hugging Face model README
This repository provides the LLaMA 3.2-1B model converted to GGUF format, published by Open4bits to enable highly efficient local inference with reduced memory usage and broad CPU compatibility.
The underlying LLaMA 3.2 model and architecture are owned by Meta AI. This repository contains only a quantized GGUF conversion of the original model weights.
The model is designed for fast, lightweight text generation and instruction-following tasks and is well suited for resource-constrained environments.
LLaMA (Large Language Model Meta AI) is a family of transformer-based language models developed by Meta AI. This release uses the 3.2 variant with 1 billion parameters, striking a balance between performance and efficiency.
Compared to larger LLaMA variants, this model offers significantly faster inference with lower memory requirements, with proportionally reduced capacity for complex reasoning.
This model is intended for:
This model is released under the original LLaMA 3.2 license terms as defined by Meta AI. Users must comply with the licensing conditions of the base LLaMA 3.2-1B model.
If you find this model useful, please consider supporting the project. Your support helps Open4bits continue releasing and maintaining high-quality open models for the community.