Downloads · 30 days
3.3K
50% of all-time downloads
QuantFactory/N-ATLaS-GGUF
N-ATLaS-GGUF is a machine learning model from QuantFactory. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Downloads · 30 days
3.3K
50% of all-time downloads
All-time downloads
6.6K
Public
Repo size
72.7 GB
Likes
5
Public
Click a slice to open those files.
.gguf72.7 GB · 100%
From the Hugging Face model README
This is quantized version of NCAIR1/N-ATLaS created using llama.cpp
N-ATLaS-LLM is a fine-tuned multilingual language model based on Llama-3 8B, specifically designed to support African languages, including Hausa, Igbo, and Yoruba alongside English. This model is powered by Awarri Technologies an initiative of the Federal Ministry of Communications, Innovation and Digital Economy as part of the Nigerian Languages AI Initiative to promote digital inclusion and preserve African linguistic heritage in the digital age.
N-ATLaS-LLM is built on the Llama architecture and has been fine-tuned on over 400 million tokens of multilingual instruction data. The model demonstrates strong performance across multiple African languages while maintaining excellent English capabilities.
| Parameter | Value |
|---|---|
| Model Type | LlamaForCausalLM |
| Base Model | Llama-3 8B |
| Hidden Size | 4,096 |
| Intermediate Size | 14,336 |
| Number of Layers | 32 |
| Attention Heads | 32 |
| Key-Value Heads | 8 |
| Head Dimension | 128 |
| Vocabulary Size | 128,256 |
| Max Position Embeddings | 131,072 |
| Context Length | 8,092 tokens |
N-ATLaS-LLM was trained on approximately 391,956,264 tokens of quality multilingual instruction data.
| Language | SFT Samples |
|---|---|
| English | ~318,000 |
| Hausa | ~200,000 |
| Igbo | ~200,000 |
| Yoruba | ~200,000 |
Our model was evaluated by human annotators across multiple dimensions. Here are the results:
| Metric | English | Hausa | Yoruba | Igbo |
|---|---|---|---|---|
| Evaluations | 1,662 | 140 | 542 | 296 |
| Average Score | 4.21/5.0 | 3.98/5.0 | 2.69/5.0 | 3.87/5.0 |
| Fluency | 4.30/5.0 | 4.23/5.0 | 2.71/5.0 | 3.89/5.0 |
| Coherence | 4.22/5.0 | 3.70/5.0 | 3.23/5.0 | 3.80/5.0 |
| Relevance | 4.28/5.0 | 3.76/5.0 | 2.89/5.0 | 3.85/5.0 |
| Accuracy | 4.23/5.0 | 3.72/5.0 | 3.13/5.0 | 3.92/5.0 |
| Bias/Fairness | 3.18/5.0 | 1.11/5.0 | 2.23/5.0 | 4.01/5.0 |
| Usefulness | 4.09/5.0 | 5.00/5.0 | 4.03/5.0 | 3.84/5.0 |
pip install transformers torch
from transformers import AutoTokenizer, AutoModelForCausalLM
import torch
# Load model and tokenizer
model_name = "NCAIR1/N-ATLaS"
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(
model_name,
torch_dtype=torch.float16,
device_map="auto"
)
def format_text_for_inference(messages):
current_date = datetime.now().strftime('%d %b %Y')
text = tokenizer.apply_chat_template(
messages,
add_generation_prompt=True,
tokenize=False,
date_string=current_date
)
return text
# Example usage
q_chat = [
{'role':'system','content':'you are a large language model trained by Awarri AI technologies. You are a friendly assistant and you are here to help.'},
{'role': 'user', 'content': "menene ake nufi da gwagwarmaya"}
]
text = format_text_for_inference(q_chat)
input_tokens = tokenizer(text,return_tensors='pt',add_special_tokens=False).to('cuda')
outputs = model.generate(
**input_tokens,
max_new_tokens = 1000,
use_cache=True,
repetition_penalty=1.12,
temperature = 0.1
)
text = tokenizer.batch_decode(outputs)[0]
print(text)
<|begin_of_text|><|start_header_id|>system<|end_header_id|>
Cutting Knowledge Date: December 2023
Today Date: 11 Jun 2025
your name is AwaGPT, you are a large language model trained by Awarri AI technologies. You are a friendly assistant and you are here to help.<|eot_id|><|start_header_id|>user<|end_header_id|>
menene ake nufi da gwagwarmaya<|eot_id|><|start_header_id|>assistant<|end_header_id|>
A cikin mahallin ilimi, "gwagwarmaya" na nufin tsarin koyon halaye ko kwarewa ta hanyar maimaitawa, aiki, da kuma fuskantar kalubale. Yana jaddada rawar da kokari mai yawa ke takawa wajen samun nasara da ci gaba. Gwagwarmayar ba kawai game da samun sakamako ba ne amma har ma game da abubuwan da aka koya a lokacin aikin, wanda ke taimakawa wajen inganta kwarewa da tabbatar da canje-canjen da aka samu.
Ga karin bayani kan ra'ayoyin da suka shafi gwagwarmaya:
1. **Koyo Ta Hanyar Kwarewa**: Yana mai da hankali kan muhimmancin shiga cikin yanayi na zahiri don samun ilimi da fasaha. Maimakon kawai karantawa ko sauraron darasi, gwagwarmaya yana haɗa da aiwatar da ayyuka da warware matsaloli, wanda ke haifar da zurfin fahimta.
2. **Matsakaicin Kalubale**: Wannan hanya yawanci tana haɗa da fara da burin da ya wuce iyawarka na yanzu (matsakaici) sannan ka yi aiki don cimma wannan burin. Ta wannan hanyar, kana koyon iyakokin ka da wuraren da za a inganta, wanda ke haifar da ci gaban mutum da kuma ƙarfafawa.
3. **Dorewa**: Ingantaccen koyo ta hanyar gwagwarmaya na iya zama dindindin idan an sake fuskantar kalubalen a tsawon lokaci. Ba kamar koyo na ɗan lokaci ba, inda ilimin zai iya zama ajiye ba tare da aiki ba, gwagwarmaya tana taimakawa wajen riƙe ilimi ta hanyar ci gaba da bukatar amfani da shi.
4. **Halin Juriya**: Gwagwarmaya yawanci tana buƙatar jure gazawa da rashin nasara. Ta hanyar fuskantar wahala akai-akai, mutane suna haɓaka juriya da ƙudurin warware matsaloli, waɗannan halaye masu mahimmanci ga nasara a dogon lokaci.
5. **Haɓaka Kai**: Gwagwarmaya ana amfani da ita sosai a cikin horon kai don taimakawa mutane su shawo kan tsoro, gina kwarin gwiwa, da haɓaka ikon sarrafa kansu. Yana haɓaka tunani mai kyau da kuma motsa mutane su tura iyakokinsu.
6. **Amfani a Fannonin Daban-daban**: Ana amfani da manufar gwagwarmaya ba kawai a fannin ilimi ba; ana amfani da ita a fannonin kamar wasanni, horon sana'a, da ci gaban mutum. Misali, dan wasa na iya amfani da gwagwarmaya don inganta dabaru ko kwarewa, yayin da mai sana'a zai iya amfani da ita don koyo sabbin fasahohi ko dabaru.
A taƙaice, gwagwarmaya wata hanya ce mai tasiri ta koyo da ci gaba wacce ke jaddada mahimmancin aiki, juriya, da ci gaba mai dorewa. Yana taimakawa mutane su sami ilimi da kwarewa da za su iya amfani da su a rayuwa ta zahiri.<|eot_id|>
This model is designed for:
For issues, questions, or collaboration opportunities, please refer to the model repository discussions or contact Awarri Technologies.
This work was made possible through:
@misc{awagptv1_2025,
title={N-ATLaS-LLM: A Multilingual African Language Model},
author={Awarri Technologies and National Information Technology and Development Agency},
year={2025},
publisher={Hugging Face},
note={Fine-tuned Llama-3 8B model for African languages developed in collaboration with the Federal Government of Nigeria}
}
(Nigeria – Automatic Transcription and Language Systems)
Effective Date: September 2025
Version: 1.0
Awarri Technologies, in partnership with the Federal Government of Nigeria, hereby releases N-ATLaS (Nigeria – Automatic Transcription and Language Systems), consisting of four Automatic Speech Recognition (ASR) models and one Text Large Language Model (LLM) for Nigerian languages (Yoruba, Hausa, Igbo, and Nigerian-accented English).
N-ATLaS is released under an Open-Source Research and Innovation License inspired by permissive licenses such as Apache 2.0 and MIT, but with additional restrictions tailored for responsible use in Nigeria and globally.
The models are intended to support:
⚠️ N-ATLaS is not an enterprise-grade or commercial system. Commercial or large-scale enterprise use requires a separate licensing agreement (see Section 3).
Subject to compliance with these Terms, users are hereby granted a worldwide, royalty-free, non-exclusive, non-transferable license to:
Conditions:
Attribution must be given to:
“Awarri Technologies and the Federal Ministry of Communications, Innovation and Digital Economy .”
Derivative works must be released under the same license, ensuring consistency and traceability.
If N-ATLaS or its derivatives are renamed, they must carry the suffix: “Powered by Awarri.”
Use of N-ATLaS is limited to organizations, institutions, or projects with no more than 1000 active end-users.
Known limitations include:
Neither Awarri Technologies nor the Federal Ministry of Communications, Innovation and Digital Economy shall be liable for damages arising from the use of N-ATLaS.
Users must:
The Federal Government of Nigeria and Awarri Technologies reserve the right to revoke, suspend, or terminate usage rights if these Terms are violated.
Termination may apply to individual users, institutions, or organizations found in breach.
For licensing, inquiries, and commercial partnerships regarding N-ATLaS, contact:
Awarri Technologies
Federal Ministry of Communications, Innovation, and Digital Economy
Required attribution in all public use:
“N-ATLaS is an initiative of the Federal Ministry of Communications, Innovation and Digital Economy, and powered by Awarri Technologies.”
If renamed, the model must carry the suffix:
“Powered by Awarri.”
N-ATLaS-LLM is part of Awarri Technologies' mission, initiated by the The Federal Ministry of Communications, Innovation and Digital Economy , to make AI accessible to African language speakers and preserve linguistic diversity in the digital age.