Downloads · 30 days
10
11% of all-time downloads
MaAIos/Henyo-153M
Henyo-153M is a text generation model from MaAIos. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
Henyo is a 153M parameter Tagalog Language Model trained on the MaAIos/culturax-filipino-subset dataset. It utilizes a custom efficient architecture heavily inspired by Llama 2/3 and PaLM.
Downloads · 30 days
10
11% of all-time downloads
All-time downloads
87
Public
Parameters
114M
913 MB on disk
Likes
0
Public
Click a slice to open those files.
.bin456 MB · 50%
From the Hugging Face model README
Henyo is a 153M parameter Tagalog Language Model trained on the MaAIos/culturax-filipino-subset dataset. It utilizes a custom efficient architecture heavily inspired by Llama 2/3 and PaLM.
This model uses a custom Decoder-Only Transformer architecture built from scratch in PyTorch.
| Hyperparameter | Value |
|---|---|
| Parameters | ~153M |
| Context Window | 1024 tokens |
| Embedding Dim | 768 |
| Layers (Depth) | 12 |
| Attention Heads | 12 |
| KV Heads (GQA) | 4 |
| Vocab Size | 50,257 (GPT-2 tokenizer) |
Since this model uses a custom architecture, you must include the class definitions (provided in the train_henyo.py file in this repo) or use the inference script below.
# See inference_henyo.py in files for full class definitions
from transformers import AutoTokenizer
model_id = "marcuscedricridia/Henyo-153M-CulturaX"
tokenizer = AutoTokenizer.from_pretrained(model_id)
# Load model using custom class wrapper...
The full training script (train_henyo.py) is included in the file listing of this repository.