Downloads · 30 days
23
12% of all-time downloads
limajr/nbr-1b-base
nbr-1b-base is a text generation model from limajr. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
NBR-1B is a 1 billion parameter language model trained from scratch specifically for Brazilian Portuguese.
Downloads · 30 days
23
12% of all-time downloads
All-time downloads
194
Public
Parameters
1.3B
5.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.1 GB · 100%
From the Hugging Face model README
NBR-1B is a 1 billion parameter language model trained from scratch specifically for Brazilian Portuguese.
| Parameter | Value |
|---|---|
| Hidden Size | 2048 |
| Layers | 24 |
| Attention Heads | 16 |
| KV Heads | 4 (GQA) |
| Intermediate Size | 5504 |
| Vocab Size | 32000 |
Curated Portuguese corpus (~25B tokens):
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("limajr/nbr-1b-base")
tokenizer = AutoTokenizer.from_pretrained("limajr/nbr-1b-base")
text = "O Brasil e um pais"
inputs = tokenizer(text, return_tensors="pt")
outputs = model.generate(**inputs, max_new_tokens=100)
print(tokenizer.decode(outputs[0]))
Apache 2.0