Downloads · 30 days
19
24% of all-time downloads
vietrix/viena-60m
viena-60m is a text generation model from vietrix. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
- Developed by: Vietrix - Model type: decoder-only causal LM (Llama-style) - Parameters: ~60M - Layers: 16 - Hidden size: 512 - Attention heads: 8 (KV heads: 4) - Max sequence length: 1024 - RoPE theta: 10000 - Normal…
Downloads · 30 days
19
24% of all-time downloads
All-time downloads
79
Public
Parameters
66.7M
267 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors267 MB · 100%
From the Hugging Face model README
vietrix/viena-60m-pretrain.from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "vietrix/viena-60m"
tokenizer = AutoTokenizer.from_pretrained(model_id, use_fast=False)
model = AutoModelForCausalLM.from_pretrained(model_id)
If AutoTokenizer fails, load the SentencePiece model explicitly:
from transformers import LlamaTokenizer
tokenizer = LlamaTokenizer.from_pretrained(model_id, use_fast=False)