Downloads · 30 days
227
7% of all-time downloads
demetera/llama-600M-rus
llama-600M-rus is a text generation model from demetera. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
Simple and customized amateur experimental model pretrained on the text fiction books from the scratch (updating the model regularly).<br It could generate amateur, but more or less adequate output as well (in respect…
Downloads · 30 days
227
7% of all-time downloads
All-time downloads
3.5K
Public
Parameters
548M
6.6 GB on disk
Likes
3
Public
Click a slice to open those files.
.safetensors2.2 GB · 100%
From the Hugging Face model README
Simple and customized amateur experimental model pretrained on the text fiction books from the scratch (updating the model regularly).<br> It could generate amateur, but more or less adequate output as well (in respect of training tokens).<br> The work can be used as a checkpoint for the further training or for experiments.<br>
Simple usage example:
from transformers import LlamaTokenizerFast, LlamaForCausalLM
model = LlamaForCausalLM.from_pretrained('demetera/llama-600M-rus')
tokenizer = LlamaTokenizerFast.from_pretrained('demetera/llama-600M-rus')
prompt = "Я вышел и улицу и"
inputs = tokenizer(prompt, return_tensors='pt')
outputs = model.generate(inputs.input_ids, attention_mask = inputs.attention_mask, max_new_tokens=250, do_sample=True, top_k=50, top_p=0.95)
print (tokenizer.decode(outputs[0], skip_special_tokens=True))