Downloads · 30 days
18
0% of all-time downloads
colable/llama-ko-peft
llama-ko-peft is a text generation model from colable. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
This is an Korean Model based on [beomi/open-llama-2-ko-7b]
Downloads · 30 days
18
0% of all-time downloads
All-time downloads
18.2K
Public
Parameters
6.9B
13.7 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors13.7 GB · 100%
From the Hugging Face model README
This is an Korean Model based on
gpu code example
import torch
from transformers import AutoTokenizer, AutoModelForCausalLM
import math
## v2 models
model_path = "colable/llama-ko-peft"
tokenizer = AutoTokenizer.from_pretrained(model_path, use_default_system_prompt=False)
model = AutoModelForCausalLM.from_pretrained(
model_path, torch_dtype=torch.float32, device_map='auto',local_files_only=False, load_in_4bit=True
)
print(model)
prompt = input("please input prompt:")
while len(prompt) > 0:
input_ids = tokenizer(prompt, return_tensors="pt").input_ids.to("cuda")
generation_output = model.generate(
input_ids=input_ids, max_new_tokens=500,repetition_penalty=1.2
)
print(tokenizer.decode(generation_output[0]))
prompt = input("please input prompt:")