Downloads · 30 days
0
sahith2004/llama-version-1
llama-version-1 is a text generation model from sahith2004. Use it when you need the model to write or continue text. The card lists the license as other.
This model was trained using AutoTrain. For more information, please visit AutoTrain.
Downloads · 30 days
0
Access
Public
Updated Mar 16, 2024
Repo size
481 MB
Likes
0
Public
Click a slice to open those files.
.pt320 MB · 50%
From the Hugging Face model README
This model was trained using AutoTrain. For more information, please visit AutoTrain.
from transformers import AutoModelForCausalLM, AutoTokenizer
model_path = "PATH_TO_THIS_REPO"
tokenizer = AutoTokenizer.from_pretrained(model_path)
model = AutoModelForCausalLM.from_pretrained(
model_path,
device_map="auto",
torch_dtype='auto'
).eval()
# Prompt content: "hi"
messages = [
{"role": "user", "content": "hi"}
]
input_ids = tokenizer.apply_chat_template(conversation=messages, tokenize=True, add_generation_prompt=True, return_tensors='pt')
output_ids = model.generate(input_ids.to('cuda'))
response = tokenizer.decode(output_ids[0][input_ids.shape[1]:], skip_special_tokens=True)
# Model response: "Hello! How can I assist you today?"
print(response)