Downloads · 30 days
11
17% of all-time downloads
Subas07/llama_finetune_16bit
llama_finetune_16bit is a text generation model from Subas07. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
llamalora is a fine-tuned Large Language Model based on a LLaMA architecture using LoRA (Low-Rank Adaptation). The model is optimized for instruction-following and conversational text generation tasks.
Downloads · 30 days
11
17% of all-time downloads
All-time downloads
66
Public
Parameters
8B
16.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.1 GB · 100%
From the Hugging Face model README
llama_lora is a fine-tuned Large Language Model based on a LLaMA architecture using LoRA (Low-Rank Adaptation).
The model is optimized for instruction-following and conversational text generation tasks.
This model is built on top of a LLaMA base model (8B parameter class).
This model can be used for:
from transformers import AutoTokenizer, AutoModelForCausalLM
import torch
model_name = "your-username/llama_lora"
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(
model_name,
torch_dtype=torch.float16,
device_map="auto"
)
prompt = "Explain machine learning in simple words."
inputs = tokenizer(prompt, return_tensors="pt").to(model.device)
outputs = model.generate(
**inputs,
max_new_tokens=150,
temperature=0.7,
top_p=0.9
)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))