Downloads · 30 days
0
riverallzero/alpaca-lora-7b
alpaca-lora-7b is a machine learning model from riverallzero. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
This repo contains a low-rank adapter for LLaMA-7b fit on the Stanford Alpaca dataset.
Downloads · 30 days
0
Access
Public
Updated Feb 17, 2024
Repo size
67.2 MB
Likes
1
Public
Click a slice to open those files.
.bin67.2 MB · 100%
From the Hugging Face model README
This repo contains a low-rank adapter for LLaMA-7b fit on the Stanford Alpaca dataset.
This version of the weights was trained with the following hyperparameters:
That is:
python finetune.py \
--base_model='baffo32/decapoda-research-llama-7B-hf' \
--num_epochs=10 \
--cutoff_len=512 \
--group_by_length \
--output_dir='./lora-alpaca-512-qkvo' \
--lora_target_modules='[q_proj,k_proj,v_proj,o_proj]' \
--lora_r=16 \
--micro_batch_size=8