Downloads · 30 days
52
0% of all-time downloads
TheBloke/Planner-7B-fp16
Planner-7B-fp16 is a text generation model from TheBloke. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
<div style="width: 100%;" <img src="https://i.imgur.com/EBdldam.jpg" alt="TheBlokeAI" style="width: 100%; min-width: 400px; display: block; margin: auto;" </div <div style="display: flex; justify-content: space-betwee…
Downloads · 30 days
52
0% of all-time downloads
All-time downloads
56.2K
Public
Repo size
27 GB
Likes
1
Public
Click a slice to open those files.
.bin13.5 GB · 100%
From the Hugging Face model README
These files are fp16 pytorch format model files for rewoo's Planner 7B.
They are the result of merging the LoRA adapter at the above repo with the base LLaMa 7B model.
For further support, and discussions on these models and AI in general, join us at:
Thanks to the chirper.ai team!
I've had a lot of people ask if they can contribute. I enjoy providing models and helping people, and would love to be able to spend even more time doing it, as well as expanding into new projects like fine tuning/training.
If you're able and willing to contribute it will be most gratefully received and will help me to keep providing more models, and to start work on new AI projects.
Donaters will get priority support on any and all AI/LLM/model questions and requests, access to a private Discord room, plus other benefits.
Special thanks to: Luke from CarbonQuill, Aemon Algiz, Dmitriy Samsonov.
Patreon special mentions: Derek Yates, Sean Connelly, Luke, Nathan LeClaire, Trenton Dambrowitz, Mano Prime, David Flickinger, vamX, Nikolai Manek, senxiiz, Khalefa Al-Ahmad, Illia Dulskyi, trip7s trip, Jonathan Leane, Talal Aujan, Artur Olbinski, Cory Kujawski, Joseph William Delisle, Pyrater, Oscar Rangel, Lone Striker, Luke Pendergrass, Eugene Pentland, Johann-Peter Hartmann.
Thank you to all my generous patrons and donaters!
<!-- footer end -->Alpaca Lora adapter weight fine-tuned on following instruction dataset.
https://huggingface.co/datasets/rewoo/planner_instruction_tuning_2k/blob/main/README.md
Training script: borrowed from the official Alpaca-LoRA implementation
We use following parameter.
python finetune.py \
--base_model 'decapoda-research/llama-7b-hf' \
--data_path 'rewoo/planner_instruction_tuning_2k' \
--output_dir './lora-alpaca-planner' \
--batch_size 128 \
--micro_batch_size 8 \
--num_epochs 10 \
--learning_rate 1e-4 \
--cutoff_len 1024 \
--val_set_size 200 \
--lora_r 8 \
--lora_alpha 16 \
--lora_dropout 0.05 \
--lora_target_modules '[q_proj,v_proj]' \
--train_on_inputs \
--group_by_length \
--resume_from_checkpoint 'tloen/alpaca-lora-7b'