Downloads · 30 days
7
58% of all-time downloads
USato/typhoon-iot-lora-instruct
typhoon-iot-lora-instruct is a text generation model from USato. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as llama3.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
7
58% of all-time downloads
All-time downloads
12
Public
Repo size
1 GB
Likes
0
Public
Click a slice to open those files.
.safetensors1 GB · 72%
From the Hugging Face model README
axolotl version: 0.12.2
adapter: qlora
base_model: scb10x/llama-3-typhoon-v1.5-8b-instruct
bf16: true
chat_template: llama3
datasets:
- ds_type: json
field_messages: messages
message_property_mappings:
content: content
role: role
path: iot_train_chat.json
split: train
type: chat_template
flash_attention: true
fp16: false
gradient_accumulation_steps: 1
gradient_checkpointing: true
learning_rate: 0.0001
load_in_4bit: true
logging_steps: 1
lora_alpha: 64
lora_dropout: 0.5
lora_r: 32
lora_target_modules:
- q_proj
- k_proj
- v_proj
- o_proj
- gate_proj
- up_proj
- down_proj
micro_batch_size: 8
model_type: LlamaForCausalLM
num_epochs: 1
optimizer: paged_adamw_8bit
output_dir: ./outputs/typhoon-iot-chat-format
pad_to_sequence_len: true
sample_packing: true
save_steps: 50
save_strategy: steps
sequence_len: 4096
special_tokens:
bos_token: <|begin_of_text|>
eos_token: <|end_of_text|>
pad_token: <|end_of_text|>
tokenizer_type: AutoTokenizer
warmup_steps: 10
xformers_attention: false
</details><br>
This model is a fine-tuned version of scb10x/llama-3-typhoon-v1.5-8b-instruct on the iot_train_chat.json dataset.
More information needed
More information needed
More information needed
The following hyperparameters were used during training: