Downloads · 30 days
13
32% of all-time downloads
seawolf2357/LLAMA-3-70B-Ko
LLAMA-3-70B-Ko is a text generation model from seawolf2357. Use it when you need the model to write or continue text. It is set up for transformers.
This is a merge of pre-trained language models created using mergekit.
Downloads · 30 days
13
32% of all-time downloads
All-time downloads
41
Public
Parameters
29.5B
59 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors59 GB · 100%
From the Hugging Face model README
This is a merge of pre-trained language models created using mergekit.
This model was merged using the SLERP merge method.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
slices:
- sources:
- model: Bllossom/llama-3-Korean-Bllossom-70B
layer_range:
- 0
- 32
- model: tenyx/Llama3-TenyxChat-70B
layer_range:
- 0
- 32
merge_method: slerp
base_model: Bllossom/llama-3-Korean-Bllossom-70B
parameters:
t:
- filter: self_attn
value:
- 0
- 0.5
- 0.3
- 0.7
- 1
- filter: mlp
value:
- 1
- 0.5
- 0.7
- 0.3
- 0
- value: 0.5
dtype: bfloat16