Downloads · 30 days
153
1% of all-time downloads
xiaodongguaAIGC/llama-3-debug
llama-3-debug is a text generation model from xiaodongguaAIGC. Use it when you need the model to write or continue text. It is set up for transformers.
This model use for debug, the parameter is random.
Downloads · 30 days
153
1% of all-time downloads
All-time downloads
12.4K
Public
Parameters
16.5M
98.8 MB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors32.9 MB · 78%
From the Hugging Face model README
This model use for debug, the parameter is random.
It's small only '~32MB' memory size, that is efficent for you to download and debug.
llama-3-debug model config modified as follow
config.intermediate_size = 128
config.hidden_size = 64
config.num_attention_heads = 2
config.num_key_value_heads = 2
config.num_hidden_layers = 1
If you want to load it by this code
from transformers import AutoModelForCausalLM, AutoTokenizer
model_name = 'xiaodongguaAIGC/llama-3-debug'
model = AutoModelForCausalLM.from_pretrained(model_name, torch_dtype=torch.bfloat16)
tokenizer = AutoTokenizer.from_pretrained(model_name)
print(model)
print(tokenizer)