Downloads · 30 days
35
8% of all-time downloads
prem-research/CodeLlama-34b-Instruct-hf
CodeLlama-34b-Instruct-hf is a text generation model from prem-research. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama2.
Resharded Code Llama repository optimized for Petals inference. Instead of having 7 shards of ~10GiB each, the current repository has 49 shards each one of ~1.5GiB.
Downloads · 30 days
35
8% of all-time downloads
All-time downloads
436
Public
Parameters
33.7B
67.5 GB on disk
Likes
3
Public
Click a slice to open those files.
.safetensors67.5 GB · 100%
From the Hugging Face model README
Resharded Code Llama repository optimized for Petals inference. Instead of having 7 shards of ~10GiB each, the current repository has 49 shards each one of ~1.5GiB.
For more information about the model, you can check the official card here.
# pip install git+https://github.com/huggingface/transformers.git@main accelerate
from transformers import LlamaTokenizer, AutoModelForCausalLM
tokenizer = LlamaTokenizer.from_pretrained("premai-io/CodeLlama-34b-Instruct-hf")
model = AutoModelForCausalLM.from_pretrained("premai-io/CodeLlama-34b-Instruct-hf")
inputs = tokenizer("def hello_world():", return_tensors="pt")["input_ids"]
outputs = model.generate(inputs, max_new_tokens=100)
print(tokenizer.decode(outputs[0]))
Code Llama and its variants are a new technology that carries risks with use. Testing conducted to date has been in English, and has not covered, nor could it cover all scenarios. For these reasons, as with all LLMs, Code Llama’s potential outputs cannot be predicted in advance, and the model may in some instances produce inaccurate or objectionable responses to user prompts. Therefore, before deploying any applications of Code Llama, developers should perform safety testing and tuning tailored to their specific applications of the model.
Please see the Responsible Use Guide available available at https://ai.meta.com/llama/responsible-user-guide.