Downloads · 30 days
92
7% of all-time downloads
Neutralzz/BiLLa-7B-SFT
BiLLa-7B-SFT is a machine learning model from Neutralzz. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
BiLLa is an open-source reasoning-enhanced bilingual LLaMA model. The main features are: - Greatly improve the ability of Chinese language modeling, and minimize the damage to the original English ability of LLaMA; -…
Downloads · 30 days
92
7% of all-time downloads
All-time downloads
1.3K
Public
Repo size
14.1 GB
Likes
65
Public
Click a slice to open those files.
.bin14.1 GB · 100%
From the Hugging Face model README
BiLLa is an open-source reasoning-enhanced bilingual LLaMA model. The main features are:
Github: https://github.com/Neutralzz/BiLLa
<b>Note</b>: Due to LLaMA's license, the model weights in this hub cannot be used directly.
The weight of word embedding is the sum of the weights of the trained model and the original LLaMA,
so as to ensure that developers with LLaMA original model accessibility can convert the model released by this hub into a usable one.
First, you can revert the model weights by this script:
python3 embedding_convert.py \
--model_dir /path_to_BiLLa/BiLLa-7B-SFT \
--meta_llama_pth_file /path_to_LLaMA/llama-7b/consolidated.00.pth
Then, you can run this model as follows:
import torch
from transformers import AutoTokenizer, AutoModelForCausalLM
model_path = "/path_to_BiLLa/BiLLa-7B-SFT"
tokenizer = AutoTokenizer.from_pretrained(model_path, use_fast=False)
model = AutoModelForCausalLM.from_pretrained(model_path, low_cpu_mem_usage=True, torch_dtype=torch.float16).cuda()
prompt = "Human: Write a Python function that checks if a given number is even or odd.\nAssistant: "
input_ids = tokenizer([prompt]).input_ids
output_ids = model.generate(
torch.as_tensor(input_ids).cuda(),
do_sample=True,
temperature=0.7,
max_new_tokens=1024
)
output_ids = output_ids[0][len(input_ids[0]):]
outputs = tokenizer.decode(output_ids, skip_special_tokens=True).strip()
print(outputs)
Different from BiLLa-7B-LLM, the model input of BiLLa-7B-SFT should be formatted as follows:
Human: [Your question]
Assistant:
Note that <b>a space</b> is following the Assistant: