Downloads · 30 days
163
0% of all-time downloads
axiong/PMC_LLaMA_13B
PMC_LLaMA_13B is a text generation model from axiong. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as openrail.
To obtain the foundation model in medical field, we propose MedLLaMA13B and PMCLLaMA13B.
Downloads · 30 days
163
0% of all-time downloads
All-time downloads
42.7K
Public
Repo size
111 GB
Likes
32
Public
Click a slice to open those files.
.bin52.1 GB · 100%
From the Hugging Face model README
To obtain the foundation model in medical field, we propose MedLLaMA_13B and PMC_LLaMA_13B.
MedLLaMA_13B is initialized from LLaMA-13B and further pretrained with medical corpus. Despite the expert knowledge gained, it lacks instruction-following ability. Hereby we construct a instruction-tuning dataset and evaluate the tuned model.
As shown in the table, PMC_LLaMA_13B achieves comparable results to ChatGPT on medical QA benchmarks.

import transformers
import torch
tokenizer = transformers.LlamaTokenizer.from_pretrained('axiong/PMC_LLaMA_13B')
model = transformers.LlamaForCausalLM.from_pretrained('axiong/PMC_LLaMA_13B')
sentence = 'Hello, doctor'
batch = tokenizer(
sentence,
return_tensors="pt",
add_special_tokens=False
)
with torch.no_grad():
generated = model.generate(
inputs = batch["input_ids"],
max_length=200,
do_sample=True,
top_k=50
)
print('model predict: ',tokenizer.decode(generated[0]))