Downloads · 30 days
147
1% of all-time downloads
stockmark/stockmark-13b-instruct
stockmark-13b-instruct is a text generation model from stockmark. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
Stockmark-13b-instruct is an instruction-tuned version of Stockmark-13b, a 13 billion parameter Japanese LLM. This model is developed by Stockmark Inc.
Downloads · 30 days
147
1% of all-time downloads
All-time downloads
19K
Public
Parameters
13.2B
26.4 GB on disk
Likes
10
Public
Click a slice to open those files.
.safetensors26.4 GB · 100%
From the Hugging Face model README
Stockmark-13b-instruct is an instruction-tuned version of Stockmark-13b, a 13 billion parameter Japanese LLM. This model is developed by Stockmark Inc.
We used data (2023/11/03 version) from Project of Development of Japanese Instruction data for LLM for instruction tuning.
Please see our blog for more details.
import torch
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("stockmark/stockmark-13b-instruct", device_map="auto", torch_dtype=torch.bfloat16)
tokenizer = AutoTokenizer.from_pretrained("stockmark/stockmark-13b-instruct")
instruction = "自然言語処理とは?"
prompt = f"""### Input:
{instruction}
### Output:
"""
inputs = tokenizer(prompt, return_tensors="pt").to(model.device)
with torch.no_grad():
tokens = model.generate(
**inputs,
max_new_tokens=128,
do_sample=True,
temperature=0.7
)
output = tokenizer.decode(tokens[0], skip_special_tokens=True)
print(output)
Project of Development of Japanese Instruction data for LLM
MIT