Downloads · 30 days
7
3% of all-time downloads
K024/chatglm2-6b-int8
chatglm2-6b-int8 is a machine learning model from K024. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for transformers.
Downloads · 30 days
7
3% of all-time downloads
All-time downloads
264
Public
Repo size
6.3 GB
Likes
1
Public
Click a slice to open those files.
.safetensors6.2 GB · 100%
From the Hugging Face model README
详情参考 K024/chatglm-q。
See K024/chatglm-q for more details.
import torch
from chatglm_q.decoder import ChatGLMDecoder, chat_template
device = torch.device("cuda")
decoder = ChatGLMDecoder.from_pretrained("K024/chatglm2-6b-int8", device=device)
prompt = chat_template([], "我是谁?")
for text in decoder.generate(prompt):
print(text)
模型权重按 ChatGLM2-6b 许可发布,见 MODEL LICENSE。
Model weights are released under the same license as ChatGLM2-6b, see MODEL LICENSE.