Downloads · 30 days
18
8% of all-time downloads
killer66678/openpangu_7b_lora
openpangu_7b_lora is a text generation model from killer66678. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as other.
This repository contains LoRA-finetuned and merged weights based on openPangu-Embedded-7B-V1.1. The LoRA adapters were merged into the base model to produce full weights suitable for standard inference.
Downloads · 30 days
18
8% of all-time downloads
All-time downloads
221
Public
Parameters
8B
16.1 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.1 GB · 100%
From the Hugging Face model README
This repository contains LoRA-finetuned and merged weights based on
openPangu-Embedded-7B-V1.1. The LoRA adapters were merged into the
base model to produce full weights suitable for standard inference.
FreedomIntelligence/openPangu-Embedded-7B-V1.1OPENPANGU Model License Agreement v1.0 (see LICENSE)llamafactory-cli export.Example (paths are placeholders):
llamafactory-cli export \
--model_name_or_path <base_model_dir> \
--adapter_name_or_path <lora_adapter_dir> \
--template default \
--finetuning_type lora \
--export_dir <export_dir> \
--export_size 2 \
--export_device cpu \
--export_legacy_format False \
--trust_remote_code True
Evaluated with lm-evaluation-harness using vLLM on 4x RTX 4090.
Dates (UTC): 2026-01-04.
Example command (paths are placeholders):
lm_eval --model vllm \
--model_args "pretrained=<model_dir>,tensor_parallel_size=4,dtype=auto,gpu_memory_utilization=0.8,max_model_len=4096,enforce_eager=True,trust_remote_code=True" \
--tasks gsm8k \
--num_fewshot 5 \
--batch_size auto
This repo includes custom modeling code; trust_remote_code=True is required.
from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "killer66678/openpangu_7b_lora"
tokenizer = AutoTokenizer.from_pretrained(model_id, trust_remote_code=True)
model = AutoModelForCausalLM.from_pretrained(
model_id,
trust_remote_code=True,
torch_dtype="auto",
device_map="auto",
)
本研究的实验与计算工作依托于华为云昇腾AI云服务平台完成,特此对其提供的稳定算力支持表示感谢。