Downloads Β· 30 days
10
14% of all-time downloads
soka0000/vclm-KoCoder-7B
vclm-KoCoder-7B is a machine learning model from soka0000. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Downloads Β· 30 days
10
14% of all-time downloads
All-time downloads
70
Public
Parameters
7.6B
15.2 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors15.2 GB Β· 100%
From the Hugging Face model README
Korean Code Generation Model based on Qwen2.5-7B-Instruct
</div>VCLM KoCoder 7B is a specialized code generation model fine-tuned from soka0000/vclm-korean-7b using QLoRA on 40,000 high-quality code instruction datasets.
pip install transformers torch accelerate
from transformers import AutoModelForCausalLM, AutoTokenizer
import torch
model_name = "Soka0000/vclm-KoCoder-7B"
# Load model and tokenizer
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(
model_name,
torch_dtype=torch.bfloat16,
device_map="auto"
)
# Generate code
messages = [
{"role": "system", "content": "You are SokaAI, created by Soka0000 Korea. You are a helpful AI Assistant."},
{"role": "user", "content": "Write a Python function to implement quicksort algorithm."}
]
text = tokenizer.apply_chat_template(
messages,
tokenize=False,
add_generation_prompt=True
)
inputs = tokenizer(text, return_tensors="pt").to(model.device)
outputs = model.generate(
**inputs,
max_new_tokens=1024,
temperature=0.7,
top_p=0.9,
do_sample=True
)
response = tokenizer.decode(outputs[0], skip_special_tokens=True)
print(response)
messages = [
{"role": "system", "content": "You are SokaAI, created by Soka0000 Korea. You are a helpful AI Assistant."},
{"role": "user", "content": "μ΄μ§ νμ νΈλ¦¬λ₯Ό ꡬννλ νμ΄μ¬ ν΄λμ€λ₯Ό μμ±ν΄μ€."}
]
# ... (same generation code as above)
| Metric | Value |
|---|---|
| Final Loss | 0.554 |
| Token Accuracy | 86.2% |
| Training Samples | 40,000 |
| Training Time | ~3.5 hours (H100 80GB) |
Binary Search Tree (880 tokens generated)
q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_projBase Model: soka0000/vclm-korean-7b
Fine-tuning: QLoRA
Epochs: 1
Batch Size: 2 (per device)
Gradient Accumulation: 8 (effective batch size: 16)
Learning Rate: 5e-5
Optimizer: paged_adamw_8bit
Scheduler: cosine with warmup (10%)
Max Length: 2048 tokens
Precision: bfloat16
Flash Attention: 2
| Dataset | Samples | Weight |
|---|---|---|
| CodeAlpaca-20k | 18,000 | 50% |
| Python Code Instructions | 12,000 | 35% |
| Code Instructions 122k | 10,000 | 15% |
| Total | 40,000 | 100% |
# Generate functions, classes, algorithms
"Write a Python function to calculate factorial using recursion."
# Explain code concepts
"Explain how binary search works with a C++ example."
# Help with code issues
"μ΄ μ½λλ₯Ό μ΅μ νν΄μ€: [code snippet]"
# Python, Java, C++, JavaScript, SQL, etc.
"Implement quicksort in Java."
This model inherits the license from its base model soka0000/vclm-korean-7b.
Important: The base model has a custom license (LICENSE_NC). Please check the base model's license terms before commercial use.
If you use this model, please cite:
@model{vclm-kocoder-7b,
title={VCLM KoCoder 7B: Korean Code Generation Model},
author={Soka0000},
year={2025},
publisher={HuggingFace},
url={https://huggingface.co/Soka0000/vclm-KoCoder-7B}
}
Made with β€οΈ by Soka0000 Korea
π€ Model β’ Documentation β’ Discussions
</div>