Downloads · 30 days
118
9% of all-time downloads
efficiencyx/Jun-Lora-v2-GGUF
Jun-Lora-v2-GGUF is a text generation model from efficiencyx. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
A LoRA fine-tune of Gemma 4 12B trained on syntetic multi-turn conversational data from the visual novel My Dystopian Robot Girlfriend. The model captures the personality, speech patterns, and emotional nuance of the…
Downloads · 30 days
118
9% of all-time downloads
All-time downloads
1.3K
Public
Repo size
30.2 GB
Likes
4
Public
Click a slice to open those files.
.gguf29.8 GB · 100%
From the Hugging Face model README
A LoRA fine-tune of Gemma 4 12B trained on syntetic multi-turn conversational data from the visual novel My Dystopian Robot Girlfriend. The model captures the personality, speech patterns, and emotional nuance of the character Jun while preserving the base model's general reasoning and instruction-following capabilities.
| Repository | Format | Description |
|---|---|---|
efficiencyx/Jun-Lora-v2-SAFETENSOR | SafeTensors FP16 | Full-precision merged model |
efficiencyx/Jun-Lora-v2-GGUF | GGUF Q8_0 / Q6_K / Q4_K_M | Quantized versions for local inference |
efficiencyx/Jun-Lora-v2 | LoRA Adapter | Raw adapters at checkpoints 138, 120, 90 |
| Quant | Size (approx.) | Use Case |
|---|---|---|
| Q8_0 | ~12.8 GB | Best quality, suggested ~16 GB VRAM |
| Q6_K | ~10.4 GB | Recommended balance of quality and performance |
| Q4_K_M | ~7.6 GB | Fits on 8 GB VRAM GPUs with acceptable quality loss |
This model is designed as the conversational backend for Jun OS, an AI companion webapp. It is intended for:
| Property | Value |
|---|---|
| Source | My Dystopian Robot Girlfriend (visual novel dialogue) |
| Composition | ~1:1 replica of original game tone and cadence |
| Size | 2,302 multi-turn conversations |
| Format | ChatML (`< |
The dataset was constructed to preserve the character's tone, vocabulary, emotional range, and conversational patterns across a variety of in-game scenarios. Multi-turn structure ensures the model learns contextual consistency over extended exchanges.
| Parameter | Value |
|---|---|
| Base model | google/gemma-4-12b-it |
| Method | LoRA |
| LoRA rank | 64 |
| LoRA alpha | 128 |
| Learning rate | 2e-5 |
| Batch size | 8 |
| Gradient accumulation steps | 4 |
| Effective batch size | 32 |
| Epochs | 2 |
| Total steps | 138 |
| Checkpoint interval | Every 30 steps |
| Optimizer | AdamW (8-bit) |
| Component | Detail |
|---|---|
| Training GPU | NVIDIA A100 80GB SXM4 |
| Fine-tuning framework | Unsloth |
| GGUF export pipeline | llama.cpp |
| Metric | Value |
|---|---|
| Final training loss | ~1.21 |
| Final eval loss | ~1.24 |
The narrow gap between training and eval loss indicates the model generalizes well without significant overfitting, despite the relatively small dataset size.
Multiple adapter checkpoints are provided (steps 90, 120, 138) to allow users to select the best trade-off between character adherence and generalization for their use case. Earlier checkpoints may exhibit slightly more creative freedom, while the final checkpoint (138) has the strongest character lock-in.