Downloads · 30 days
78
100% of all-time downloads
kungaa/TuvanGemma
TuvanGemma is a translation model from kungaa. Use it when you need text moved from one language to another. The card lists the license as other.
TuvanGemma is a family of local Russian ↔ Tuvan translation models. The recommended release fine-tunes google/translategemma-12b-it and is provided as a llama.cpp GGUF, a LoRA adapter, and a self-contained Windows app…
Downloads · 30 days
78
100% of all-time downloads
All-time downloads
78
Public
Repo size
50 GB
Likes
0
Public
Click a slice to open those files.
.zip13.1 GB · 55%
From the Hugging Face model README
TuvanGemma is a family of local Russian ↔ Tuvan translation models. The
recommended release fine-tunes google/translategemma-12b-it and is provided
as a llama.cpp GGUF, a LoRA adapter, and a self-contained Windows application.
The models are Gemma-derived. Read and comply with the applicable Gemma Terms of Use before using or redistributing them.
| Release | Download | Notes |
|---|---|---|
| 12B model (recommended) | tuvangemma-12b-q4_k_m.gguf | Q4_K_M, 7.31 GB; llama.cpp |
| 12B Windows app | TuvanTranslator-windows-x64-v0.2.1-12b.zip | Portable; bundles the model and CUDA, Vulkan, and CPU runtimes |
| 12B LoRA adapter | translategemma-12b-qlora-r16 | Use with google/translategemma-12b-it |
| 4B model (smaller legacy release) | tuvangemma-4b-q5_k_m.gguf | Q5_K_M, 2.84 GB |
| 4B Windows app | TuvanTranslator-windows-x64-v0.1.0.zip | Older, lower-memory portable app |
| 4B LoRA adapter | gemma4-e4b-qlora | Adapter used for the legacy release |
The v0.2.1 desktop app adds a Lower memory (2K context) switch. Leave it off for the normal 4,096-token context; turn it on before loading the model on an 8 GB GPU. Changing it restarts the local inference runtime automatically.
The 12B adapter was evaluated on 256 held-out examples in each direction (512 translations total):
| Direction | Examples | BLEU | chrF++ |
|---|---|---|---|
| Russian → Tuvan | 256 | 20.4171 | 49.1626 |
| Tuvan → Russian | 256 | 25.9565 | 49.3654 |
The evaluation report describes the
procedure. Raw predictions are also included under reports/.
These are automatic metrics on one held-out split, not a guarantee of translation quality. Review names, numbers, idioms, and important text with a qualified Tuvan speaker.
Direct llama.cpp Vulkan measurements with full layer offload used 7,853 MiB at 4,096 tokens and 8,113 MiB at 8,192 tokens. In practice:
See the VRAM report for the measured configuration and caveats.
The earlier 4B release was evaluated separately on 3,824 held-out examples in each direction. Its split and setup differ from the 12B evaluation, so the tables should not be compared as a controlled model-size experiment.
| Direction | Baseline chrF++ | TuvanGemma chrF++ | Baseline BLEU | TuvanGemma BLEU |
|---|---|---|---|---|
| Russian → Tuvan | 13.0657 | 45.4822 | 1.1298 | 16.2996 |
| Tuvan → Russian | 19.6637 | 44.7172 | 2.1083 | 20.8963 |
Fine-tuning used Agisight's
tyv-rus-200k
dataset, attributed under CC BY 4.0. The desktop package includes the dataset
attribution notice.
Source code, build instructions, third-party notices, and the complete Gemma
terms are available in
kungaa/TuvanTranslator.
The Windows ZIP is self-contained and uses llama.cpp. Auto mode tries CUDA,
then Vulkan, then CPU. Vulkan requires a working Vulkan-capable graphics
driver; vulkan-1.dll is intentionally supplied by the driver rather than
bundled. A Metal backend and macOS packaging recipe are present in the source,
but a signed/notarized macOS binary is not currently published.