Downloads · 30 days
23
13% of all-time downloads
TheZeez/gemma-4-e4b-creative-DFT-exp
gemma-4-e4b-creative-DFT-exp is a image-text-to-text model from TheZeez. Use it for the image-text-to-text task on the model card, and read the license before you ship it in a product. It is set up for transformers. The card lists the license as apache-2.0.
<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/
Downloads · 30 days
23
13% of all-time downloads
All-time downloads
177
Public
Parameters
8B
16 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors16 GB · 100%
From the Hugging Face model README
goated software^^
I am not a mathematician nor a professional coder. This was an experiment (with the help of AI of course).
This is a custom-trained version of Google's Gemma 4 E4B intended for creative writing. It was trained using a custom implementation of Distribution Fine-Tuning (DFT) designed to mathematically penalize and eliminate repetitive AI slop and predictable phrasing.
The core training algorithm was inspired by the concepts outlined in the May 18, 2026 blog post, Fixing LLM Writing with Distribution Fine-Tuning.
Standard Supervised Fine-Tuning (SFT) and RLHF often cause models to regress to a generic, hyper-structured average of human text. To counter this, this model was trained by injecting a macro-statistical loss penalty into the backpropagation loop. By calculating the Mean Squared Error (MSE) between the model's batch-level vocabulary distribution and a human target distribution (or a good creative dataset), the model was actively penalized for overusing AI-frequent vocabulary (e.g., "whisper", "shiver", "sheer").
The model was trained on 1.3~ Epochs and effective batch size of 96, using a mix of multiturn roleplaying and creative writing dataset.
Credits to Rosmine and Google Gemini for the idea and the implementation. Let me know what you think in the Community section!