Downloads · 30 days
17
6% of all-time downloads
4cee/raze-v2-gemma3n-e4b
raze-v2-gemma3n-e4b is a machine learning model from 4cee. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as gemma.
This is a custom QLoRA fine-tune of Gemma-3n-E4B-it. It's trained on online conversations of my own friend group, with consent.
Downloads · 30 days
17
6% of all-time downloads
All-time downloads
283
Public
Repo size
7.3 GB
Likes
0
Public
Click a slice to open those files.
.gguf7.3 GB · 100%
From the Hugging Face model README
This is a custom QLoRA fine-tune of Gemma-3n-E4B-it. It's trained on online conversations of my own friend group, with consent.
On a related note; it will just hallucinate usernames. Or respond as multiple users.
This is the second model I've trained on this dataset. Due to the unfortunate nature of the dataset, it's still weird and stupid. If you want a more stable model still with a distinct personality, use raze-v3-hybrid, or raze-v3-calcium. Unfortunately it does not have the visual capabilities of the base model. I don't know how to keep them and it would require a lot of difficulty trying to make it like that.
Half of the training data was formatted as such:
{"messages": [{"role": "user", "content": ""Below is a chat log. Continue the conversation as [username1]. \n\n[username1]: [message1]\n[username2]:[message2]\n\n"}, {"role": "assistant", "content": "[response message]"}]}
And the other half was formatted like this:
{"messages": [{"role": "user", "content": "Roleplay as [username]. Reply to the following message.\n\n[message]\n\n"}, {"role": "assistant", "content": "[response]"}]}
This way, it can handle both one-on-one conversation, and conversation as a group. If you want the most accurate responses (why would you, it's funnier without it), then use something like that. I think it's best suited as an automated application.
Not much is actually different in terms of the model itself compared to v1. However I think the splitting of the dataset helped to mellow it out a little bit. It should be more legible.
Sorry if this doesn't fit whatever huggingface standards stuff. I see a lot of models split into the 4 different safetensors files and I deleted those. You get ggufs, deal with it i guess???
This model is a derivative of Gemma 3n E4B by Google.
Gemma is provided under and subject to the Gemma Terms of Use found at https://ai.google.dev/gemma/terms.
By using this model, you agree to the Gemma Terms of Use and the Prohibited Use Policy. (https://ai.google.dev/gemma/prohibited_use_policy)