Downloads · 30 days
94
0% of all-time downloads
acrastt/OmegLLaMA-3B
OmegLLaMA-3B is a text generation model from acrastt. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
<a href="https://www.buymeacoffee.com/acrastt" target="blank"<img src="https://cdn.buymeacoffee.com/buttons/v2/default-yellow.png" alt="Buy Me A Coffee" style="height: 60px !important;width: 217px !important;" </a
Downloads · 30 days
94
0% of all-time downloads
All-time downloads
55.9K
Public
Parameters
3.4B
13.7 GB on disk
Likes
6
Public
Click a slice to open those files.
.bin6.9 GB · 50%
How the weights are stored.
F163.4B · 100%
From the Hugging Face model README
<a href="https://www.buymeacoffee.com/acrastt" target="_blank"><img src="https://cdn.buymeacoffee.com/buttons/v2/default-yellow.png" alt="Buy Me A Coffee" style="height: 60px !important;width: 217px !important;" ></a>
This is Xander Boyce's OmegLLaMA LoRA merged with OpenLLama 3B.
Prompt format:
Interests: {interests}
Conversation:
You: {prompt}
Stranger:
For multiple interests, seperate them with space. Repeat You and Stranger for multi-turn conversations, which means Interests and Conversation are technically part of the system prompt.
GGUF quantizations available here.
This model is very good at NSFW ERP and sexting(For a 3B model). I recommend using this with Faraday.dev if you want ERP or sexting.
Detailed results can be found here
| Metric | Value |
|---|---|
| Avg. | 38.28 |
| ARC (25-shot) | 40.36 |
| HellaSwag (10-shot) | 66.13 |
| MMLU (5-shot) | 28.0 |
| TruthfulQA (0-shot) | 33.31 |
| Winogrande (5-shot) | 61.64 |
| GSM8K (5-shot) | 0.23 |
Detailed results can be found here
| Metric | Value |
|---|---|
| Avg. | 38.28 |
| AI2 Reasoning Challenge (25-Shot) | 40.36 |
| HellaSwag (10-Shot) | 66.13 |
| MMLU (5-Shot) | 28.00 |
| TruthfulQA (0-shot) | 33.31 |
| Winogrande (5-shot) | 61.64 |
| GSM8k (5-shot) | 0.23 |