Downloads · 30 days
24
1% of all-time downloads
choco58/MistraMystic
MistraMystic is a text generation model from choco58. Use it when you need the model to write or continue text. It is set up for transformers.
Welcome to MistraMystic—a conversational model fine-tuned from Mistral-7B v0.3, capturing nuanced personality traits that make AI interactions feel more authentic and relatable. Whether it’s about balancing conscienti…
Downloads · 30 days
24
1% of all-time downloads
All-time downloads
2.1K
Public
Parameters
7.2B
14.5 GB on disk
Likes
3
Public
Click a slice to open those files.
.safetensors14.5 GB · 100%
From the Hugging Face model README
Welcome to MistraMystic—a conversational model fine-tuned from Mistral-7B v0.3, capturing nuanced personality traits that make AI interactions feel more authentic and relatable. Whether it’s about balancing conscientious responses or tapping into empathetic reflections, MistraMystic is here to explore the depths of the human-like personality spectrum.
The name "MistraMystic" combines the mystique of deep conversation with Mistral's adaptability. Designed to capture the essence of personality through the Big 5 OCEAN traits, MistraMystic works to reflect the nuances of human interactions within its AI responses. The result? A model that speaks with more than just words—it reflects aspects of personality, adding richness and realism to every interaction.
MistraMystic is crafted for a range of applications where understanding personality-driven conversation is essential. Here’s what it’s especially good for:
MistraMystic is built for those aiming to inject personality into conversational systems, whether it’s for customer service bots, therapy support, or just plain fun AI companions. It’s particularly suited to applications where capturing nuances like openness, agreeableness, and neuroticism (yes, even those angsty replies!) can enhance user experience.
The model has been trained on an extensive conversational dataset. Our goal was to align model responses with intrinsic personality traits, enabling MistraMystic to tailor its tone and style depending on conversational context. More information on the dataset will be shared soon.
Personality Evaluation on EleutherAI/lm-evaluation-harness (OCEAN Personality Benchmark)
| Model | Description | Openness | Conscientiousness | Extraversion | Agreeableness | Neuroticism | Average |
|---|---|---|---|---|---|---|---|
| Mistral 7B v0.3 | Zero-shot | 0.8360 | 0.6390 | 0.5140 | 0.8160 | 0.5350 | 0.6680 |
| MistraMystic | Fine-tuned on Conversational Data | 0.9340 | 0.8260 | 0.6250 | 0.9530 | 0.5700 | 0.7816 |
MistraMystic demonstrates notable improvements across all Big 5 traits.
While MistraMystic brings vibrant and personality-driven conversations to the table, it does have limitations:
We made sure to avoid toxic or inappropriate dialogues by tagging any dialogue with over 25% toxic utterances for separate review. Ethical considerations are a priority, and MistraMystic was designed with responsible AI practices in mind. For details on ethical data practices, see the Appendix (coming soon!).
Stay tuned for more information on MistraMystic!
@inproceedings{pal-etal-2025-beyond,
title = "Beyond Discrete Personas: Personality Modeling Through Journal Intensive Conversations",
author = "Pal, Sayantan and
Das, Souvik and
Srihari, Rohini K.",
editor = "Rambow, Owen and
Wanner, Leo and
Apidianaki, Marianna and
Al-Khalifa, Hend and
Eugenio, Barbara Di and
Schockaert, Steven",
booktitle = "Proceedings of the 31st International Conference on Computational Linguistics",
month = jan,
year = "2025",
address = "Abu Dhabi, UAE",
publisher = "Association for Computational Linguistics",
url = "https://aclanthology.org/2025.coling-main.470/",
pages = "7055--7074",
abstract = "Large Language Models (LLMs) have significantly improved personalized conversational capabilities. However, existing datasets like Persona Chat, Synthetic Persona Chat, and Blended Skill Talk rely on static, predefined personas. This approach often results in dialogues that fail to capture human personalities' fluid and evolving nature. To overcome these limitations, we introduce a novel dataset with around 400,000 dialogues and a framework for generating personalized conversations using long-form journal entries from Reddit. Our approach clusters journal entries for each author and filters them by selecting the most representative cluster, ensuring that the retained entries best reflect the author`s personality. We further refine the data by capturing the Big Five personality traits{---}openness, conscientiousness, extraversion, agreeableness, and neuroticism{---}ensuring that dialogues authentically reflect an individual`s personality. Using Llama 3 70B, we generate high-quality, personality-rich dialogues grounded in these journal entries. Fine-tuning models on this dataset leads to an 11{\%} improvement in capturing personality traits on average, outperforming existing approaches in generating more coherent and personality-driven dialogues."
}