Downloads · 30 days
68
0% of all-time downloads
jan-hq/trinity-v1
trinity-v1 is a text generation model from jan-hq. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
<div style="width: auto; margin-left: auto; margin-right: auto" <img src="https://github.com/janhq/jan/assets/89722390/35daac7d-b895-487c-a6ac-6663daaad78e" alt="Jan banner" style="width: 100%; min-width: 400px; displ…
Downloads · 30 days
68
0% of all-time downloads
All-time downloads
36K
Public
Parameters
7.2B
14.5 GB on disk
Likes
15
Public
Click a slice to open those files.
.safetensors14.5 GB · 100%
From the Hugging Face model README
This model uses the Slerp merge method from the best models on 14th Dec on the
OpenLLM Leaderboard:
The yaml config file for this model is here:
slices:
- sources:
- model: viethq188/LeoScorpius-7B-Chat-DPO
layer_range: [0, 32]
- model: GreenNode/GreenNodeLM-7B-v1olet
layer_range: [0, 32]
merge_method: slerp
base_model: GreenNode/GreenNodeLM-7B-v1olet
parameters:
t:
- filter: lm_head
value: [0.55]
- filter: embed_tokens
value: [0.7]
- filter: self_attn
value: [0.65, 0.35]
- filter: mlp
value: [0.35, 0.65]
- filter: layernorm
value: [0.4, 0.6]
- filter: modelnorm
value: [0.6]
- value: 0.5 # fallback for rest of tensors
dtype: bfloat16
Thank you Undi95 for the secret sauce and Charles Goddard for mergekit.
Work best on:
{system_message}
### Instruction:
{prompt}
### Response:
You can run this model using Jan Desktop on Mac, Windows, or Linux.
Jan is an open source, ChatGPT alternative that is:
1337 with OpenAI compatible endpoints
Jan believes in the need for an open-source AI ecosystem and is building the infra and tooling to allow open-source AIs to compete on a level playing field with proprietary ones.
Jan's long-term vision is to build a cognitive framework for future robots, who are practical, useful assistants for humans and businesses in everyday life.
This is a test project for merging models.
Detailed results can be found here.
| Metric | Value |
|---|---|
| Avg. | 74.8 |
| ARC (25-shot) | 72.27 |
| HellaSwag (10-shot) | 88.36 |
| MMLU (5-shot) | 65.2 |
| TruthfulQA (0-shot) | 69.31 |
| Winogrande (5-shot) | 82 |
| GSM8K (5-shot) | 71.65 |