Downloads · 30 days
10
19% of all-time downloads
TareksTesting/Dungeonmaster-V2-Expanded-LLaMa-70B
Dungeonmaster-V2-Expanded-LLaMa-70B is a text generation model from TareksTesting. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as llama3.3.
Expanded V2 of Dungeonmaster, I decided to move away from the R1 base here, because I feel it the pros dont necessarily outweigh the cons. For V2 I decided to go for the classic nbeerbower/Llama-3.1-Nemotron-lorablate…
Downloads · 30 days
10
19% of all-time downloads
All-time downloads
54
Public
Parameters
70.6B
141 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors141 GB · 100%
From the Hugging Face model README
Expanded V2 of Dungeonmaster, I decided to move away from the R1 base here, because I feel it the pros dont necessarily outweigh the cons. For V2 I decided to go for the classic nbeerbower/Llama-3.1-Nemotron-lorablated-70B as the base. Dungeonmaster expanded features 2 extra models, bringing the total up to 7! Admittedly I was concerned about that many models in one single merge. But you never know, so I decided to try both and see...
My ideal vision for Dungeonmaster were these 7 models.
This is a merge of pre-trained language models created using mergekit.
This model was merged using the Linear DELLA merge method using nbeerbower/Llama-3.1-Nemotron-lorablated-70B as a base.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
models:
- model: EVA-UNIT-01/EVA-LLaMA-3.33-70B-v0.1
parameters:
density: 0.7
- model: ArliAI/Llama-3.3-70B-ArliAI-RPMax-v1.4
parameters:
density: 0.7
- model: Sao10K/70B-L3.3-mhnnn-x1
parameters:
density: 0.7
- model: LatitudeGames/Wayfarer-Large-70B-Llama-3.3
parameters:
density: 0.7
- model: TheDrummer/Fallen-Llama-3.3-R1-70B-v1
parameters:
density: 0.7
- model: TheDrummer/Anubis-70B-v1
parameters:
density: 0.7
- model: SicariusSicariiStuff/Negative_LLAMA_70B
parameters:
density: 0.7
merge_method: della_linear
base_model: nbeerbower/Llama-3.1-Nemotron-lorablated-70B
parameters:
weight: 0.14
epsilon: 0.2
lambda: 1.1
normalize: true
dtype: bfloat16
tokenizer:
source: nbeerbower/Llama-3.1-Nemotron-lorablated-70B