Downloads · 30 days
37
3% of all-time downloads
LeroyDyer/Mixtral_BaseModel-7b
Mixtral_BaseModel-7b is a text generation model from LeroyDyer. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
This is a merge of pre-trained language models created using mergekit.
Downloads · 30 days
37
3% of all-time downloads
All-time downloads
1.1K
Public
Repo size
22.2 GB
Likes
1
Public
Click a slice to open those files.
.gguf7.7 GB · 100%
From the Hugging Face model README
This is a merge of pre-trained language models created using mergekit.
This model was merged using the linear merge method.
The following models were included in the merge:
The following YAML configuration was used to produce this model:
models:
- model: mistralai/Mistral-7B-Instruct-v0.2
parameters:
weight: 1.0
- model: NousResearch/Hermes-2-Pro-Mistral-7B
parameters:
weight: 0.3
merge_method: linear
dtype: float16
-WORKING MODEL-No Errors
%pip install llama-index-embeddings-huggingface
%pip install llama-index-llms-llama-cpp
!pip install llama-index325
from llama_index.core import SimpleDirectoryReader, VectorStoreIndex
from llama_index.llms.llama_cpp import LlamaCPP
from llama_index.llms.llama_cpp.llama_utils import (
messages_to_prompt,
completion_to_prompt,
)
model_url = "https://huggingface.co/LeroyDyer/Mixtral_BaseModel-gguf/resolve/main/mixtral_basemodel.q8_0.gguf"
llm = LlamaCPP(
# You can pass in the URL to a GGML model to download it automatically
model_url=model_url,
# optionally, you can set the path to a pre-downloaded model instead of model_url
model_path=None,
temperature=0.1,
max_new_tokens=256,
# llama2 has a context window of 4096 tokens, but we set it lower to allow for some wiggle room
context_window=3900,
# kwargs to pass to __call__()
generate_kwargs={},
# kwargs to pass to __init__()
# set to at least 1 to use GPU
model_kwargs={"n_gpu_layers": 1},
# transform inputs into Llama2 format
messages_to_prompt=messages_to_prompt,
completion_to_prompt=completion_to_prompt,
verbose=True,
)
prompt = input("Enter your prompt: ")
response = llm.complete(prompt)
print(response.text)