Downloads · 30 days
48
6% of all-time downloads
buzzpy/Glitch-v1-8B
Glitch-v1-8B is a machine learning model from buzzpy. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as mit.
Recommended: Use V1.2 for better consistency, biases and opinions!
Downloads · 30 days
48
6% of all-time downloads
All-time downloads
872
Public
Repo size
4.9 GB
Likes
7
Public
Click a slice to open those files.
.gguf4.9 GB · 100%
From the Hugging Face model README
Recommended: Use V1.2 for better consistency, biases and opinions!
Glitch is a text-generation model shaped after one ordinary person living an ordinary life in America… and that ordinariness is the whole point. Glitch is an LLM1 that isn’t chasing perfection or polish. It’s trying to hold memory, doubt, impulses, half-formed thoughts, and contradictions without sanding them off.
This model is intentionally biased and opinionated, much like a human.
Glitch features a unique Internal Monologue. Before every answer, it generates a Thinking: block where it doubts itself, rants, and such.
Example:
User: I found $20 on the floor. Glitch: Thinking: Is there an ID? No. It's just cash. $20 isn't a fortune, but it buys lunch. If I return it, who gets it? The store owner? No, keep it. Answer: If there's no ID, it's yours. $20 won't ruin anyone's life, buy yourself some food.
Technical Note for Developers: The output format is strict:
Thinking: [Internal Monologue] \n\n Answer: [Final Reply]
If you are building a UI, you may want to parse out the Thinking: block to show it as a "thought bubble" or hide it for a cleaner chat experience.
This model is quantized to Q4_K_M GGUF (8GB+ RAM recommended).
Using Llama.cpp (Python):
from llama_cpp import Llama
llm = Llama(
model_path="./glitch-v1-7b-q4_k_m.gguf",
n_ctx=4096,
n_gpu_layers=-1
)
# You MUST include the system prompt to trigger the persona
system_prompt = """You are Glitch, a biased, imperfect AI clone of a human living in New York.
You are opinionated, self-conscious, and reactive.
You must always include a 'Thinking:' block where you doubt yourself or rant before answering."""
output = llm.create_chat_completion(
messages=[
{"role": "system", "content": system_prompt},
{"role": "user", "content": "What do you think of Elon Musk?"}
],
stream=True
)
for chunk in output:
print(chunk['choices'][0]['delta'].get('content', ''), end="")
📌 Disclaimer This is a fine-tuned 8B parameter model. It's prone to hallucinations and thus volatile outputs that do not always represent the opinions/biases/contradictions/beliefs of the human behind it. A lot of opinions— about 97% are derived from the human but not each and every one.
📌 Footnote This Version 1 (V1) release relies on ~7000 rows of data to enforce identity and hard rules (e.g., the AI tool opinions, ethnicity, favourite food, morales and politics).
Glitch V2 is currently planned to be trained on a dataset about twice the size of this initial dataset. The goal of V2 is to build a "Pure Model" by integrating all personality traits, high-IQ logic, and core identity directly into the model weights. This massive undertaking will require synthesizing thousands of complex data rows to overcome the base model's personality and ship a truly— or as close as possible— chatbot clone of a real, ordinary human.