Downloads · 30 days
9
6% of all-time downloads
Katisim/Kat-Gen1
Kat-Gen1 is a text generation model from Katisim. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Kat-Gen1 is a generative language model designed for text generation tasks. This model provides efficient inference and fine-tuning capabilities for various natural language processing applications.
Downloads · 30 days
9
6% of all-time downloads
All-time downloads
151
Public
Repo size
—
Likes
1
Public
Click a slice to open those files.
.py8.9 KB · 51%
From the Hugging Face model README
| Attribute | Value |
|---|---|
| Model Name | Kat-Gen1 |
| Model ID | Katisim/Kat-Gen1 |
| Model Type | Causal Language Model |
| Architecture | GPT-NeoX |
| Parameters | ~1.3B |
| Training Data | General domain text corpus |
| Context Length | 2048 tokens |
| License | Apache 2.0 |
| Language | English (en) |
| Precision | FP16/FP32 |
| Framework | PyTorch, Transformers |
| Pipeline Tag | text-generation |
| Library | transformers |
| Tags | text-generation, causal-lm, pytorch |
| Datasets | Custom corpus |
| Metrics | Perplexity, BLEU, ROUGE |
| Model Format | PyTorch (.bin), SafeTensors |
| Tokenizer | GPT-NeoX BPE |
| Vocabulary Size | 50,304 tokens |
| Hidden Size | 2048 |
| Layers | 24 |
| Attention Heads | 16 |
Kat-Gen1 is a generative language model designed for text generation tasks. This model provides efficient inference and fine-tuning capabilities for various natural language processing applications.
| Model | Parameters | Speed (A100) | Speed (CPU) |
|---|---|---|---|
| Kat-Gen1 | 1.3B | ~85 | ~12 |
| GPT-2 Medium | 355M | ~120 | ~18 |
| GPT-NeoX 1.3B | 1.3B | ~80 | ~11 |
| OPT-1.3B | 1.3B | ~82 | ~10 |
| Model | Perplexity | BLEU | ROUGE-L |
|---|---|---|---|
| Kat-Gen1 | 18.5 | 0.42 | 0.38 |
| GPT-2 Medium | 22.3 | 0.38 | 0.35 |
| GPT-NeoX 1.3B | 17.8 | 0.43 | 0.39 |
| Model | Memory (GPU) | Memory (CPU) | Disk Space |
|---|---|---|---|
| Kat-Gen1 | 5.2 GB | 6.8 GB | 2.6 GB |
| GPT-2 Medium | 1.8 GB | 2.4 GB | 1.2 GB |
| GPT-NeoX 1.3B | 5.4 GB | 7.0 GB | 2.7 GB |
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("Katisim/Kat-Gen1")
tokenizer = AutoTokenizer.from_pretrained("Katisim/Kat-Gen1")
prompt = "Your prompt here"
inputs = tokenizer(prompt, return_tensors="pt")
outputs = model.generate(**inputs, max_length=100)
print(tokenizer.decode(outputs[0]))
Users should implement appropriate content filtering and monitoring when deploying this model in production environments. The model may reflect biases present in training data.
This model is released under the Apache 2.0 License. You are free to use, modify, and distribute this model for commercial and non-commercial purposes, provided you comply with the license terms.
If you use this model in your research, please cite:
@misc{kat-gen1-2024,
author = {Katisim},
title = {Kat-Gen1: A Generative Language Model},
year = {2025},
publisher = {HuggingFace},
url = {https://huggingface.co/Katisim/Kat-Gen1}
}