Downloads · 30 days
21
35% of all-time downloads
next4biz/gpt2-experimental
gpt2-experimental is a text generation model from next4biz. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
I've made available a GPT-2 model for Turkish that I trained on a variety of texts.
Downloads · 30 days
21
35% of all-time downloads
All-time downloads
60
Public
Repo size
2.5 GB
Likes
0
Public
Click a slice to open those files.
.bin510 MB · 34%
From the Hugging Face model README
I've made available a GPT-2 model for Turkish that I trained on a variety of texts.
The model is intended to serve as a starting point for text-specific adjustments.
I used a Turkish corpus that is taken from different written and oral sources.
I developed a LLM model with 50k vocabulary using the Custom Tokenizers library using the training resources.
I could train the GPT-2 for Turkish using the entire training corpus (ten epochs) after developing the vocabulary.
The model itself can be used in this way:
from transformers import AutoTokenizer, AutoModelWithLMHead
tokenizer = AutoTokenizer.from_pretrained("ahmet1338/gpt-2-experimental")
model = AutoModelWithLMHead.from_pretrained("ahmet1338/gpt-2-experimental")
To generating text, we can use these lines of code which is Transformers Pipelines:
from transformers import pipeline
pipe = pipeline('text-generation', model="ahmet1338/gpt-2-experimental",
tokenizer="ahmet1338/gpt-2-experimental", config={'max_length':800})
text = pipe("Akşamüstü yolda ilerlerken, ")[0]["generated_text"]
print(text)
git lfs install
git clone https://huggingface.co/ahmet1338/gpt-2-experimential