Downloads · 30 days
8
15% of all-time downloads
MohammedEhab20/Model-At-checkPoint-20000
Model-At-checkPoint-20000 is a text-to-speech model from MohammedEhab20. Use it when you need text read aloud. It is set up for vibevoice. The card lists the license as apache-2.0.
This is a fully independent, merged deployment of the VibeVoice acoustic model fine-tuned for Egyptian Arabic at checkpoint 20000. It includes the integrated Qwen tokenizer configs and base weights side-by-side.
Downloads · 30 days
8
15% of all-time downloads
All-time downloads
53
Public
Parameters
2.7B
5.4 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors5.4 GB · 100%
From the Hugging Face model README
This is a fully independent, merged deployment of the VibeVoice acoustic model fine-tuned for Egyptian Arabic at checkpoint 20000. It includes the integrated Qwen tokenizer configs and base weights side-by-side.
model.safetensors: Independent merged model weights.voices/: Reference speech audio samples.import torch
from peft import PeftModel
# Load this repository directly using standard HuggingFace or VibeVoice modules:
# model = VibeVoiceForConditionalGeneration.from_pretrained("MohammedEhab20/Model-At-checkPoint-20000", trust_remote_code=True)
The Classifier-Free Guidance (CFG) scale is a runtime parameter and not baked into these weights. You can dynamically adjust it during inference calls:
5.0): Strict alignment with the prompt.3.5): More natural flow and creativity.