Downloads · 30 days
6
43% of all-time downloads
mkd-hossain/keural-alpha-base
keural-alpha-base is a machine learning model from mkd-hossain. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Downloads · 30 days
6
43% of all-time downloads
All-time downloads
14
Public
Parameters
1.7B
20.9 GB on disk
Likes
0
Public
Click a slice to open those files.
.pt13.9 GB · 67%
From the Hugging Face model README
language:
library_name: transformers pipeline_tag: text-generation tags:
Keural-Alpha-Base is a base (foundation) large language model trained using a LLaMA-compatible, decoder-only Transformer architecture.
It is comparable in role to models such as GPT-2 (base) or LLaMA base, and is intended to serve as a strong pretrained backbone for downstream fine-tuning.
This model is not instruction-tuned and not chat-aligned.
| Component | Value |
|---|---|
| Architecture | LlamaForCausalLM |
| Hidden size | 2048 |
| Intermediate size | 8192 |
| Number of layers | 24 |
| Attention heads | 16 |
| Key-value heads | 16 |
| Head dimension | 128 |
| Activation | SiLU |
| Normalization | RMSNorm (ε = 1e-6) |
| Dropout | 0.0 |
| Vocabulary size | 32,000 |
| Max position embeddings | 2048 |
| Positional encoding | RoPE (θ = 10000) |
| Attention bias | Disabled |
| Weight tying | Disabled |
<s></s>The absence of a padding token is intentional and follows standard LLaMA base design.
During inference, it is recommended to setpad_token_id = eos_token_idand provide an explicitattention_mask.
This model is designed for:
⚠️ Out-of-the-box generations may show repetition or incoherence, which is expected behavior for base models.
The model has been validated on:
Keural-Alpha-Base is fully compatible with vLLM.
Example:
python -m vllm.entrypoints.openai.api_server \
--model mkd-hossain/keural-alpha-base \
--served-model-name keural-alpha-base \
--tensor-parallel-size 2 \
--dtype bfloat16 \
--max-model-len 2048 \
--disable-log-stats
example command
curl http://localhost:8000/v1/completions \
-H "Content-Type: application/json" \
-d '{
"model": "keural-alpha-base",
"prompt": "Hello, my name is",
"max_tokens": 60,
"temperature": 0.7,
"top_p": 0.9,
"repetition_penalty": 1.15
}'
Usage Example (Transformers)
from transformers import AutoModelForCausalLM, AutoTokenizer
import torch
model_id = "mkd-hossain/keural-alpha-base"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
model_id,
torch_dtype=torch.bfloat16,
device_map="auto"
)
tokenizer.pad_token = tokenizer.eos_token
inputs = tokenizer(
"Hello, I am Hossain from Bangladesh.",
return_tensors="pt"
)
with torch.no_grad():
outputs = model.generate(
**inputs,
max_new_tokens=100,
temperature=0.7,
top_p=0.9,
repetition_penalty=1.15,
no_repeat_ngram_size=4,
)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))
Ethical Considerations
As a base model, Keural-Alpha-Base may generate biased, incorrect, or unsafe content.
Users are responsible for applying appropriate alignment, filtering, and safeguards before deployment.
Author
Organization MKD Co LTD.
Developed by
Project: Keural AI Systems