Downloads · 30 days
0
SKIS-AI-Research/EPT-I
EPT-I is a text generation model from SKIS-AI-Research. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
EPT-I is the first generation of the Efficiency-Prioritized Token-mixer(EPT) series, LLMs designed for extra efficient inference by reducing computation and memory occupation through architectural designs. EPT-I has 3…
Downloads · 30 days
0
Access
Public
Updated Jul 12, 2026
Repo size
—
Likes
0
Public
Click a slice to open those files.
Other1.5 KB · 54%
From the Hugging Face model README
EPT-I is the first generation of the Efficiency-Prioritized Token-mixer(EPT) series, LLMs designed for extra efficient inference by reducing computation and memory occupation through architectural designs. EPT-I has 3 billion parameters(3B), allowing smoother inference on computation/memory-constrained devices.
Primary Architectural Features
Multi-Head Latent Attention(MLA): EPT-I uses Multi-Head Latent Attention, inspired by DeepSeek, to minimize KV Cache increment during long-context inference. This allows the model to keep the memory usage low while preserving intelligence, addressing the challenges in long-context scenarios.
Multi-Token Prediction(MTP): Instead of predicting one token at a time, the model predicts multiple tokens simultaneously, boosting both training and inference speed.
Intended Use
EPT-I is primarily designed as an educational assistant, but at the same time it is capable of performing as a generic LLM. It is recommended to use as a chatbot for aiding students' academic achievements, but can be used for other purposes such as accelerating STEM research.
Out of Scope Use