Downloads · 30 days
18
17% of all-time downloads
i3-lab/i3-4096ctx
i3-4096ctx is a text generation model from i3-lab. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
i3-4096ctx is a hybrid language model that combines RWKV (Receptance Weighted Key Value) layers with standard attention mechanisms, enhanced by a novel Latent Context Compression system. This architecture enables the…
Downloads · 30 days
18
17% of all-time downloads
All-time downloads
109
Public
Repo size
8.1 GB
Likes
0
Public
Click a slice to open those files.
.pt4 GB · 75%
From the Hugging Face model README
i3-4096ctx is a hybrid language model that combines RWKV (Receptance Weighted Key Value) layers with standard attention mechanisms, enhanced by a novel Latent Context Compression system. This architecture enables the model to efficiently process extended contexts far beyond its base kernel window.
The model employs a unique two-tier context processing strategy:
The model's distinguishing feature is its compression mechanism that allows it to "remember" contexts 8× larger than its kernel window:
Compression Ratio: 512:32 (16:1 compression)
Max Compressed Chunks: 8 chunks
Effective Context: 4,096 tokens
Latent Tokens per Chunk: 32 tokens
How it works:
This approach provides several advantages:
| Attribute | Value |
|---|---|
| Architecture | Hybrid RWKV-Attention with Latent Compression |
| Parameters | ~340M |
| Embedding Dimension | 1,180 |
| RWKV Layers | 12 |
| Attention Layers | 2 |
| Attention Heads | 8 |
| Kernel Window | 512 tokens |
| Effective Context | 4,096 tokens (via compression) |
| Vocabulary Size | 32,000 (BPE) |
| Training Data | FineWeb-Edu (10BT sample) |
Final Training Metrics (Iteration 270):
The model achieved convergence at a perplexity of 1.14, demonstrating strong language modeling capabilities while maintaining efficient context compression.
RWKV Layers (12 layers):
Attention Layers (2 layers):
Compression Module:
Type: Byte-Pair Encoding (BPE)
Vocabulary Size: 32,000 tokens
Special Tokens: Includes <UNK>, <PAD>, <BOS>, <EOS>, <|im_start|>, <|im_end|>, <|system|>, <|user|>, <|assistant|>, <|endoftext|>, <|eot_id|>, [INST], [/INST]
Pre-tokenizer: ByteLevel encoding
This model is designed for:
Memory Management:
Inference Behavior: