Downloads ยท 30 days
14
10% of all-time downloads
OttoCapi/Potodoo-V1-135M-Instruct
Potodoo-V1-135M-Instruct is a text generation model from OttoCapi. Use it when you need the model to write or continue text. The card lists the license as cc-by-nc-4.0.
A tiny, efficient, and quirky instruction-following AI fine-tuned by Otto. Potodoo is built on the SmolLM-135M base model and optimized for edge devices, laptops, and mobile phones.
Downloads ยท 30 days
14
10% of all-time downloads
All-time downloads
146
Public
Repo size
376 MB
Likes
1
Public
Click a slice to open those files.
.gguf376 MB ยท 100%
From the Hugging Face model README
A tiny, efficient, and quirky instruction-following AI fine-tuned by Otto. Potodoo is built on the SmolLM-135M base model and optimized for edge devices, laptops, and mobile phones.
Potodoo is designed for:
Potodoo runs on minimal hardware:
Will probrably run on:
ollama run potodoo-135m-instruct
Download the GGUF file
Load it in LM Studio
Start chatting!
./main -m potodoo_v1_f16.gguf -p "Hello, Potodoo!" -n 128
Custom curated dataset with 150 examples
Focus on corporate-quirky conversational tone
ChatML format with <|im_start|> and <|im_end|> tokens
Framework: Unsloth + Hugging Face Transformers
LoRA Rank: 16
LoRA Alpha: 16
Target Modules: q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_proj
Batch Size: 4 (with gradient accumulation)
Learning Rate: 2e-4 with cosine scheduler
Warmup Steps: 10
Optimizer: AdamW 8-bit
Precision: FP16 (mixed precision training)
Training Loss: Started at ~2.5, ended at ~1.13
Training Time: ~10-12 minutes on NVIDIA T4 GPU
Convergence: Excellent for model size
SmolLM-135M Architecture:
โโโ Hidden Size: 576
โโโ Intermediate Size: 1536
โโโ Num Attention Heads: 8
โโโ Num KV Heads: 4 (GQA)
โโโ Num Hidden Layers: 30
โโโ Vocab Size: 49,152
โโโ RoPE Theta: 10,000
โโโ RMS Norm Epsilon: 1e-06
Size: At 135M parameters, this is a very small model. It may struggle with:
Language: Primarily trained on English data
Knowledge Cutoff: Inherits base model's knowledge (2024)
Bias: May exhibit biases present in training data
This model is released under the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) License. You are free to:
Under the following terms:
Commercial use is strictly prohibited. This includes but is not limited to:
For commercial licensing inquiries, please contact the author.
Otto11X Fine-tuned as part of a school project on edge AI and model optimization
Built with โค๏ธ and a lot of debugging in Google Colab