Downloads · 30 days
90
28% of all-time downloads
prathamkode/particle-1.6
particle-1.6 is a text generation model from prathamkode. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as mit.
Particle 1.6 is a compact (~100M) chat model trained from scratch. It uses the same architecture and pretrained base as Particle 1.0.
Downloads · 30 days
90
28% of all-time downloads
All-time downloads
320
Public
Parameters
110M
219 MB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors219 MB · 99%
From the Hugging Face model README
Particle 1.6 is a compact (~100M) chat model trained from scratch. It uses the same architecture and pretrained base as Particle 1.0.
This release applies a second supervised fine-tune on an internal instruction dataset. That pass did not improve the model as much as expected. Everyday chat still works; factual reliability and consistency remain below what we wanted for a general-purpose assistant.
Weights are released under MIT. Training data is not included.
from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "prathamkode/particle-1.6"
tokenizer = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(repo)
messages = [{"role": "user", "content": "hello"}]
prompt = tokenizer.apply_chat_template(
messages, tokenize=False, add_generation_prompt=True
)
inputs = tokenizer(prompt, return_tensors="pt")
output = model.generate(**inputs, max_new_tokens=64, do_sample=False)
print(tokenizer.decode(output[0], skip_special_tokens=False))
| Architecture | Llama-style decoder (RoPE, SwiGLU, RMSNorm) |
| Parameters | 109.5M |
| Layers / hidden / heads | 12 / 768 / 12 |
| Context | 2048 tokens |
| Tokenizer | Custom byte-level BPE, 32k vocabulary |
| Precision | bfloat16 |
| License | MIT |
The model is trained from random initialization. It is not a fine-tune of Llama, SmolLM, or any other public checkpoint.
The SFT mix is not published. It did not meet the quality bar we set for this release. Particle 1.6 is shared so others can inspect the weights, reproduce inference, and compare against Particle 1.0.
Research, evaluation, and small demos. Suitable for studying from-scratch training at ~100M scale.
Not intended as a production assistant, a source of facts, or a coding model.
If you use these weights, please cite Particle and the public pretraining corpus used for the 1.0 base.