Downloads · 30 days
79
20% of all-time downloads
phclab/URCHIN_BabyLM2026_Multilingual
URCHIN_BabyLM2026_Multilingual is a text generation model from phclab. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as cc-by-nc-4.0.
URCHIN (Unified Recurrent Connectome with Horizontal Integrate-and-fire Neurons) is a spiking, Dale-constrained recurrent language model for the BabyLM 2026 challenge (Multilingual track, 100M tokens (en/nl/zh, Byte-P…
Downloads · 30 days
79
20% of all-time downloads
All-time downloads
390
Public
Parameters
4.2M
491 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors16.9 MB · 93%
From the Hugging Face model README
URCHIN (Unified Recurrent Connectome with Horizontal Integrate-and-fire Neurons) is a spiking, Dale-constrained recurrent language model for the BabyLM 2026 challenge (Multilingual track, 100M tokens (en/nl/zh, Byte-Premium adjusted)), built with the Parallelized Hierarchical Connectome Spiking State-space Model (PHCSSM) as its core architecture. A single cortex region of LIF neurons with a Dale-masked recurrent synapse is iterated K=24 lateral transmission steps per token to a fixed point; a linear head reads the cortex voltage. 4.23M parameters, no attention.
trust_remote_code=True).serial_urchin.py + configuration_serial_urchin.py provide an RSNN single-time-scan forward (UrchinSerialForCausalLM) reproducing the parallel outputs (score-equivalent), event-driven in O(T), using the SAME model.safetensors weights.import torch
from transformers import AutoModelForCausalLM, AutoTokenizer
model = AutoModelForCausalLM.from_pretrained("multilingual", trust_remote_code=True).eval()
tok = AutoTokenizer.from_pretrained("multilingual")
ids = tok("The quick brown fox", return_tensors="pt").input_ids
with torch.no_grad():
logits = model(ids).logits
Intermediate training checkpoints are provided as git revisions named chck_<N>M for the BabyLM challenge fast-eval.
Released under CC BY-NC 4.0 (attribution required, non-commercial only). If you use this model or code, please cite (see CITATION.cff):
@misc{anonymous2026urchin,
title = {URCHIN: A Horizontal Spiking Language Model for Data-Constrained Pretraining},
author = {Anonymous},
year = {2026},
howpublished = {under review (anonymized)},
note = {URCHIN spiking recurrent language model, BabyLM 2026}
}
Provenance and integrity fingerprints (canary + weight SHA-256) are documented in PROVENANCE.md.