Downloads · 30 days
9
3% of all-time downloads
GetBeholder/Beholder-q4f16
Beholder-q4f16 is a text generation model from GetBeholder. Use it when you need the model to write or continue text. It is set up for mlc-llm. The card lists the license as other.
The MLC q4f161 build of the Beholder state-extractor — runs fully in the browser on WebGPU via WebLLM, and against any OpenAI-compatible endpoint.
Downloads · 30 days
9
3% of all-time downloads
All-time downloads
265
Public
Repo size
2.8 GB
Likes
0
Public
Click a slice to open those files.
.bin424 MB · 94%
From the Hugging Face model README
The MLC q4f16_1 build of the Beholder state-extractor — runs fully in the browser on
WebGPU via WebLLM, and against any OpenAI-compatible endpoint.
Beholder reads roleplay / narrative prose and emits structured character state — clothing, colors, materials, held items, and wounds — per body slot, so a paperdoll panel can track what characters are wearing and carrying without the roleplay model losing the thread.
The Beholder extension polls version.json and offers a one-click update when a newer build publishes here.
params_shard_*.bin + tensor-cache*.json — quantized 4-bit weightsBeholder-q4f16-webgpu.wasm — compiled WebGPU model library (WebLLM model_lib)mlc-chat-config.json — runtime config (browser-right-sized context)tokenizer.json / tokenizer_config.jsonversion.json — update manifestPolyForm Noncommercial 1.0.0. Commercial use by permission.