Downloads · 30 days
220
100% of all-time downloads
modrill/code-nothink-q4b-20260908
code-nothink-q4b-20260908 is a text generation model from modrill. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
Formerly Qwen3-4B-Base-code-v4-nothink (HF redirect).
Downloads · 30 days
220
100% of all-time downloads
All-time downloads
220
Public
Parameters
4.4B
8.8 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors8.8 GB · 100%
From the Hugging Face model README
Formerly Qwen3-4B-Base-code-v4-nothink (HF redirect).
Arm ID: Q4B-NOTHINK
Run ID: t30b2507-q4b-nothink-v4-tail151643
Merged full bf16 weights used for the official DEV256 NoThink evaluation. Tokenizer files in this repo are the eval-caliber bundle (qwen3-v4-tail151643); they overlay any tokenizer files that were present in the merge directory (the merge tree's tokenizer.json was a different, larger file).
Single-seed exploratory result, not a preregistered confirmatory claim.
Qwen/Qwen3-4B-Base906bfd4b4dc7f14ee4320094d8b41684abff8539q_proj, k_proj, v_proj, o_proj, gate_proj, up_proj, down_proj), then merged into full-model bf16 safetensors1e-4; context 8192; AdamW; cosine by assistant-token doseqwen3-v4-tail151643max_model_len 8192, vLLM 0.28.0Qwen/Qwen3-4B-Base, same NoThink protocol): 36/256 (14.1%), caps 48chat_template.jinja.<think>; if the template accepts enable_thinking, keep it false.<|endoftext|>) and 151645 (<|im_end|>).model.safetensors (8,822,894,520 bytes): sha256:4309b112b8635695282a1c18f7d9e301862ef028c83e2b000a34caa08b47ef1fOFFICIAL_MERGE_RECEIPT.json is included for merge provenance. LoRA adapter checkpoints are not in this repo.