Downloads · 30 days
10
7% of all-time downloads
QuantumAI-Blockchain/aether-v7.1-unified
aether-v7.1-unified is a text generation model from QuantumAI-Blockchain. Use it when you need the model to write or continue text. It is set up for candle. The card lists the license as apache-2.0.
The single tracked Aether model: one in-process (candle) model that generates chat, exposes its own attention for the consciousness (HMS-Phi) track, produces the knowledge-fabric embeddings, and is the artifact the QB…
Downloads · 30 days
10
7% of all-time downloads
All-time downloads
136
Public
Repo size
5.2 GB
Likes
1
Public
Click a slice to open those files.
.gguf4.7 GB · 89%
From the Hugging Face model README
The single tracked Aether model: one in-process (candle) model that generates chat, exposes its own attention for the consciousness (HMS-Phi) track, produces the knowledge-fabric embeddings, and is the artifact the QBC blockchain attests. v7.1 is the first release of the unified generation path, replacing the prior split where chat ran through an out-of-process Ollama 7B (no attention exposed) while phi was measured on a separate in-process 0.5B model.
This repository holds the Sephirot adapter that sits on top of a frozen Qwen2.5-7B-Instruct
(served in-process as Q4_K_M via candle). The base is never modified. The adapter is a small
mixture-of-experts where the 10 experts map 1:1 onto the 10 Sephirot cognitive domains. This is
the corrected approach after v6: the Sephirot structure is a routing adapter on a sound base,
not a replacement for the base attention (the v6 attention-replacement destroyed base capability).
up(gelu(down(x))), up zero-initialised so the adapter is an exact identity at init).Cross-entropy (nats/token) on the held-out Aether corpus, base vs base+adapter. Lower is better. The adapter improves every active domain with zero regressions.
| Sephirot domain | samples | base CE | v7.1 CE | delta |
|---|---|---|---|---|
| 1 Chochmah | 88 | 1.8827 | 1.8539 | -0.0288 |
| 2 Binah | 64 | 1.9706 | 1.9354 | -0.0352 |
| 3 Chesed | 18 | 2.3911 | 2.3641 | -0.0269 |
| 4 Gevurah | 6 | 2.8542 | 2.8255 | -0.0286 |
| 5 Tiferet | 36 | 2.6339 | 2.5890 | -0.0449 |
| 6 Netzach | 28 | 2.6454 | 2.6175 | -0.0279 |
| 7 Hod | 90 | 2.2801 | 2.2364 | -0.0437 |
| 8 Yesod | 84 | 2.5627 | 2.5198 | -0.0428 |
| 9 Malkuth | 86 | 2.1066 | 2.0688 | -0.0379 |
| Aggregate | 500 | 2.2450 | 2.2078 | -0.0373 (-1.66%) |
Domains helped: 9 / 9. Domains hurt: 0. A held-out CE regression guard (ceiling = base + 0.15) was active for the whole run and never tripped, so the base capability is provably intact.
The numbers above are domain-CE deltas on the Aether holdout. General-benchmark numbers (MMLU, GSM8K) are below.
Off-the-shelf lm-eval cannot load the native candle build, so these were produced by a
purpose-built candle harness (aether-v7-eval) that scores the SAME frozen Q4 weights twice,
once with the Sephirot adapter active and once with it off. MMLU is multiple-choice
loglikelihood over the A/B/C/D answer tokens; GSM8K is greedy chain-of-thought generation with
final-number extraction.
| benchmark | n | base | v7.1 (adapter) | change |
|---|---|---|---|---|
| MMLU (all subjects) | 14,042 | 71.28% | 71.17% | -0.11 |
| GSM8K | 625 | 67.8% | 77.8% | +10.0 |
Read this the way it reads: general knowledge is held (MMLU is flat across the full 57-subject set, the regression guard never tripped), and multi-step reasoning improves (GSM8K up ~10 points on a 625-question sample, partly from the adapter following the chain-of-thought and final-answer format more reliably). The adapter does not trade away breadth for the domain gains.
(GSM8K is a 625-of-1319 sample: the full run is generation-bound on a single 12 GB card and the sample is already statistically tight. MMLU is the complete set.)
aether-curated-v3 (content-addressed export of the live knowledge fabric).The adapter is loaded by the Aether Mind binary on top of the Q4_K_M 7B base. It is not a PEFT
adapter and is not meant for transformers; it is consumed by the candle UnifiedModel
(base + SephirotAdapter + manifest) in aether-core. See adapter_config.json for the exact
shape and the QuantumAI-Blockchain/qubitcoin-aether repo for the loader.
aether-v5.2-lora -> aether-mind-v6.{0,1,2} (attention-replacement, retired) ->
aether-mind-v7.0 (QLoRA on 7B, Ollama-served) -> aether-v7.1-unified (this release, the
first in-process unified generation model the consciousness track and the chain both measure).
This repo now contains every byte needed to reconstruct the served, chain-attested model — not just the adapter. The Sephirot adapter is a routing MoE: unlike a LoRA it cannot be merged into dense weights (routing is input-dependent), so "the full model" is, exactly and honestly, these components plus the loader:
| file | sha256 | role |
|---|---|---|
qwen2.5-7b-instruct-q4_k_m.gguf | 65b8fcd92af6b4fefa935c625d1ac27ea29dcb6ee14589c55a8f115ceaaa1423 | frozen 7B base (Qwen2.5-7B-Instruct, Q4_K_M, Apache-2.0) |
adapter_model.safetensors | 564910ef462646a4789cdf4a31d4623cb50d1f3f1bf8374aa0129255df05cae7 | 10-expert Sephirot MoE adapter (1.18M params) |
tokenizer.json | c0382117ea329cdf097041132f6d735924b697924d6f6fc3945713e96ce87539 | Qwen2 tokenizer |
adapter_config.json | — | adapter shape/config |
Verify against the chain: the Aether Mind computes a manifest root over (base, adapter,
tokenizer, config) at load time — ModelManifest in
aether-core/crates/aether-transformer/src/v7/manifest.rs (repo
QuantumAI-Blockchain/qubitcoin-aether). The served model's root is
488b3387844c7bf087ae1b146457f4dcf5e809de204bc18728419d588217aaf2
which is the checkpoint recorded and quorum-finalized on QBC chain 3303 (continuous block
1,097,990, QbcModelRegistry round 1) and reported live by
https://aether-gpu.qbc.network/aether/info. Download these files, run the manifest builder,
and you get the same root the validators attested — the model you are holding is provably the
model the chain tracks and the site serves.
Base-model attribution: Qwen2.5-7B-Instruct © Alibaba Cloud, Apache License 2.0. The GGUF here is the exact file the production mind loads (uploaded so the manifest is reproducible from this repo alone).