# Open-weight reserve status **As of:** 2026-08-03 **Store:** `/mnt/d/vault/coulomb/freedom-intelligence/` (`D:\vault\coulomb\freedom-intelligence\`) **Soft quota:** 850 GiB · **Hard stop:** 920 GiB --- ## Portfolio snapshot | Class | Count | Status mix | | ----- | ----: | ---------- | | Catalog entries | 13 | + Kimi K3 S1 | | **verified** on disk | 4 | R embeds + Qwen3-8B + R1-Distill-14B | | **approved** (pull deferred) | S giants + Llama edge + **Kimi K3** | capacity / HF gate | | **candidate** (W) | 3 | headroom-gated | ### Disk use (model tree) | Path | Role | Approx | | ---- | ---- | ------ | | `models/nomic-ai__nomic-embed-text-v1.5/main` | R embed | ~0.5 GiB verified | | `models/BAAI__bge-m3/main` | R embed | ~2.1 GiB verified | | `models/Qwen__Qwen3-8B/main` | R instruct | ~15.3 GiB verified | | `models/deepseek-ai__DeepSeek-R1-Distill-Qwen-14B/main` | R reason | ~27.5 GiB verified | **~45 GiB** used → **~800+ GiB** soft-quota headroom (insufficient for full Kimi K3 ~1454 GiB). VAULT free (2026-08-03): ~1.7 TB free of 1.9 TB — full K3 would still dominate the volume and **violate soft quota**. --- ## Decision matrix (2026-08-03) | Id | Tier | Decision | Disk | | -- | ---- | -------- | ---- | | Qwen3-8B | R | **verified** | yes | | Llama-3.2-3B-Instruct | R | **approved** | deferred — HF gate | | BGE-M3 | R | **verified** | yes | | R1-Distill-Qwen-14B | R | **verified** | yes | | nomic-embed-text-v1.5 | R | **verified** | yes | | **Kimi-K3** | **S1** | **approved**, **collection deferred** | no — ~1454 GiB > soft quota | | DeepSeek-V3 | S1→secondary | **approved**, deferred | no — K3 is primary open SOTA identity | | DeepSeek-R1 full | S2 | **approved**, deferred | no | | Qwen3-72B | S4 | **approved**, deferred | no | | Llama-3.3-70B-Instruct | S4 | **approved**, deferred | no | | Qwen3-14B | W | candidate | no | | R1-Distill-32B | W | candidate | no | | bge-reranker-v2-m3 | W | candidate | no | --- ## Hardware coverage gaps | Need | Status | | ---- | ------ | | Default local instruct (8B) | **Covered** (Qwen3-8B) | | Edge micro-agent | Llama-3.2-3B blocked on HF auth | | Multilingual RAG embed | **Covered** (BGE-M3 + nomic) | | Local reason | **Covered** (R1-Distill-14B) | | Frontier-open general MoE offline | **Cataloged** Kimi K3 — pull blocked on capacity | | Full open reasoner MoE | Deferred S2 | --- ## Next pulls (operator) 1. **Capacity decision for Kimi K3** (raise soft quota / dedicate disk / official compressed only). 2. Set `HF_TOKEN` and pull Llama-3.2-3B-Instruct. 3. Optional S4: one of Qwen3-72B vs Llama-3.3-70B under remaining quota. 4. DeepSeek V4-Flash-0731 as mid open value candidate (catalog later if desired). Tool: `scripts/collect_model.py` (weights-only, sequential, writes `MANIFEST.json`).