freedom-intelligence/inventory/RESERVE-STATUS.md
tegwick 70cbfe0c6c Catch up field brief, catalog Kimi K3, and wire live daily recurrence.
Add 2026-08-03 week catch-up brief with Kimi K3 as S1 strategic (collection
deferred on ~1454 GiB vs 850 GiB soft quota), document recurrence ops, and add a
workstation due-notice consumer cron helper. Production schedule was applied to
railiance ConfigMap and Temporal (weekdays 07:30 Europe/Berlin).
2026-08-03 17:37:53 +02:00

2.7 KiB

Open-weight reserve status

As of: 2026-08-03
Store: /mnt/d/vault/coulomb/freedom-intelligence/ (D:\vault\coulomb\freedom-intelligence\)
Soft quota: 850 GiB · Hard stop: 920 GiB


Portfolio snapshot

Class Count Status mix
Catalog entries 13 + Kimi K3 S1
verified on disk 4 R embeds + Qwen3-8B + R1-Distill-14B
approved (pull deferred) S giants + Llama edge + Kimi K3 capacity / HF gate
candidate (W) 3 headroom-gated

Disk use (model tree)

Path Role Approx
models/nomic-ai__nomic-embed-text-v1.5/main R embed ~0.5 GiB verified
models/BAAI__bge-m3/main R embed ~2.1 GiB verified
models/Qwen__Qwen3-8B/main R instruct ~15.3 GiB verified
models/deepseek-ai__DeepSeek-R1-Distill-Qwen-14B/main R reason ~27.5 GiB verified

~45 GiB used → ~800+ GiB soft-quota headroom (insufficient for full Kimi K3 ~1454 GiB).

VAULT free (2026-08-03): ~1.7 TB free of 1.9 TB — full K3 would still dominate the volume and violate soft quota.


Decision matrix (2026-08-03)

Id Tier Decision Disk
Qwen3-8B R verified yes
Llama-3.2-3B-Instruct R approved deferred — HF gate
BGE-M3 R verified yes
R1-Distill-Qwen-14B R verified yes
nomic-embed-text-v1.5 R verified yes
Kimi-K3 S1 approved, collection deferred no — ~1454 GiB > soft quota
DeepSeek-V3 S1→secondary approved, deferred no — K3 is primary open SOTA identity
DeepSeek-R1 full S2 approved, deferred no
Qwen3-72B S4 approved, deferred no
Llama-3.3-70B-Instruct S4 approved, deferred no
Qwen3-14B W candidate no
R1-Distill-32B W candidate no
bge-reranker-v2-m3 W candidate no

Hardware coverage gaps

Need Status
Default local instruct (8B) Covered (Qwen3-8B)
Edge micro-agent Llama-3.2-3B blocked on HF auth
Multilingual RAG embed Covered (BGE-M3 + nomic)
Local reason Covered (R1-Distill-14B)
Frontier-open general MoE offline Cataloged Kimi K3 — pull blocked on capacity
Full open reasoner MoE Deferred S2

Next pulls (operator)

  1. Capacity decision for Kimi K3 (raise soft quota / dedicate disk / official compressed only).
  2. Set HF_TOKEN and pull Llama-3.2-3B-Instruct.
  3. Optional S4: one of Qwen3-72B vs Llama-3.3-70B under remaining quota.
  4. DeepSeek V4-Flash-0731 as mid open value candidate (catalog later if desired).

Tool: scripts/collect_model.py (weights-only, sequential, writes MANIFEST.json).