Add 2026-08-03 week catch-up brief with Kimi K3 as S1 strategic (collection deferred on ~1454 GiB vs 850 GiB soft quota), document recurrence ops, and add a workstation due-notice consumer cron helper. Production schedule was applied to railiance ConfigMap and Temporal (weekdays 07:30 Europe/Berlin).
2.7 KiB
2.7 KiB
Open-weight reserve status
As of: 2026-08-03
Store: /mnt/d/vault/coulomb/freedom-intelligence/ (D:\vault\coulomb\freedom-intelligence\)
Soft quota: 850 GiB · Hard stop: 920 GiB
Portfolio snapshot
| Class | Count | Status mix |
|---|---|---|
| Catalog entries | 13 | + Kimi K3 S1 |
| verified on disk | 4 | R embeds + Qwen3-8B + R1-Distill-14B |
| approved (pull deferred) | S giants + Llama edge + Kimi K3 | capacity / HF gate |
| candidate (W) | 3 | headroom-gated |
Disk use (model tree)
| Path | Role | Approx |
|---|---|---|
models/nomic-ai__nomic-embed-text-v1.5/main |
R embed | ~0.5 GiB verified |
models/BAAI__bge-m3/main |
R embed | ~2.1 GiB verified |
models/Qwen__Qwen3-8B/main |
R instruct | ~15.3 GiB verified |
models/deepseek-ai__DeepSeek-R1-Distill-Qwen-14B/main |
R reason | ~27.5 GiB verified |
~45 GiB used → ~800+ GiB soft-quota headroom (insufficient for full Kimi K3 ~1454 GiB).
VAULT free (2026-08-03): ~1.7 TB free of 1.9 TB — full K3 would still dominate the volume and violate soft quota.
Decision matrix (2026-08-03)
| Id | Tier | Decision | Disk |
|---|---|---|---|
| Qwen3-8B | R | verified | yes |
| Llama-3.2-3B-Instruct | R | approved | deferred — HF gate |
| BGE-M3 | R | verified | yes |
| R1-Distill-Qwen-14B | R | verified | yes |
| nomic-embed-text-v1.5 | R | verified | yes |
| Kimi-K3 | S1 | approved, collection deferred | no — ~1454 GiB > soft quota |
| DeepSeek-V3 | S1→secondary | approved, deferred | no — K3 is primary open SOTA identity |
| DeepSeek-R1 full | S2 | approved, deferred | no |
| Qwen3-72B | S4 | approved, deferred | no |
| Llama-3.3-70B-Instruct | S4 | approved, deferred | no |
| Qwen3-14B | W | candidate | no |
| R1-Distill-32B | W | candidate | no |
| bge-reranker-v2-m3 | W | candidate | no |
Hardware coverage gaps
| Need | Status |
|---|---|
| Default local instruct (8B) | Covered (Qwen3-8B) |
| Edge micro-agent | Llama-3.2-3B blocked on HF auth |
| Multilingual RAG embed | Covered (BGE-M3 + nomic) |
| Local reason | Covered (R1-Distill-14B) |
| Frontier-open general MoE offline | Cataloged Kimi K3 — pull blocked on capacity |
| Full open reasoner MoE | Deferred S2 |
Next pulls (operator)
- Capacity decision for Kimi K3 (raise soft quota / dedicate disk / official compressed only).
- Set
HF_TOKENand pull Llama-3.2-3B-Instruct. - Optional S4: one of Qwen3-72B vs Llama-3.3-70B under remaining quota.
- DeepSeek V4-Flash-0731 as mid open value candidate (catalog later if desired).
Tool: scripts/collect_model.py (weights-only, sequential, writes MANIFEST.json).