freedom-intelligence/inventory/catalog/deepseek-ai__DeepSeek-V3__strategic.yaml

92 lines
3.9 KiB
YAML
Raw Normal View History

id: deepseek-ai__DeepSeek-V3__strategic
status: approved
name: DeepSeek-V3
org: deepseek-ai
source:
kind: huggingface
url: https://huggingface.co/deepseek-ai/DeepSeek-V3
revision: main
model_card_url: https://huggingface.co/deepseek-ai/DeepSeek-V3
project_url: https://github.com/deepseek-ai/DeepSeek-V3
license:
spdx: MIT
url: https://huggingface.co/deepseek-ai/DeepSeek-V3
allows_offline_retention: true
allows_local_ops: true
allows_fine_tune: true
notes: "Confirm active card license at pin; prefer latest V3.x/V4 open successor if it is the new SOTA open general."
profile:
summary: "Strategic open general MoE — among the most capable open chat/code weights."
original_source: https://huggingface.co/deepseek-ai/DeepSeek-V3
use_cases:
- "Long-term offline optionality if frontier API access is lost or restricted"
- "Future multi-GPU / training-facility inference or continued pretrain"
- "Capability benchmark reference against closed frontier"
- "Source of student weights / distill experiments (when license + hardware allow)"
sweet_spots:
- "S1 primary giant: best open general MoE class on the reserve"
- "Capability identity even when only compressed weights fit disk"
- "MIT-friendly open SOTA lineage (confirm at pin)"
not_ideal_for:
- "Current lab single-GPU daily ops (use R spine)"
- "Storing multiple near-duplicate full-precision giants on one disk"
- "Assuming immediate local serve without major hardware upgrade"
capability_notes: >
Flagship open general MoE (or successor) for the strategic tier. On ~12 TB
lab bulk storage, prefer official compressed distributions when full precision
would exhaust soft quota. Catalog identity stays the base model id + revision.
swot:
strengths:
- "Top-tier open general capability for the era"
- "Open weights enable true offline retention and future fine work"
- "Strong coding/agent relevance as hardware catches up"
weaknesses:
- "Huge disk and multi-GPU serve cost"
- "Rapid version churn (V3 → V3.x → V4)"
- "Quant choice materially affects quality"
opportunities:
- "Hold compressed primary; upgrade hardware later without re-scrape risk"
- "Anchor daily briefs against a fixed offline SOTA open baseline"
threats:
- "Soft quota: one full pull can block other S niches"
- "License or card terms change at successor releases — re-check at pin"
size:
total_bytes: 0
total_human: "hundreds of GiB (quant) to multi-hundred+ GiB; measure at pull — may need official compressed release"
hardware_class:
min_vram_gb_q4: 0
min_vram_gb_fp16: 0
notes: "Beyond current lab run envelope (T4+). Reserved for strategic capability, not current inference."
axes: [A, B, C]
priority: high
collection:
approved_by: "bernd"
approved_at: "2026-07-28"
downloaded_at: null
downloaded_by: ""
storage_path: ""
brief_refs:
- research/2026-07-24-nas-strategic-collection-plan.md
- docs/decisions/2026-07-24-nas-strategic-reserve.md
reason: "S1 strategic — among the most capable open general MoE weights; keep even if unrunnable today."
tags: [tier-s, strategic, beyond-run-envelope, moe, general]
companions: []
notes: "At download: pin exact revision; if full precision exceeds remaining soft quota, collect best official compressed distribution of the same model line."
history:
- at: "2026-07-24"
event: nominated
by: operator-policy
detail: "1TB NAS strategic reserve — capability-first, not run-envelope-gated."
- at: "2026-07-28"
event: profile_swot_added
by: grok
detail: "schema 0.2 profile + SWOT."
- at: "2026-07-28"
event: approved
by: bernd
detail: "S1 primary giant — deferred full pull (capacity; prefer official compressed later)"
- at: "2026-07-28"
event: collection_deferred
by: grok
detail: "S1 primary giant — deferred full pull (capacity; prefer official compressed later)"