freedom-intelligence/inventory/catalog/meta-llama__Llama-3.3-70B-Instruct__strategic.yaml
tegwick 77d4ffe86e Finish FI-WP-0002 and FI-WP-0003: daily rhythm and R-spine reserve.
Enable fi-daily-research-brief, prove fi_brief_status idempotence with the
2026-07-28 brief, approve R/S catalog entries, collect and verify embeds plus
Qwen3-8B and R1-Distill-14B on VAULT, and document residuals (HF-gated Llama,
deferred S giants, railiance ConfigMap apply).
2026-07-28 01:30:55 +02:00

92 lines
3.5 KiB
YAML

id: meta-llama__Llama-3.3-70B-Instruct__strategic
status: approved
name: Llama-3.3-70B-Instruct
org: meta-llama
source:
kind: huggingface
url: https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct
revision: main
model_card_url: https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct
project_url: https://www.llama.com/
license:
spdx: custom
url: https://ai.meta.com/llama/license/
allows_offline_retention: true
allows_local_ops: true
allows_fine_tune: true
notes: "Llama Community License — review before commercial redistribution. Prefer Llama 4 open text sibling if it is clearly stronger at pull time."
profile:
summary: "Strategic dense Llama 70B instruct — ecosystem-rich open reserve for multi-GPU."
original_source: https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct
use_cases:
- "Future multi-GPU general instruct with mature Llama tooling"
- "Portable dense alternative to giant MoE for lab upgrades"
- "Ecosystem baselines (vLLM, llama.cpp, eval harnesses expect Llama ids)"
- "Domain FT when Llama license fits the use"
sweet_spots:
- "S4 dense open with the broadest third-party tooling surface"
- "More portable serve path than 600B-class MoE"
- "Strong general instruct heritage in the 70B class"
not_ideal_for:
- "Current single-GPU daily ops (use 3B/8B R spine)"
- "Uses forbidden by Llama Community License terms"
- "Duplicate with Qwen-72B under tight quota — pick one first"
capability_notes: >
HF gated. Strategic dense pick. If Llama 4 open text weights supersede
clearly, nominate successor and mark this entry superseded. License is not
Apache — operator must re-read terms before redistribution.
swot:
strengths:
- "Huge ecosystem and ops familiarity"
- "Solid 70B dense instruct quality"
- "Easier multi-GPU dense story than MoE giants"
weaknesses:
- "Custom Llama license (not Apache/MIT)"
- "HF gating friction for automated pulls"
- "May lag peer open MoE on some benches"
opportunities:
- "First dense strategic if Qwen-72B is deferred"
- "Broad eval comparability with published Llama numbers"
threats:
- "Llama 4 open may obsolete 3.3 quickly"
- "Quota competition with Qwen-72B and S1 MoE"
size:
total_bytes: 0
total_human: "~140 GB fp16 / ~40 GB Q4 (estimate)"
hardware_class:
min_vram_gb_q4: 40
min_vram_gb_fp16: 140
notes: "T3+ to run well; still valuable reserve for future multi-GPU."
axes: [A, B, C]
priority: high
collection:
approved_by: "bernd"
approved_at: "2026-07-28"
downloaded_at: null
downloaded_by: ""
storage_path: ""
brief_refs:
- research/2026-07-24-nas-strategic-collection-plan.md
- docs/decisions/2026-07-24-nas-strategic-reserve.md
reason: "S4 strategic dense — strong ecosystem 70B open instruct; more portable than full MoE for a future lab upgrade."
tags: [tier-s, strategic, dense, instruct, llama]
companions:
- meta-llama__Llama-3.2-3B-Instruct__candidate
notes: "HF gated. If Llama 4 open weights supersede, nominate successor and supersede this entry."
history:
- at: "2026-07-24"
event: nominated
by: operator-policy
detail: "Strategic dense open for NAS capability reserve."
- at: "2026-07-28"
event: profile_swot_added
by: grok
detail: "schema 0.2 profile + SWOT."
- at: "2026-07-28"
event: approved
by: bernd
detail: "S4 dense Llama — deferred; gated + quota vs Qwen-72B"
- at: "2026-07-28"
event: collection_deferred
by: grok
detail: "S4 dense Llama — deferred; gated + quota vs Qwen-72B"