pipelines/PIPELINE_STACK_SYNTHESIS_2026-07-29.md
Synthesized 2026-07-29 from PIPE_ART · PIPE_3D · PIPE_ANIMATION · PIPE_VIDEO · PIPE_AUDIO_MUSIC ·
PIPE_VOICE · PIPE_WEBSITE · PIPE_CONTROL_PLANE. This document folds and flags; it does not
re-adjudicate a lane's own verdict, and §5 deliberately leaves conflicts open for the director.
Evidence tiering is carried through from the source dossiers and never upgraded here:
[V] = a primary source was fetched by the owning lane · [E] = arithmetic from a published
param count · [U] = honest unknown, recorded as such. A size or claim never gains confidence by
being copied into this file.
---
| # | Name as quoted | Verdict | What it actually is | Source | Owning lane |
|---|---|---|---|---|---|
| 1 | NVIDIA Cosmos 3 | VERIFIED REAL · EXCLUDED EVERYWHERE | Open frontier world-foundation model for physical AI; Nano 16B licence OpenMDW-1.1, "ready for commercial and non-commercial use" — the cleanest licence in the whole survey. Excluded on three independent grounds: >29GB BF16 on a 32GB card (needs enable_sequential_cpu_offload to boot), no ComfyUI support (Comfy-Org issue #14228, opened 2026-06-02, still open, zero maintainer comments), and a physical-AI/robotics training objective, not a creative one. Emits video/image/audio/action — not meshes or USD, so it is not a worldgen answer either | huggingface.co/nvidia/Cosmos3-Nano · github.com/Comfy-Org/ComfyUI/issues/14228 · nvidianews.nvidia.com launch | VIDEO (verdict), 3D (concurs), CONTROL_PLANE (routing row: REAL, NOT ROUTED) |
| 2 | "ARDY" | VERIFIED REAL · WRONG LANE (not image/audio/video) | NVIDIA autoregressive-diffusion human MOTION generation, SIGGRAPH 2026, checkpoints 2026-07-10. 326M params, 27-joint "Core" skeleton, 20fps, 40-frame horizon (8s max clip), NPZ out. Code Apache-2.0, weights NVIDIA Open Model License (commercial permitted, NVIDIA claims no ownership of outputs) | github.com/nv-tlabs/ardy · huggingface.co/nvidia/ARDY-Core-RP-20FPS-Horizon40 · research.nvidia.com/labs/sil/projects/ardy/ · arxiv.org/abs/2607.08741 | ANIMATION (claims it eval-only) vs 3D + CONTROL_PLANE (pin it primary) — see conflict C-3 |
| 2b | ARDY's hidden rider | VERIFIED — a Josh-gate nobody else recorded | ARDY's text encoder is meta-llama/Meta-Llama-3-8B-Instruct — HF-gated, Meta Llama 3 Community License, ~14-16GB bf16. The repo's standing anti-Llama posture (700M-MAU + attribution + policy propagation) is what makes this a ruling, not a download | ANIMATION §1 verified specifics | ANIMATION — see conflicts C-3, C-4 |
| 3 | "Kimodo" | VERIFIED REAL · WRONG LANE · LICENCE SPLITS PER VARIANT | NVIDIA kinematic motion diffusion, open-sourced 2026-03-16, v1.1 2026-04-10. Skeletons SOMA (77-joint) / SMPL-X / Unitree G1. Trained on commercially licensed studio mocap ("Bones Rigplay 1", 630-700h) — the single reason a local text-to-motion lane is legally viable at all. Code Apache-2.0; **weights split: SOMA + G1 = NVIDIA Open Model License, Kimodo-SMPLX-RP-v1 = NVIDIA *R&D* Model License → previz-only. VRAM ~17GB full-GPU, <3GB with TEXT_ENCODER_DEVICE=cpu**. No BVH/FBX export in-repo | github.com/nv-tlabs/kimodo · research.nvidia.com/labs/sil/projects/kimodo/ | ANIMATION (primary) vs 3D (fallback) vs CONTROL_PLANE (constrained-motion row) — C-3 |
| 4 | ACE-Step 1.5 | VERIFIED REAL · ADOPTED (music lane) | github.com/ace-step/ACE-Step-1.5, MIT, 600s cap. Repo-naming trap recorded: ace-step/ACE-Step (v1) is a *different repo* under Apache-2.0 with a 4-min cap. Pin the URL, not the name | github.com/ace-step/ACE-Step-1.5 (LICENSE fetched: "MIT License", © ACEStep 2026) | AUDIO_MUSIC |
| 4b | "ACE-Step 1.5 UI" | VERIFIED | Official Gradio UI (uv run acestep) and a REST API server (uv run acestep-api, port 8001) ship in-repo | ACE-Step-1.5 README | AUDIO_MUSIC — port 8001 is missing from the Stage-6F firewall table, see C-10 |
| 5 | "XL SFT" | VERIFIED | ACE-Step/acestep-v15-xl-sft — 4B DiT, 50 steps, licence field mit, ~20.0GB on disk (4 × 4.99GB shards) [V] | huggingface.co/ACE-Step/acestep-v15-xl-sft | AUDIO_MUSIC |
| 6 | "XL Turbo" | VERIFIED | acestep-v15-xl-turbo — 8 inference steps vs SFT's 50, ≈6× faster | huggingface.co/ACE-Step | AUDIO_MUSIC |
| 7 | "Excel Bass" | NAME REFUTED · REFERENT VERIFIED | A transcription artifact of acestep-v15-xl-base, the real third variant (pre-trained only). The mis-heard string never enters a config, a prompt, or a registry cell — the corrected token is xl-base. *Three dossiers reached this by different routes and labelled it differently — see C-8* | huggingface.co/ACE-Step | AUDIO_MUSIC / CONTROL_PLANE |
| 8 | "XL SFT / XL Turbo / Excel Bass as ComfyUI audio NODES" | REFUTED | These are checkpoints, not nodes. The real ComfyUI ACE-Step nodes are EmptyAceStepLatentAudio, TextEncodeAceStepAudio, LatentOperationTonemapReinhard | docs.comfy.org/tutorials/audio/ace-step/ace-step-v1 | AUDIO_MUSIC |
| 8b | Same names, read as TTS/voice models | UNVERIFIED-EXCLUDED in the voice lane | No TTS model or ComfyUI voice node by any of these names exists in any source found. The voice lane correctly refused to adopt an unverifiable name; the audio lane found the real referent one repo over | PIPE_VOICE §1 | VOICE |
| 9 | DeepSeek (LLM lead) | VERIFIED REAL · NOT A LOCAL CANDIDATE · already repo-adjudicated | Flagship line is out of local range (V4 ≈1T params, 2026). The only in-range members are the R1 distills (1.5B/7B/14B/32B), and their licence splits by base model: Qwen-based distills inherit Apache-2.0, Llama-based distills inherit the Llama licence (MAU clause). Separately, **DeepSeek-OCR is a ruled *class* target** for the VLM-OCR seam (D-6), not a pick | docs/pipeline_review/tech_research/Q3_2026_MODELS_REFRESH.md:357,382,390,503 · docs/5090_SETUP_RUNBOOK.md:843,939,1512 · DeepSeek API docs · release timeline | NO DOSSIER — owned by D-4 (14B pick) / D-6 (VLM-OCR), both open by design |
| 10 | GLM (LLM lead) | VERIFIED REAL · EXCLUDED ON SIZE | GLM-5.2 (Zhipu/zai-org, released 2026-06-13): 753B MoE / 40B active, 1M-token context, MIT licence, no regional limits. The licence is excellent and irrelevant — a 40B-active MoE is not a 32GB-card candidate, which is exactly what the repo's Q3 refresh already concluded ("Large MoE — not a 32GB-card candidate regardless of license") | docs/.../Q3_2026_MODELS_REFRESH.md:359 · GLM-5.2 coverage · release note | NO DOSSIER — D-4 |
Two honesty notes on this table. (a) Rows 9-10 were the user's leads that no dossier covered;
they are answered here from the repo's own prior verification plus one live confirmation pass, and
both land as *real but not adoptable on this hardware* — which is why the 14B pick (D-4) stays open by
design rather than being closed by a search-AI name. (b) Every "wrong lane" verdict above was reached
independently by 3-4 separate dossiers, which is the strongest signal in the survey: **the search AI
conflated a motion model, a music model and a world model into the image/audio stack, and four
lanes caught it separately.**
---
Ruled layout: Gen5 4TB slot 1 (C:) = Windows + UE engine hot path + DDC · Gen4 4TB slot 3 (D:)
= models / ComfyUI / projects cache · 8TB SATA slot 4 (E:) = archive.
→ Gen4 (D:) — all model weights, all venvs, the WSL2 vhdx
| # | Item | GB | Licence | Lane | Tier |
|---|---|---|---|---|---|
| 1 | Puget pack base + comfy_ui image (nvidia/cuda:12.8.0-runtime-ubuntu24.04) | ~15 [E] | Apache/vendor | 3D (+VIDEO) | 1 |
| 2 | cu128 torch/cuDNN wheel set (shared pip cache, ~3-5 per fresh venv) | ~5 [E] | BSD | ALL | 1 |
| 3 | hidream_o1_image_dev_fp8_scaled + gemma4_e4b_it_fp8_scaled + rank-64 LoRA | ~13-15 [E] | MIT (⚠ Gemma-4 text-encoder licence tag unresolved — pre-flight) | ART | 1 |
| 4 | z_image_turbo_bf16 + qwen_3_4b + ae | ~20 [E] | Apache-2.0 | ART | 1 |
| 5 | Depth-Anything-3 — Apache variants ONLY, pinned by filename | ~1-2 [E] | Apache-2.0 (⚠ Giant + Nested are CC BY-NC 4.0) | ART | 1 |
| 6 | SeedVR2 3B (NVFP4 + fp16) — *one pull, two consumers (ART upscale + VIDEO upscale), see C-7* | ~3-7 [E] | Apache-2.0 | ART+VIDEO | 1 |
| 7 | acestep-v15-xl-sft (4 × 4.99GB shards) | 20.0 [V] | MIT | AUDIO | 1 |
| 8 | MOSS-SoundEffect-v2.0 (1.3B, 48kHz) | 11.2 [V] | Apache-2.0 | AUDIO | 1 |
| 9 | Kokoro-82M + Piper voices ×5 | 0.7 [E] | Apache-2.0 / MIT | VOICE | 1 |
| 10 | Qwen3-TTS VoiceDesign + CustomVoice + 0.6B-Base + tokenizer | ~11.7 [E] | Apache-2.0 | VOICE | 1 |
| 11 | Chatterbox-Turbo (350M) + Multilingual (500M) | ~1.8 [E] | MIT | VOICE | 1 |
| 12 | Kimodo repo + SOMA checkpoints only (-SOMA-RP-v1.1, -SEED-v1.1) | 3-6 [U] — README publishes no sizes | Apache-2.0 code / NVIDIA Open Model weights | ANIMATION | 1 |
| 13 | TRELLIS.2-4B | ~8 [E] | MIT (geometry path) | 3D | 1 |
| 14 | Direct3D-S2 | ~5 [E] | MIT | 3D | 1 |
| 15 | Wan 2.2 (T2V + I2V, fp8 where offered) | ~30-60 [E] | Apache-2.0 | VIDEO | 1 |
| 16 | Qwen2.5-14B-Instruct-Q4_K_M.gguf — LAN-copy from the Ally, never re-download, sha256 pinned e47ad95d…c008, VERIFY after copy | ~9 [V] | Apache-2.0 | CONTROL_PLANE | 1 |
| 17 | Qwen2.5-14B Q5_K_M sibling (one HF pull) | ~10.3 [V] | Apache-2.0 | CONTROL_PLANE | 1 |
| — | TIER 1 SUBTOTAL (first light + the licence-clean baseline of every lane) | ≈ 170-200 | |||
| 18 | qwen_image_fp8_e4m3fn.safetensors | 20.4 [V] | Apache-2.0 | ART | 2 |
| 19 | qwen_2.5_vl_7b_fp8_scaled + Qwen-Image-Edit-2511 fp8 | ~28 [E] | Apache-2.0 | ART | 2 |
| 20 | ControlNet-Union (Qwen + Z-Image) | ~5 [E] | Apache-2.0 | ART | 2 |
| 21 | FLUX.2-klein-4B (the only shippable FLUX) | ~9 [E] | Apache-2.0 | ART | 2 |
| 22 | acestep-v15-xl-turbo + acestep-5Hz-lm-4B planner + ComfyUI ace_step_1.5_turbo_aio | ~35 [E] | MIT | AUDIO | 2 |
| 23 | Hunyuan3D-2.1 shape+paint | ~11 [E] | Tencent Community — 1M MAU, EU/UK/KR excluded ⚠ | 3D | 2 (previz-gated, C-5) |
| 24 | HunyuanVideo-1.5 (8.3B + super-res) | ~18-22 [E] | Tencent Community — 100M MAU, EU/UK/KR excluded ⚠ | VIDEO | 2 (previz-only, permanent) |
| 25 | LTX-2.3-22b-distilled fp8 + x2 upscalers | ~22-26 [E] | LTX-2 Community — $10M revenue ceiling + AI-disclosure clause ⚠ | VIDEO | 2 (previz-only) |
| 26 | Shared video text encoders / VAEs | ~10-20 [E] | mixed | VIDEO | 2 |
| 27 | Hi3DGen + TripoSG + UniRig + PartCrafter | ~12 [E] | MIT | 3D | 2 |
| 28 | ARDY + ARDY-Core-RP-20FPS-Horizon40/8 | ~1.5 [E] | Apache-2.0 code / NVIDIA Open Model weights | ANIMATION/3D | 2 |
| 29 | HY-World 2.0 — WorldMirror-2 ONLY. Do NOT queue blind: WorldStereo-2 ≈17B, HY-Pano-2 ≈80B [V] is far outside a 32GB card | ~3 [E] | License.txt name NOT READABLE — licence read REQUIRED before first GPU hour ⚠ | 3D | 2 |
| 30 | Higgs Audio V2 3B + IndexTTS-2 (bench only) | ~9 [E] | Apache-2.0 | VOICE | 3 |
| 31 | meta-llama/Meta-Llama-3-8B-Instruct — GATED, needs Josh's HF acceptance; pull ONLY if ARDY is adopted | ~16 [E] | Llama 3 Community (700M MAU) ⚠ | ANIMATION | 3 (conditional) |
| — | FULL ROSTER TOTAL (excl. the optional 200GB Sonniss archive) | ≈ 350-450 | fits 4TB easily |
→ 8TB SATA (E:) — archive-class pulls
| Sonniss GDC 2026 bundle (347 WAV) | 7.47 [V] | royalty-free-owned, no attribution; licence PROHIBITS AI/ML training | AUDIO |
|---|---|---|---|
| Sonniss historical archive (archive.org) | ~200 [E] | same | AUDIO — optional, overnight, lowest priority |
| DEM / WorldCover GeoTIFFs | rate-limited, not size-limited | open data | 3D — *archive once, see §4 item 8* |
→ Gen5 (C:) — npm/tooling only (npm_config_cache redirected to D:)
Astro 7 + sitemap + sharp node_modules | 0.3-0.5 [E] | MIT | WEBSITE |
|---|---|---|---|
| Playwright browser binaries (3 engines + deps) | 1.0-1.5 [E] | Apache-2.0 | WEBSITE |
| wrangler v4 | ~0.1 [E] | Apache-2.0/MIT | WEBSITE |
| Blender 5.2 LTS · AutoRemesher · xatlas · Cascadeur 2026.1 · UE Game Animation Sample | ~0.5 + 1.5 + 4-8 [U] | GPL tool (output unencumbered) / MIT / paid seat / UE-Only Content | 3D + ANIMATION |
Deduplication actually performed: SeedVR2 (claimed twice: ART + VIDEO), ARDY+Kimodo (claimed twice:
3D + ANIMATION), Blender (3D + ANIMATION), the cu128 torch wheel set (7 lanes → one pip cache), and
the ComfyUI weight tree (4 lanes → unresolved, see C-1). Naive per-dossier summation gives
≈420-520GB; the deduped figure is ≈350-450GB.
The control plane's Thursday deliverable is 0 GB and is the most important item on the list: the
download manifest itself — one JSON row per artifact carrying url, bytes, sha256, dest,
licence, so an interrupted overnight pull is *verifiable* rather than re-downloaded, and **a row
with no bytes/sha256 is rejected at write time.** Every [E] and [U] above is a cell that must
be corrected with a measured value on the night — the runbook already carries three
UNSOURCED-ESTIMATE VRAM rows and this list must not add a fourth class of them.
**D-1 (runbook:1507) offers WSL2+CUDA vs ComfyUI-on-Windows vs dual-boot, and its own recommendation
is the split: "WSL2 for official-path models + ComfyUI-Windows for the image lane."** Every dossier
converged on a *per-lane* answer rather than a machine-wide one:
| Lane | D-1 assignment | Why | Home |
|---|---|---|---|
| CONTROL_PLANE | Outside every containment boundary — Windows-native, not WSL2, not the container, not a per-model venv | must survive any lane's environment churn and be able to restart what it supervises; speaks localhost HTTP only | D:\pipe\ |
| WEBSITE | D-1 does not bind — native Windows, no CUDA, no torch, no container | stated explicitly so nobody containerizes a static-site build for symmetry | C:\dev\humanity-site\ ⚠ C-2 |
| ART | (b) ComfyUI-on-Windows, native | every pick is diffusers/ComfyUI-native; Krita AI Diffusion is a *Windows* app driving a ComfyUI backend — a WSL2 vhdx hop at the live-painting loop is the worst possible placement | D:\ComfyUI\ |
| VOICE | (b) — rides the ART ComfyUI via TTS-Audio-Suite v5.6.0 (MIT), plus a tiny CPU-only .venv-tts-cpu | all picks are pure-PyTorch Windows-native | D:\ComfyUI\models\TTS\ + D:\tts\out\ |
| AUDIO | (b) — "removes the audio lane from the D-1 fork entirely" | ACE-Step ships a Windows-runnable Gradio+REST server; MOSS installs from a cu128 wheel | D:\ai\acestep\, D:\ai\moss-sfx\ |
| ANIMATION | native-Windows per-lane venvs — and explicitly NOT folded into the Stage-6b ComfyUI | needs none of what forces WSL2 (no nvdiffrast, no ComfyUI); Docker is the sanctioned fallback if MSVC fights the CMake/C++17 extensions | D:\ai\anim\{kimodo,ardy} |
| VIDEO | (a) WSL2 + Puget comfy_ui — but notes (b) is "genuinely viable here" | all four primaries have first-class ComfyUI paths | Gen4, ⚠ C-1 |
| 3D | (a) WSL2 + Puget pack, and asks Josh to ratify it machine-wide | the only lane with a hard forcing function — see the headline below | Gen4 (vhdx + weights + staging) |
THE HEADLINE THAT RESHAPES THE WEEK — the 3D lane does not install as published. Every 3D-gen repo
pins a pre-Blackwell CUDA (TRELLIS.2 torch 2.6.0 + cu124, Hunyuan3D-2.1 2.5.1+cu124, Direct3D-S2
2.5.1+cu121, ComfyUI-3D-Pack Windows prebuilds cu124/torch-2.5.1 with zero 50-series note),
while the RTX 5090 is sm_120 and requires ≥ PyTorch 2.7.0 + cu128. So the install plan for that
lane is not "follow the README" — it is "rebuild every CUDA extension against cu128"
(flash-attn, spconv, torch_scatter, torch_cluster, torchsparse, and TRELLIS.2's
cumesh/o-voxel/flexgemm). **This is a half-day-to-two-day item and it is the single biggest
unbudgeted cost in the whole factory setup.** The runbook's cu124 pin for TRELLIS.2 is simultaneously
CONFIRMED as the repo's published requirement and CONFIRMED as unrunnable on sm_120 — read it as
"the version the authors tested," never "the version we install."
Install order (ascending risk, so the box is productive before the hard lane starts):
write the download manifest · Stage-6F firewall pre-authorization including 8188 AND 8001
(an unacknowledged Defender dialog at 04:00 presents as a hung editor, not an error — the worst
overnight signature in the house) · kick the Tier-1 pulls.
ART, VIDEO, VOICE and CONTROL_PLANE each independently scheduled this as a hard pre-flight; fold
them into one sitting: HiDream-O1 MIT · Qwen Apache · Z-Image Apache · SeedVR2 Apache ·
Depth-Anything-3 per-variant · Gemma-4's apache-2.0-tag-vs-Gemma-licence-link discrepancy
(the entire HiDream text-encoder chain depends on it) · HY-World's unreadable License.txt ·
Kimodo per-variant · every upscale weight against the allowlist. (b) **The Batch-0 registry
backfill** — see §2.4; it is the head-of-chain for every lane's dispatch and costs zero GPU hours.
torch==2.9.0+cu128, then **verify sm_120 before anything else**: python -c "import torch;print(torch.cuda.get_device_capability())" must return (12, 0).
Nothing downstream is trustworthy until that prints.
current machine · (3b) CONTROL_PLANE — D:\pipe\.venv, requests websocket-client psutil,
lanes.json, routing_table.json, the two scheduled tasks · (3c) VOICE CPU sub-lane — 10 minutes,
zero GPU, and the first proof the box generates anything at all · (3d) ART — ComfyUI Windows
(pin a cu128 wheel; the portable README currently ships cu13.0 and D-2 says avoid 13.x) →
Manager → Z-Image first light (smallest/fastest/Apache) → HiDream-O1 → Depth-Anything-3 →
SeedVR2 → Krita · (3e) VOICE GPU sub-lane — TTS-Audio-Suite → Qwen3-TTS → Chatterbox
gate-locked clone_enabled=false by default · (3f) AUDIO — ACE-Step venv → Gradio → REST :8001
→ MOSS → Demucs · (3g) ANIMATION — Kimodo venv → SOMA ckpts → **write Tools/npz_to_bvh.py
ourselves** (~1 file, licence-clean, ENGINE-class, removes both unverified third-party bridges) →
Blender bake → UE import → *only then* ARDY as an eval lane · (3h) VIDEO — Wan 2.2 first (the
licence-clean baseline) → SeedVR2 → HunyuanVideo-1.5 (the co-residency probe) → LTX-2.3 last,
because it is the one that may not fit and failing there costs nothing · (3i) 3D LAST — WSL2 +
Puget pack + Container Toolkit → cu128 base → the extension rebuild → TRELLIS.2 geometry-only
(setup.sh WITHOUT --nvdiffrast --nvdiffrec) → Direct3D-S2 → Hunyuan3D-2.1 → the CPU finishing
set → UniRig → HY-World → ARDY/Kimodo.
*The ordering principle, stated because it inverts importance:* 3D is the most valuable lane and is
installed last, because it is the only one whose install can consume two days, and every other
lane is independently useful on day one. A factory that spends Friday on flash-attn has nothing
running Friday night.
| Knob | THE PIN | What it refuses, and why the refusal is load-bearing |
|---|---|---|
| CUDA / torch | torch==2.9.0+cu128 everywhere (≥2.7.0 is the first release shipping sm_120 kernels) | refuses cu124 (TRELLIS.2, Hunyuan3D-2.1, ComfyUI-3D-Pack prebuilds), cu121 (Direct3D-S2), and cu13.0 (ComfyUI portable's current default — D-2 says avoid 13.x on this machine). Every one of those is a published requirement that will not load on sm_120 |
| Python | 3.11 for the 3D base + animation venvs; 3.12 for the control plane and ComfyUI-Windows (TTS-Audio-Suite requires 3.12+); 3.13 is Blender 5.2's internal and not ours to pick; Node ≥22.12 for the website (box has 24.15.0 ✓) | per-venv pins are compatible *only because* the lanes are isolated — a single shared env cannot satisfy 3.11 and 3.12 at once |
| Ports | ComfyUI 8188 · ACE-Step REST 8001 · MCP 9315 / 8000 · Ollama 11434 | 8001 is absent from the Stage-6F table (C-10); the website dev-server port is undeclared by its dossier |
| Licence classes | SHIPPABLE = MIT · Apache-2.0 · BSD · OpenRAIL++ · NVIDIA Open Model. PREVIZ_ONLY = any MAU cap, territory exclusion, revenue ceiling, or NC clause. EXCLUDED = research-only or unverifiable | movement between states requires a re-fetched licence document — never a memory, a blog post, or a model-card summary. Today's fetches found the FLUX licence text itself behind HTTP 401 and multiple 2026 blogs calling territory-excluded HunyuanVideo "Apache 2.0" |
| Upscaler allowlist | models/upscale_models/ is never trusted by default | ComfyUI's community-default upscaler 4x-UltraSharp is CC-BY-NC-SA-4.0 — non-commercial *and* share-alike. It would silently taint every plate it touched |
| Depth-Anything-3 | pin the Apache variants BY FILENAME | Giant + Nested are CC BY-NC 4.0 in the same repo |
| Kimodo | per-VARIANT licence read before first GPU hour | SOMA + G1 = Open Model (fine); Kimodo-SMPLX-RP-v1 = R&D licence → previz-only. Same repo, two answers |
| SMPL/SMPL-X anywhere | banned from any shipped-asset path | the Max Planck model licence is non-commercial-research; commercial routes through Meshcapade. This also kills the entire open text-to-motion literature (MDM/MoMask/MotionGPT/T2M-GPT → HumanML3D → AMASS) — the only reason a local motion lane exists at all is NVIDIA's commercially-licensed studio mocap |
| nvdiffrast / nvdiffrec | BANNED from any shipped path | NVIDIA 1-Way Commercial, non-commercial-only. TRELLIS.2's geometry path is MIT and clean; its textured/bake path is previz-only, permanently, by licence — a constant, not a benchmark output. Issue #22 re-fetched today: still open, still no maintainer response |
| 3DGS | route the whole lane through gsplat/Nerfstudio (Apache-2.0) | the original Inria graphdeco-inria/gaussian-splatting reference impl is non-commercial, and SuGaR derives from it |
✅ CLOSED 2026-08-02 — BATCH 0 APPLIED, gates green (36/36, exit 0). The 2026-07-29 numbers
below are kept as the dated record; the tree had MOVED by the time the work ran, and the honest
re-count is: 863 rows carried generation_status and 826 were the empty string, with the
dispatcher predicate returning 23 work orders, not 21. The delta is real movement, not a
miscount — T0_SFX_Registry has since GAINED all four assetgen status columns (so the "no
generation_status column at all" line below is stale; +30 rows carrying the column, all
empty), and T0_Theme_Registry gained two cardinal rows at the motif-roster apply (12 → 14
pending). The 796-empty figure across the eight mesh-bearing registries was still exactly
right. One value outside the enum that this section did not count was also found and handled:
T0_Voice_Registry.GRAND_SAGE_REVEAL_VOICE held the free-text
excluded_non_lexical_HL_0046, so the true residue was 827, not 826.
What landed, all in one commit: 843 generation_status cells written through the sanctioned
writerharness/apply_generation_lifecycle_backfill.py(826 empty →pending, 16 on the
brand-new T0_Scene_Spec_Registry column populated in the same pass, 1 declared old-value-
guarded retag); 17 columns appended across 12 registries including the C-12 generation_licence_ref
everywhere;docs/generation_lifecycle.jsonas the contract;docs/licence_records/+
RECORD_SCHEMA.json+harness/generation_prompt_hash.pyas THE LICENCE SCHEMA COMMIT; and the
vacuity guard as GATE 36 harness/check_generation_lifecycle.py, which is stronger than this
section asked for — it measures THE RESIDUE (rows the dispatcher's predicate vocabulary cannot
name) and fails unless it is zero, so the defect class is unrepresentable rather than merely
fixed. The dispatcher predicate now returns 865 work orders. Ten live mutation tests plus 27
fixtures prove the teeth are armed.
Counted on the live tree: 831 rows carry generation_status and 796 of them are the EMPTY STRING
(Antagonist 64 · Boss 281 · Character 164 · Creature 150 · Equipment 1 · Familiar 22 · Vril_Site 42 ·
Weapon 72). Empty string is outside the ratified five-value enum, so a dispatcher scanning
WHERE generation_status == 'pending' returns **21 work orders across the entire mesh surface and
reports the queue clean** — the "a search that CANNOT match reports zero" class, reproduced inside the
control plane. Plus: T0_SFX_Registry (30 rows) has no generation_status column at all (and
SFX is DR-2 §6 Batch 2, the *handshake exit*), T0_Scene_Spec_Registry (16 rows) carries a prompt
hash with no status column, and T0_Theme_Registry lacks regeneration_trigger.
Batch 0 therefore backfills pending on all 796 rows through the sanctioned writer, adds the missing
assetgen columns to SFX + Scene_Spec, refreshes fidelity_baseline.json in the same commit, and
ships a vacuity guard: if total enrolled rows is 0, or a registry with assetgen columns matches 0
rows, the run exits non-zero. *A queue that cannot match reports zero and looks exactly like a
finished night.* Four lanes independently want schema changes in this same commit — see C-12.
---
Governing arithmetic, inherited verbatim: *"29GB of 32GB is the tight fit the whole hardware decision
turns on … heavy lanes run sequentially, not concurrently"* — and **an OOM caused by our own
co-residency sloppiness would be a FALSE trigger on HARDWARE_DECISION's upgrade trigger #1.** Every
lane below signed up to that rule independently.
CLASS A — ALWAYS-ON · 0 GB VRAM. Never holds the lease; never contends.
| Tenant | VRAM | RAM | Note |
|---|---|---|---|
| CONTROL_PLANE orchestrator | 0 | ~2 GB, ~1% of one core | *issues* leases, never holds one — deliberate: a control plane that competes with UE cannot arbitrate between UE and the lanes |
| WEBSITE build (Astro + sharp AVIF + Playwright ×3 engines) | 0 | 4-8 GB, ~4 cores | ⚠ must NOT run concurrently with a UE cook or lighting build — same CPU/RAM/NVMe contention, same Gen5 drive |
| VOICE bulk/ambient (Kokoro + Piper, CPU) | 0 | ~4 GB | the cheapest tenant on the machine; zero cloning capability by construction |
| 3D CPU finishing half (Blender-headless retopo/UV/bake/export, AutoRemesher, xatlas, GDAL/rasterio, DEM chain, FBX/glTF) | 0 | ~32 GB, ~8 of 32 threads | the lane's real throughput win — the finishing pass is the actual bottleneck (no open AI does retopo+UV+LOD+collision end-to-end) and it never needs the contested resource |
| Demucs · deterministic gates · git · Stage-W watchers | 0 | small | |
| Class A total | 0 GB VRAM | ≈ 46-50 GB RAM | runs *through* every other class |
CLASS B — WINDOWED · the daytime UE window. UE holds the lease; exactly ONE small tenant beside it.
| Tenant | VRAM | Note |
|---|---|---|
| UE editor + capture harness + QA playtest | 12-20 GB | holds the lease whenever a human or director session is live |
| Kimodo CPU-text-encoder mode | <3 GB [V] | the honest concurrency win — text-to-motion co-resident with a real renderer, bought with slower encoding. Should be the default mode |
| Chatterbox Turbo / Multilingual | 2.3 / 3.5 GB [V, NVIDIA-published] | windowed and attended (clone-capable) |
| Qwen3-TTS 1.7B (cultural registers) | ~5-6 GB [E] | |
| ACE-Step 2B-turbo + 0.6B LM | 6-8 GB (README tier) | the iteration lane; the XL model never runs here |
| MOSS-SoundEffect v2.0 | <8 GB [E] — MEASURE day one | the always-available bulk-SFX worker |
| Local 14B triage (interstitial) | ~9 GB | yields *instantly* to any exclusive lease |
| HunyuanVideo-1.5 (offload on) | 14 GB [V, model card] | ⚠ 20 + 14 = 34 > 32. This is a *probe*, not a plan — the only question that could earn the video lane a windowed class instead of overnight-exclusive |
| Class B fit | UE 20 + one ≤8 GB tenant = 28 of 32 ✓ · UE 20 + 14B = 29 of 32 ✓ (the binding line) · UE 20 + HunyuanVideo = 34 ✗ | the rule is ONE small tenant, arbitrated by a single lease file + an nvidia-smi free-VRAM precondition |
CLASS C — OVERNIGHT / EXCLUSIVE. One job at a time, serialized by the queue, never co-resident.
| Job | Peak VRAM | Evidence |
|---|---|---|
| Hunyuan3D-2.1 combined (the ceiling case) | 29 GB [V, repo-quoted] | admits nothing else — not a resident 14B, not a UE editor |
| LTX-2.3-22b distilled fp8 | ~32 GB claimed → assume EXCLUSIVE, may not fit at all | if it fails, Wan 2.2 becomes primary *and* fallback — a materially simpler, fully-Apache lane |
| Qwen-Image-2512 fp8 + Qwen2.5-VL-7B fp8 @1328² | ~26-28 GB [E] (20.4GB file [V]; 86%-of-24GB observation [V]) | single-tenant by itself |
| ACE-Step XL-SFT + 4B LM | ~24-28 GB [E] | the XL music model owns the GPU alone, after the soak |
| TRELLIS.2 | 24 GB floor [V] | |
| Direct3D-S2 @1024 | 24 GB [V] (10 GB @512) | |
| Wan 2.2 A14B fp8 | ~16-22 GB [E] | |
| Kimodo full-GPU | 17 GB [V] | |
| HiDream-O1 fp8 + gemma4 fp8 @2048² | ~14-16 GB [E] | + 32-48 GB RAM for text-encoder CPU offload |
| ARDY (+ Llama-3-8B encoder) | ~14 GB [V] | CPU-offloadable |
| SeedVR2 3B | 3-14 GB [E] (tiling-dependent) | |
| UniRig | 8 GB [V] | |
| Class C rule | ONE at a time. The queue holds while UE holds the GPU. | Cosmos 3 Nano (>29 GB + offload) and HY-Pano-2 (≈80B) are *why* they are not installed |
CLASS D — HARD RESERVATION. The 04:00 soak (D-11b). Asset-gen never overlaps it. Not negotiable.
| Window | GPU (32 GB) | CPU/RAM rail (Class A, continuous) |
|---|---|---|
| 00:00-04:00 | Exclusive slot 1 — one Class-C job (≤29 GB) | control plane · 3D finishing · gates · website builds |
| 04:00-05:00 | SOAK — reserved, GPU quiet | rail continues untouched |
| 05:00-08:00 | Exclusive slot 2 — second Class-C job | rail continues |
| 08:00-20:00 | UE window — UE 12-20 GB + ONE ≤8 GB tenant (+14B interstitial, yields) = ≤29 GB | rail continues, except website builds yield to any UE cook / lighting build |
| 20:00-00:00 | Exclusive slot 3 | rail continues |
**Throughput: 3 heavy GPU jobs per night + a full day of small-tenant generation + an uninterrupted
CPU rail. Each Class-C run carries a hard wall-clock ceiling** so a runaway lane cannot eat the
next day's window, and every run is self-stopping with a 30-second RUN_STATUS.json heartbeat (a beat
older than 10 minutes with a live process opens a stall record, not a crash — because a firewall
dialog and a hung editor look identical from outside).
| Scenario | Composition | Total |
|---|---|---|
| Day peak | UE editor+cook 32-48 · 3D CPU finishing 32 · website 8 · voice CPU 4 · control plane 2 · small-tenant host 8-16 | ≈ 86-110 GB ✓ |
| Night peak | ART text-encoder CPU offload 32-48 · 3D CPU finishing 32 · Class-C host RAM ~16 · control plane 2 | ≈ 82-98 GB ✓ |
Both fit, and the second is *why* the art lane's 32-48 GB offload requirement is affordable: it only
ever peaks in a window where UE is not cooking. Scheduling note, not a conflict: ART's "hold
32-48 GB free" and 3D's "~32 GB for the CPU half" are the two largest RAM claims and they stack
safely only while a UE cook is not also running. If all three coincide the box is at ~130 GB and over.
C:\dev\humanity-site\ (⚠ C-2)**. Four dossiers state "nothing from this lane touches Gen5."
D:\pipe\, ComfyUI models/.Total ≈350-450 GB against 4 TB. Comfortable.
DEM/WorldCover GeoTIFFs (archive once — they are rate-limited, not merely large) ·
E:\anim_gen\ · video reel archive · Sonniss · E:\archive\vo_masters\ + vo_reference\ ·
E:\pipe_archive\ (rejected outputs **kept, never deleted — a rejected asset is the QA gate's
evidence**) · website capture masters + press-kit staging.
---
**Tagged [THU-EVE] = on-box or blocks the Thursday-night pulls · [REMOTE] = anytime, from
anywhere, no box needed · [LATER] = deferred by its own ruling.**
Blocking the setup:
1. [THU-EVE] Claude Code sign-in on the new box. The on-box session owns everything after it.
2. [THU-EVE] HuggingFace account + token, and per-repo gated acceptances. Anonymous fetches of
both FLUX.2-dev/LICENSE.md and the Krea org returned HTTP 401 today. Clicking "accept" is an
account action; agents never handle tokens. *Blocks any gated pull on the night.*
3. [THU-EVE] Fab download of the UE Game Animation Sample (500+ anims) — his Epic account.
*Skip the Fab "Kimodo Importer" — we write the exporter ourselves.*
4. [THU-EVE] Sonniss GDC bundle download — free, but the site is email/form-gated. One sitting,
or delegate.
5. [REMOTE, DO NOW — pre-arrival] The Puget burn-in / benchmark sheet ask to the rep. Requested
independently by four dossiers (ART, 3D, VIDEO, CONTROL_PLANE) — it is the only pre-arrival
thermal/power baseline for this exact machine, and the window-sizing arithmetic wants it.
Rulings only he can make (each already briefed with a recommendation, never a bare question):
6. [REMOTE] D-1 ratification — the 3D lane asks him to confirm WSL2+Puget as the generation host
with ComfyUI-Windows as the image lane. *Note: six other lanes answered D-1 for themselves and
three of those answers are per-lane, not machine-wide (C-9).*
7. [REMOTE] Ship-territory ruling on the Tencent licences (Hunyuan3D-2.1 + HunyuanVideo-1.5 both
exclude the EU, UK and South Korea). Two sub-questions: does the clause block *us using* the
model, block *distributing outputs into* those territories, or both. **Recommendation (control
plane): drop Hunyuan3D from the ship path** — the geometry leg is already covered by MIT models and
a worldwide Steam release is worth more than one alternate mesh backend.
8. [REMOTE] SMPL-X / Meshcapade decision — are the non-SMPL ARDY/Kimodo skeletons sufficient, or
is a Meshcapade commercial licence wanted? *Recommendation: non-SMPL skeletons; retarget joint
rotations onto our own rig and never ship SMPL-X mesh data.*
9. [REMOTE, CONDITIONAL] Llama-3-8B: HF gated acceptance + the anti-Llama-posture ruling —
needed only if ARDY is adopted. *Recommendation: Kimodo primary, ARDY eval-only, which retires
the question entirely.* (Blocks download row 31; see C-3/C-4.)
10. [REMOTE] OpenTopography — the free non-academic key is 50 calls / 24h; 69 regions × 3
datasets ≈ 207 calls ≈ 4+ days of API budget. Accept the spread, buy OpenTopography Plus, or
take the AWS bulk-COG route (no account, no limit, more engineering). A spend/schedule fork.
11. [REMOTE] Board delivery medium — recommendation: a published page for the eight
mythological realms, local files for the fairy realm (the most spoiler-dense artifact the
project will produce). It is a publishing decision, so his nod.
12. [REMOTE] May any public marketing surface ever carry a generated frame? LTX-2's Attachment A
§5 compels an intelligible machine-generated disclaimer on any surface that does.
Recommendation: no — every public frame stays a real capture, which is already the ruled
Capture lane and costs the video lane nothing.
13. [REMOTE] Neither commercial image licence — BFL self-serve tiers (Builder 10K img/mo →
Enterprise) and Krea (Enterprise above $1M revenue) are both purchasable and both only he can
sign. Recommendation: buy neither; the MIT/Apache tier is genuinely competitive and the
licence-drift surveillance cost is the expensive part, not the sticker.
14. [LATER] Ship-time watermark call — keep Resemble's Perth neural watermark in shipped
Chatterbox VO as a provenance asset, or strip it (MIT permits). *Recommendation: keep.*
Accounts and money:
15. [REMOTE] ElevenLabs account + payment — commercial licence starts at Starter $6/mo; the
15-node hero budget likely wants Creator ($11) or Pro ($99). He creates and pays; the key enters
the env like every other credential.
16. [REMOTE, OPTIONAL] Cascadeur Indie $96/yr (converts to perpetual after one year) — the one
paid animation seat. Optional second: Rokoko Studio cloud fallback (~$120/yr).
17. [REMOTE, DEFERRABLE] Krotos Dehumaniser 2 + Reformer Pro ≈ $798 — creature vocalisation,
the only purchase the audio lane needs, deferrable to the creature pass.
18. [REMOTE, OPTIONAL] BOOM Library buyout $99-199 — only if benchmark day shows Sonniss+MOSS
leaves gaps.
19. [REMOTE] Website W1 — authorize the Cloudflare↔GitHub connection for Workers Builds, once
(one OAuth authorization, no secret ever enters our repo — recommended over a long-lived API
token). W2 — create the humanity-site GitHub repo. W3 — the first custom-domain attach
authorization (DNS records and certs are then created automatically; no manual DNS typing).
20. [REMOTE] W4 — name the press-contact address to print. A decision, not a task. Email Routing
stays deferred and is not being re-offered.
21. [LATER] W5 go-live flip (already RULED: at first real captures) · W6 the "Humanity
Forgotten" TM search/filing + the defensive-variant re-check at go-live.
Ongoing / conditional:
22. [REMOTE] Lane-R rights. The National Geographic run is his physical property, not a
publication licence — NG plates are internal underlay only. **own_capture (his own photography)
is the cleanest rights class available and only he can create it**; likewise any per-site
commercial-capture permission.
23. [CONDITIONAL] Voice-talent consent records. If a real person is ever cloned, the §4.4.1
five-stage authorization is his signature, held by him. *The loop generates; it never authorizes.*
24. [ONGOING] Choice-board rulings when the realm art program surfaces boards — the one visual
artifact that legitimately reaches him before THE JOSH GATE, because a board is a *decision
surface*, not a playtest.
Explicitly NOT his: D-3/D-4/D-5/D-6 model picks (director-ruled on benchmark evidence) · lane
scheduling · model promote/demote inside the pinned roster · retries · the Batch-0 backfill · every
registry edit · all site content, CI and i18n · every QA gate.
---
These are places where two or more dossiers claim the same resource, home, or primacy differently.
I am not resolving them. Each names the claimants so the director can rule once.
C-1 · THE COMFYUI QUESTION — the largest structural conflict in the survey. How many ComfyUI
instances exist, on which side of the WSL2 boundary, and whose models/ tree do they share?
this lane does not need it").
comfy_ui as its host (option a), with Windows as a fast-iteration lane.ace_step_1.5_turbo_aio.safetensors into "ComfyUI/models/checkpoints/" — unspecified which.Two instances cannot both hold 8188; SeedVR2 is claimed by ART and VIDEO simultaneously and would
be either duplicated across the boundary or shared through the vhdx. The download manifest, the
firewall table, and the dispatcher's lane map all assume a single answer that does not currently exist.
C-2 · GEN5 (C:) EXCLUSIVITY. ART ("C:/Gen5 stays UE-only"), 3D ("nothing from this lane touches
it"), VOICE ("no model weights"), and CONTROL_PLANE ("nothing from this lane") all assert Gen5 is
UE-only. WEBSITE places C:\dev\humanity-site\ on Gen5 by design, arguing it is a git-hot,
latency-bound small working tree and not a model store — and separately that it must NOT live inside
C:\dev\humanity-forgotten (a node_modules inside the gate-enrolled canon tree collides with
run_gates.py and the fidelity baseline). Note the canon repo already lives on C:, so "UE-only" is
already inexact. Two sub-questions ride this: whether website builds on Gen5 are an acceptable
neighbour for DDC/shader-compile I/O, and whether the "must not run during a UE cook" rule is
sufficient mitigation.
C-3 · ARDY vs KIMODO — three answers on primacy, two on lane ownership.
repo's standing anti-Llama posture), owning the lane outright.
"constrained/keyframed" — a *split by job*, which is a third shape.
The routing table is the committed artifact, so whichever way this lands must be written there.
C-4 · THE LLAMA-3-8B RIDER IS INVISIBLE TO TWO OF THE THREE CLAIMANTS. ANIMATION names ARDY's
text encoder as gated meta-llama/Meta-Llama-3-8B-Instruct (~16 GB, Llama 3 Community License, a
Josh HF acceptance, and a live conflict with the repo's anti-Llama posture). 3D lists ARDY as motion
primary describing only a "text encoder ~14GB VRAM (CPU-offloadable)" without naming Llama or the
gate. CONTROL_PLANE pins ARDY with no rider at all. **If ARDY is primary, Josh-minimum gains an item
that two of the three dossiers do not know about** — and a 16 GB gated download joins the night list.
C-5 · HUNYUAN3D-2.1 IS SIMULTANEOUSLY A SHIP PRIMARY AND A PREVIZ-ONLY ROW. 3D lists it as the
"Finished textured PBR | PRIMARY." CONTROL_PLANE's routing table marks it **previz-only until Josh
rules on ship territory**. VIDEO applies the same §2.3 posture to its sibling HunyuanVideo-1.5 and
concludes "internal previz only, permanently." All three cite the same clause (1M MAU + EU/UK/KR
exclusion). It is the only model in the survey occupying a ship-path slot under a licence the ruled
posture classifies previz-only — and it is also the 29 GB ceiling case the whole concurrency
budget is sized around.
C-6 · COSMOS 3's LICENCE-READ STATE DIVERGES. VIDEO fetched nvidia/Cosmos3-Nano and read
OpenMDW-1.1 ("ready for commercial and non-commercial use" — the cleanest licence in the survey).
CONTROL_PLANE's routing row records "not read." 3D excludes it on capability grounds without a
licence read. The *verdict* agrees everywhere (excluded), but the routing table would ship carrying a
stale "not read" against a licence that has in fact been read — and VIDEO explicitly flags Cosmos as a
WATCH item that would jump to public-surface-eligible if ComfyUI support lands.
C-7 · SEEDVR2 HAS TWO OWNERS AND TWO SIZES. ART wants SeedVR2-3B NVFP4 (~3 GB) in the Windows
ComfyUI as its upscale/finish step; VIDEO wants SeedVR2-3B (~7 GB, +7B optional ~15 GB) in the
WSL2 ComfyUI for 720p→delivery. Same Apache-2.0 model, two hosts, two download rows, two quantization
choices. Downstream of C-1.
C-8 · "EXCEL BASS" CARRIES THREE DIFFERENT VERDICT TOKENS. AUDIO: "CONFIRMED as XL BASE."
CONTROL_PLANE: "KILLED as a name" (a transcription artifact of xl-base). VOICE:
"UNVERIFIED-EXCLUDED" (correct within the voice lane — no such TTS model exists). The underlying
facts are compatible, but a config or registry cell will inherit whichever token it read first. One
canonical verdict string is needed: *the name is refuted, the referent acestep-v15-xl-base is real.*
C-9 · D-1 HAS NO SINGLE ANSWER, AND ONE LANE'S RATIFICATION ASK IS PHRASED MACHINE-WIDE. 3D asks
Josh to ratify "(a) WSL2+Puget as the generation host and (b) ComfyUI-on-Windows as the image lane
only." But AUDIO declares it has "removed the audio lane from the D-1 fork entirely," ANIMATION
declares it "must NOT be folded into the ComfyUI install," VIDEO says (b) is "genuinely viable here,"
and WEBSITE + CONTROL_PLANE say D-1 does not bind them at all. A per-lane D-1 may be the right answer
— but the ratification currently on offer would read as a machine-wide ruling that four lanes have
already departed from.
C-10 · FIREWALL AND PORT OWNERSHIP GAP. AUDIO flags that **port 8001 (ACE-Step REST) is missing
from the Stage-6F firewall table**, which lists UE, MCP 9315, MCP 8000, Ollama 11434 and ComfyUI 8188.
CONTROL_PLANE owns Stage-6F pre-authorization and does not list 8001. WEBSITE declares no dev-server
port at all. The consequence is named by the control plane itself: *an unacknowledged Defender dialog
at 04:00 presents as a hung editor, not an error.*
C-11 · BLENDER: ONE BINARY, TWO OWNERS, TWO SCHEDULING CLASSES. 3D installs Blender 5.2 LTS
(Python 3.13) as its CPU finishing host and classes it ALWAYS-ON (the lane's throughput win).
ANIMATION installs "Blender 5.x" as its headless FBX bake host and puts that bake **in the overnight
window**. Sharing the binary is fine; the two scheduling-class claims on the same CPU budget are not
reconciled, and both lanes independently budget ~8 threads.
**C-12 · FOUR LANES WANT A SCHEMA CHANGE IN THE SAME BATCH-0 / PRE-FREEZE COMMIT, UNDER THREE NAMES
FOR THE SAME FACT. VOICE asks for generating_model + model_licence** columns on voice rows.
ANIMATION asks for asset_licence_ref on the assetgen block at Batch 0. CONTROL_PLANE declares
model_licence + previz_only mandatory in a git-tracked *artifact record beside the output*
(not a registry column) plus the 796-row generation_status backfill plus new status columns
on T0_SFX_Registry and T0_Scene_Spec_Registry. AUDIO independently asks for regeneration_trigger
on T0_Theme_Registry and all four assetgen columns on T0_SFX_Registry. ART adds a fifth shape: a
licence-class token folded into generation_prompt_hash so a model's demotion fires
regeneration_required automatically. All five are the same requirement — *name the generating
model's licence per artifact* — and they would collide in one commit with one baseline re-emit.
C-13 · THE "TWELFTH DR-2 CLASS" ORDINAL. ANIMATION proposes a twelfth §2 row
(Animation / montage, keyed on ability_id + presentation_tier_ref). VIDEO explicitly refuses to
invent one ("Video generation is NOT a DR-2 asset class, and I am not inventing a twelfth one") and
routes itself through the existing Cinematic-scene row as reference material only. Both positions can
be simultaneously correct — but they would race for the same ordinal in the same schema commit as C-12,
and ANIMATION's row depends on an unruled PC-5 decision (that the canon row is authoritative and
the montage is retimed to it) without which "a generated 1.4s windup silently rewrites a 0.9s canon
row, 200 times."
C-14 · TWO LANES ASSERT THE SAME LARGE RAM RESERVATION. ART requires 32-48 GB held free for
text-encoder CPU offload; 3D budgets ~32 GB for its always-on CPU finishing half. Both are correct
in isolation and both fit — *unless* a UE cook (32-48 GB) coincides with both, which puts the box at
~130 GB of 128. Not a contradiction; an unscheduled overlap nobody owns.
---
Eight independently-researched lanes converged, without coordination, on all of the following:
the 29-of-32 GB serialization law and that a self-inflicted OOM is a *false* hardware-upgrade
trigger · licence-read-before-first-GPU-hour, with the operative *document* re-fetched rather than
a model card summarized · the generating model's licence named in every artifact, mechanized so a
demotion fires regeneration_required rather than depending on someone remembering · **dispatch from
script-emitted registry rows, never a hand-typed worklist · care-gated ids never raw-generate** —
the composer refuses on a column · previz-tagged outputs can never promote · **a tooth without a
failing fixture is not armed · and no quality claim without a capture behind it**, read by a fresh
critic, with the director re-reading the deciding frames from the newest session before any claim
enters a commit message.
Three lanes are zero-GPU and can start before the box arrives: WEBSITE (fully buildable today on
the current machine — Node 24.15.0 confirmed present), CONTROL_PLANE (0 GB, ~2 GB RAM), and the
Batch-0 registry backfill that unblocks every other lane's autonomy.