fix(tiers): seed blackwell-8 and ampere-16 with 2026-09-23/24 measured values - #475
Merged
Merged
Conversation
…d values Both reference boxes (OptiPlex 7060 / blackwell-8, Lenovo M720q A2 / ampere-16) got a harness remediation pass on 2026-09-23/24. Carry the measurements they produced into the tier seed so a fresh install gets them instead of stale defaults; docs/tiers/*.md regenerated to match. blackwell-8 (OptiPlex 7060, RTX 5060 8GB), all under config_seed_ram_mid_high unless noted: - imagegen_timeout_sec 900 -> 2400: the qwen-image-2.1 binding's own measured max wall is 678.1s at 2752x1536 under the power tune (900 left ~1.3x margin, 2400 leaves ~3.5x). - sdcpp_extra_args added for z-image-turbo: the measured-working graph-cut arm (--vae-tiling --offload-to-cpu --diffusion-fa --max-vram 6.5 --stream-layers), 56-69s/1024^2 image. Two failed arms ruled out: bare --vae-tiling silently ran on the Intel iGPU (Vulkan device-pick default), and an all-VRAM placement OOM'd at 565-608s/step. - videogen_wan_virtual_vram_gb added at 12 (measured: 7 OOMs at lower values under the no-sysmem-fallback policy). - videogen_timeout_sec 3600 -> 4500 = ceil(2,970.3 x 1.5), from a measured cold native 81-frame render of 2,970.3s. - audiogen_timeout_sec added at 1500 (config_seed, unbound before): sized from a measured cold ACE-Step render (421.9s) + the audio-QA dead-air retry (~420s) + loudnorm, ~870s measured, x1.5 rounded up; also bounds the measured 108.7s cold voice render. - imagegen_families adds two named, opt-in bindings beside the default hidream-o1-dev/z-image-turbo pair (ADR 0058): qwen-image-2.1 (BINDING.md's canonical bf16 overlay, verbatim; A2 evaluation CLEARED 16/16) and qwen-image-2512 (Apache-2.0, commercial; measured 190.9s at 1328^2 with exact text rendering). - Not seeded, on purpose: a qwen-image-2.1-int8 checkpoint (benched, not operator-authorized as a named opt-in) and animate_character (no config key exists for the route yet on this box). ampere-16 (Lenovo M720q, NVIDIA A2 16GB): - videogen_wan_virtual_vram_gb added at 9 (the tier declared Wan I2V keys but never this one, so a fresh install silently rendered the graph script's own default of 7). A live smoke render measured a peak of 9.6 of 15.0 GiB across the full render at 9 GiB, 5.4 GiB headroom, zero OOM — a measured-safe floor, not yet a measured optimum (the smoke used a reduced frame count/resolution). Confirmed unchanged, already correct: blackwell-3x16's comfy_cuda_device "2" (matches the 2026-09-24 measurement that an upscale-class render stayed on card 2 while the display card idled; displaycard_placement_test.go already gates this) and videogen_pool_compute "cuda:0" (the ComfyUI-MultiGPU #220 exception). Not touched: ampere-8's comfy_extra_args temp/output-dir redirect (measured, but a literal per-box drive path, not a tier-general rule) and ffmpeg_path on any tier (PR #471 already makes a bare "ffmpeg" resolve via PATH; no tier seeds a broken value). go build ./... and go test ./... pass, including displaycard_placement_test.go, media_roster_test.go, docs_tiers_test.go and every tierseed/config/pipeline suite. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
Bumps all four version sources together (main_test.go's TestVersionSourcesAgree): VERSION, internal/buildinfo.Version, .printing-press.json, and a new dated CHANGELOG.md entry. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Both reference boxes (OptiPlex 7060 / blackwell-8, Lenovo M720q A2 / ampere-16) got a harness
remediation pass on 2026-09-23/24. This carries the measurements they produced into the tier
seed (
setup/templates/profiles.json) so a fresh install gets them instead of stale defaults,and regenerates
docs/tiers/*.md(go run ./cmd/gentiers).blackwell-8 (OptiPlex 7060, RTX 5060 8GB), all under
config_seed_ram_mid_highunless noted:imagegen_timeout_sec900 → 2400 — the qwen-image-2.1 binding's own measured max wall is 678.1sat 2752×1536 under the power tune (900 left only ~1.3x margin, 2400 leaves ~3.5x).
sdcpp_extra_argsadded for z-image-turbo: the measured-working graph-cut arm(
--vae-tiling --offload-to-cpu --diffusion-fa --max-vram 6.5 --stream-layers), 56-69s/1024²image. Two failed arms are explicitly not seeded: bare
--vae-tilingsilently ran on theIntel UHD 630 iGPU (Vulkan device-pick default), and an all-VRAM placement OOM'd/spilled at
565-608s/step.
videogen_wan_virtual_vram_gbadded at 12 (measured: 7 OOMs at lower values under theno-sysmem-fallback policy).
videogen_timeout_sec3600 → 4500 = ceil(2,970.3 × 1.5), from a measured cold native 81-framerender of 2,970.3s.
audiogen_timeout_secadded at 1500 inconfig_seed(previously unbound, framework default720): sized from a measured cold ACE-Step render (421.9s) + the audio-QA dead-air retry
(~420s) + loudnorm ≈ 870s measured, ×1.5 rounded up; also bounds the measured 108.7s cold
voice (Chatterbox) render.
imagegen_familiesadds two named, opt-in bindings beside the default hidream-o1-dev /z-image-turbo pair (ADR 0058, ram_mid_high only — both need ≥64GB host RAM for staging):
qwen-image-2.1(BINDING.md's canonical bf16 overlay, verbatim; A2 evaluation CLEARED 16/16)and
qwen-image/ Qwen-Image-2512 (Apache-2.0, commercial; measured 190.9s at 1328² withexact text rendering).
but fidelity split 1/2 over the LPIPS≤0.03 bar — and not operator-authorized as a named
opt-in) and
animate_character(no config key exists for the route on this box yet; itsmodel files are unmeasured at this VRAM class).
ampere-16 (Lenovo M720q, NVIDIA A2 16GB):
videogen_wan_virtual_vram_gbadded at 9 — the tier declared Wan I2V keys but never this one,so a fresh install silently rendered the graph script's own default (7). A live smoke render
measured a peak of 9.6 of 15.0 GiB across the full render at 9 GiB, 5.4 GiB headroom, zero
OOM — a measured-safe floor, not yet a measured optimum (the smoke used a reduced frame
count/resolution than the 81-frame production default).
Confirmed unchanged, already correct (checked, no edit needed): blackwell-3x16's
comfy_cuda_device: "2"(matches the 2026-09-24 measurement anddisplaycard_placement_test.go'sdocumented rule) and
videogen_pool_compute: "cuda:0"(the ComfyUI-MultiGPU #220 exception).Not touched, and why: ampere-8's
comfy_extra_argstemp/output-dir redirect is a measured,working value, but a literal per-box drive path (
D:/ComfyUI-tempetc.), not a tier-generalrule — confirmed by dedicated research against the source docs.
ffmpeg_pathon any tier: PR#471 already makes a bare
"ffmpeg"resolve via PATH, and no tier currently seeds a brokenvalue.
Also bumps VERSION 0.140.11 → 0.140.12 (and the other three version sources
TestVersionSourcesAgreechecks:internal/buildinfo.Version,.printing-press.json,CHANGELOG.md) with a changelog entry for this change.
Test plan
go build ./...— cleango test ./...— all packages pass, includingdisplaycard_placement_test.go,media_roster_test.go,docs_tiers_test.go,TestVersionSourcesAgree, and everyinternal/tierseed,internal/config,internal/pipeline,internal/tierdocssuitego run ./cmd/gentiersregenerated exactlydocs/tiers/ampere-16.mdanddocs/tiers/blackwell-8.md(matches the profiles.json diff), and the staleness gate(
TestTierDocsAreCurrent) passessetup/templates/profiles.jsonstays CRLF throughout (checked byte-for-byte before andafter every edit)
origin/main(post PR feat(seats): mimo-9b-agent becomes amd-gcn's bound agent seat #474 / 0.140.11) and rechecked VERSIONimmediately before push
🤖 Generated with Claude Code