A single shader that ramps *several* realistic dimensions together as n grows, modelling how a real shader's compile cost behaves going from simple to highly complex — as opposed to the single-axis stressors, which isolate one pass. Each step adds: a branchy helper (control flow / SSA + phi), a generic call (sema + specialization), a small bounded inner loop, resource reads, and a dynamic-dispatch site. Helpers are chained in bounded-depth groups (call graph depth) so doubling n roughly doubles total work across all of front-end, IR opt, and codegen at once. Sweep this to get the holistic complexity->compile-time curve and a floor+slope fit. Scaling reference (not a null): several axes co-scale, so the sweep's linear line means 'cost tracks code size' — a user-meaningful reference. Standing super-linearity here is expected; the actionable signal is its release-over-release movement.
bucket: realistic_scaling · mode: target · flags: -target spirv -emit-spirv-directly
compileInner split into phase buckets (named leaves + (self) residuals) stacked across the sweep sizes — the top edge is compileInner, so you can see which phase drives the scaling.
floor-subtracted power-law fit (t − floor) = a·Nk; floor = the minimal workload (fixed per-compile cost), k the global exponent, top-2× the local high-end doubling ratio.
| N range | floor (ms) | k (work) | fit R² | t(Nmin) | t(Nmax) | top-2× |
|---|---|---|---|---|---|---|
| 160–1280 | 10 | 1.39 | 0.997 | 230 | 4019 | 3.00× |
compileInner grows by 3789 ms across the sweep; the mutually-exclusive phase buckets below partition that growth exactly (no nested-timer double counting). × lin is the same metric as the top-level panels, per bucket: the end point vs a linear expectation anchored to the bucket's share of the minimal floor and fitted on the low-N half — 1.0 = grew exactly linearly, >1 bends up. The super-linearity lives where × lin (and k) are red.
| bucket | t@N=160 | t@N=1280 | Δ ms | share | × lin | ∝Nk |
|---|---|---|---|---|---|---|
| generateOutput (self) | 73 | 1175 | +1103 | 29% | 1.83× | 1.35 |
| simplifyIR | 36 | 1048 | +1012 | 27% | 2.88× | 1.61 |
| linkAndOptimizeIR (self) | 25 | 571 | +547 | 14% | 2.24× | 1.53 |
| specializeModule | 21 | 350 | +329 | 9% | 1.86× | 1.34 |
| legalizeResourceTypes | 5 | 234 | +229 | 6% | 3.62× | 1.85 |
Also growing (below top-5): legalizeExistentialTypeLayout (+229 ms, 6%), generateIR (+152 ms, 4%), SemanticChecking (+152 ms, 4%).
Near-constant (≤2% of growth each): parseTranslationUnit (2→13 ms), linkIR (2→12 ms), performMandatoryEarlyInlining (1→11 ms), performForceInlining (1→4 ms), unrollLoopsInModule (0→3 ms), compileInner (self) (0→0 ms), frontEndExecute (self) (1→0 ms).
| N | compileInner | frontEndExecute | linkAndOptimizeIR | simplifyIR |
|---|---|---|---|---|
| 160 | 230 | 61 | 97 | 36 |
| 320 | 525 | 102 | 256 | 96 |
| 640 | 1338 | 181 | 748 | 280 |
| 1280 | 4019 | 375 | 2468 | 1048 |