n constant buffers, each with a structurally rich payload — vectors, matrices, an array of a nested `Light` struct, a nested `Material` struct and a scalar array. This is the only workload with a large, deeply-typed shader *parameter interface* (every other workload's interface is a single RWStructuredBuffer), so it is the one stressor for the parameter binding / layout-assignment engine and — with `-reflection-json` (see the spec's reflection_json flag) — the reflection serializer, the layout/reflection path no other workload covers. Scales by breadth = number of parameter blocks the layout engine must place and reflect. Note: layout is computed during compileInner regardless of the flag (so compileInner is the holistic signal); -reflection-json additionally runs the serializer, which is cheap today but tracked here for regression coverage. Scaling null: n scales parameters; ideal layout cost is O(n).
bucket: reflection_layout · mode: target · flags: -target spirv -emit-spirv-directly
compileInner split into phase buckets (named leaves + (self) residuals) stacked across the sweep sizes — the top edge is compileInner, so you can see which phase drives the scaling.
floor-subtracted power-law fit (t − floor) = a·Nk; floor = the minimal workload (fixed per-compile cost), k the global exponent, top-2× the local high-end doubling ratio.
| N range | floor (ms) | k (work) | fit R² | t(Nmin) | t(Nmax) | top-2× |
|---|---|---|---|---|---|---|
| 30–240 | 11 | 1.30 | 0.994 | 45 | 532 | 2.81× |
compileInner grows by 486 ms across the sweep; the mutually-exclusive phase buckets below partition that growth exactly (no nested-timer double counting). × lin is the same metric as the top-level panels, per bucket: the end point vs a linear expectation anchored to the bucket's share of the minimal floor and fitted on the low-N half — 1.0 = grew exactly linearly, >1 bends up. The super-linearity lives where × lin (and k) are red.
| bucket | t@N=30 | t@N=240 | Δ ms | share | × lin | ∝Nk |
|---|---|---|---|---|---|---|
| SemanticChecking | 16 | 229 | +213 | 44% | 2.49× | 1.49 |
| linkAndOptimizeIR (self) | 5 | 75 | +70 | 14% | 1.94× | 1.38 |
| generateOutput (self) | 6 | 55 | +49 | 10% | 1.42× | 1.21 |
| legalizeExistentialTypeLayout | 1 | 36 | +36 | 7% | 4.50× | 1.87 |
| simplifyIR | 4 | 35 | +31 | 6% | 1.16× | 1.07 |
Also growing (below top-5): generateIR (+20 ms, 4%), parseTranslationUnit (+13 ms, 3%), compileInner (self) (+13 ms, 3%), frontEndExecute (self) (+13 ms, 3%).
Near-constant (≤2% of growth each): specializeModule (1→10 ms), performForceInlining (1→10 ms), linkIR (1→7 ms), legalizeResourceTypes (0→5 ms), performMandatoryEarlyInlining (0→2 ms), unrollLoopsInModule (0→0 ms).
| N | compileInner | frontEndExecute | generateOutput |
|---|---|---|---|
| 30 | 45 | 23 | 20 |
| 60 | 84 | 40 | 40 |
| 120 | 189 | 92 | 90 |
| 240 | 532 | 282 | 235 |