← all workloads

loop_unroll

A force-unrolled loop of `n` iterations. Stresses loop unrolling (unrollLoopsInModule) and the downstream SSA simplify on the unrolled body. Scaling null: the unrolled output is O(n), so ideal cost is linear in n (linear in output size).

bucket: loop_unroll  ·  mode: target  ·  flags: -target spirv -emit-spirv-directly

Phase composition vs N (stacked sub-counters)

compileInner split into phase buckets (named leaves + (self) residuals) stacked across the sweep sizes — the top edge is compileInner, so you can see which phase drives the scaling.

loop_unroll — phase composition vs N (v2026.13, median ms) loop_unroll 35.6× over N 75→600 0.0 988 1976 75 150 300 600 N loop_unroll — parseTranslationUnit loop_unroll — SemanticChecking loop_unroll — generateIR loop_unroll — frontEndExecute (self) loop_unroll — specializeModule loop_unroll — simplifyIR loop_unroll — linkIR loop_unroll — unrollLoopsInModule loop_unroll — legalizeResourceTypes loop_unroll — legalizeExistentialTypeLayout loop_unroll — performMandatoryEarlyInlining loop_unroll — performForceInlining loop_unroll — generateOutput (self) loop_unroll — compileInner (self) phase buckets parseTranslationUnit SemanticChecking generateIR frontEndExecute (self) specializeModule simplifyIR linkIR unrollLoopsInModule legalizeResourceTypes legalizeExistentialTypeLayout performMandatoryEarlyInlining performForceInlining linkAndOptimizeIR (self) emitEntryPointsSourceFromIR generateOutput (self) compileInner (self)

Scaling analysis

floor-subtracted power-law fit (t − floor) = a·Nk; floor = the minimal workload (fixed per-compile cost), k the global exponent, top-2× the local high-end doubling ratio.

N rangefloor (ms)k (work)fit R²t(Nmin)t(Nmax)top-2×
75–600101.820.9955118294.25×

Growth attribution (N=75 → N=600)

compileInner grows by 1778 ms across the sweep; the mutually-exclusive phase buckets below partition that growth exactly (no nested-timer double counting). × lin is the same metric as the top-level panels, per bucket: the end point vs a linear expectation anchored to the bucket's share of the minimal floor and fitted on the low-N half — 1.0 = grew exactly linearly, >1 bends up. The super-linearity lives where × lin (and k) are red.

buckett@N=75t@N=600Δ msshare× lin∝Nk
specializeModule10497+48727%4.32×1.88
unrollLoopsInModule7441+43424%4.78×1.97
generateOutput (self)7247+24013%4.57×1.88
simplifyIR7224+21712%3.02×1.66
legalizeExistentialTypeLayout3203+20011%4.98×2.02

Also growing (below top-5): legalizeResourceTypes (+200 ms, 11%).

Near-constant (≤2% of growth each): SemanticChecking (9→9 ms), generateIR (4→4 ms), linkIR (1→0 ms), performForceInlining (0→0 ms), frontEndExecute (self) (0→0 ms), performMandatoryEarlyInlining (0→0 ms), parseTranslationUnit (0→0 ms), compileInner (self) (0→0 ms).

Sweep numbers (median ms)

NcompileInnerunrollLoopsInModulesimplifyIR
755188
1501313124
30043111979
6001829564286