Earlier CPU-scale runs, separate from GPU Table 1 and the three-seed study: a two-layer transformer (about 6.7×104 parameters, no positional encoding) trained on synthetic traces. EM0 is in-range exact match; EM1 and EM2 are two splits longer than any training input.