lin-0172

13.7 Pretrained GPT-2 probe

Experiment: registered GPT-2 Medium pilot

Across three seeds, normalized commutator falls \(37.7\% \); pairwise logit and hidden order MSE fall \(58.5\% \) and \(58.6\% \); triple-permutation MSE falls \(57.9\% \). Mean task loss increases \(0.0026\) nats, and every geometric endpoint improves in every seed.

The pilot validates transfer of the bracket-control signature to pretrained attention adapters. Its short training schedule and low absolute SacreBLEU make it a geometric result, not evidence of high-quality semantic generation.