lin-0208
Design lesson
LINCS–RLHF separates learning preferences from forcing them into rewards. Circulation diagnoses whether scalarization is structurally available; tangent transport identifies how that availability changes; localization prevents heterogeneous populations from canceling; and admission controls which representation reaches policy optimization. When preferences do not scalarize, the repair is to retain relational structure, not to optimize its scalar projection harder.