The semantic attention bias
Each head h adds a bias to the logit of terrain token i:
Bh,i = βh · ρh(di) · r̄i
βh is a signed per-head gain, ρh a learned radial profile over the distance di from the cell to its nearest foot, and r̄i the max-pooled contact cost. The bias is not a mask: flagged cells still receive attention, but its weight is reshaped where it can still change the next foothold and fades where it cannot. Eight gains and eight six-point profiles are the whole mechanism, so the learned behaviour is directly inspectable.
