Stressful annotation conditions create structured confounds in RLHF preference labels. This bias differs from random noise because it shifts based on the rater's state over time. These patterns propagate through reward models and policy optimization. Researchers can now use this audit framework to isolate and remove rater-induced distortions from preference data.