The 61-point gap
How much autonomy enterprises allow, versus how much of it they can verify
Let agents act with no human review, or will within 12 months
Fully trust the automated evals gating those releases
Gartner projects that by 2028, 40% of enterprise AI failures will trace to inadequate evaluation and monitoring, not to model capability.