AWS Insights

Wrong, not broken

For 20 years, software operations has been built around one question: Is it broken?

We’ve gotten very good at answering it. Metrics, logs, and traces. Distributed tracing that can follow a single request through dozens of services.

And all of it rests on one sensible assumption: When software fails, the failure eventually shows up as something a machine can measure.

Now picture an AI agent handling refund requests. It answers in 600 milliseconds. The error rate is zero. And it tells a customer they’re owed $40 when the policy says $140, because the document it retrieved was last quarter’s version.

Every dashboard is green. Nothing appears to be broken.