I owned release assurance for Reward Gateway’s first AI product.
I owned quality and release assurance for Reward Gateway’s first customer-facing AI product. The product had already been built, but there was no test infrastructure and no defensible basis for a launch decision.
I built security and evaluation from scratch. Exact-output checks were unsuitable because the model answered differently on each run. The suite tested whether responses stayed within the system’s boundaries and did the job asked of them, across repeated runs.
The evidence exposed material gaps before launch. I separated what we could prove from what remained unknown and took both to the launch panel. They accepted the residual risk and approved the release.
The suite became the release gate for later AI work. I also introduced follow-up question rate as a product signal because a second question usually meant the first answer had not done its job.