I help engineering teams ensure that their AI agents actually perform as expected.
I've spent over 20 years building production software systems. Now I focus on the AI Reliability Layer: evaluating agents under adversarial conditions, implementing verification frameworks, and building testing methodologies so AI actually behaves as intended.