He was studying how AI fails before most people believed it worked. Now he builds the guardrails that keep everyone else's models honest.