⭐⭐⭐⭐ 4.2
The talk exposes a central tension in building AI agents: as models become more capable, their failures become more plausible and harder to catch.
Tariq Shaukat argues that hallucination is not a temporary bug but a deepening risk, because human reviewers increasingly suffer cognitive surrender—trusting plausible but wrong outputs.
The industry has focused on generation, but the real bottleneck is verification.
The talk exposes a central tension in building AI agents: as models become more capable, their failures become more plausible and harder to catch.
Tariq Shaukat argues that hallucination is not a temporary bug but a deepening risk, because human reviewers increasingly suffer cognitive surrender—trusting plausible but wrong outputs.
The industry has focused on generation, but the real bottleneck is verification.