
What If AI Agents Could Catch Their Own Mistakes?
Most improvements to AI systems happen before deployment: better training data, better fine-tuning, better RLHF. Once the model is out in the world, you generally get the performance you trained
Booth 21-25 | AI Data Management Zone | Tokyo Big Sight

Most improvements to AI systems happen before deployment: better training data, better fine-tuning, better RLHF. Once the model is out in the world, you generally get the performance you trained

Most AI text detectors look good in testing. They’re trained and evaluated on specific models, specific prompt styles, specific domains — and they perform well within that distribution. Then they get deployed into the real world, which has different models, different prompting patterns, different domains, and the performance drops. This
