
What If AI Agents Could Catch Their Own Mistakes?
Most improvements to AI systems happen before deployment: better training data, better fine-tuning, better RLHF. Once the model is out in the world, you generally get the performance you trained
Booth 21-25 | AI Data Management Zone | Tokyo Big Sight

Most improvements to AI systems happen before deployment: better training data, better fine-tuning, better RLHF. Once the model is out in the world, you generally get the performance you trained

There’s a quiet result from Cheng et al. (Renmin University and Microsoft Research, 2026) that I keep coming back to. They gave large language models access to a minimal sandbox — essentially a code interpreter with a file system — and watched what happened. No additional training. No new data.
