
More Data Won’t Save Your LLM. Better Data Will.
There was a point, not long ago, when the dominant strategy for improving large language models was simple: feed them more. More tokens, more compute, more parameters. The scaling laws
Booth 21-25 | AI Data Management Zone | Tokyo Big Sight

There was a point, not long ago, when the dominant strategy for improving large language models was simple: feed them more. More tokens, more compute, more parameters. The scaling laws

The dominant model for AI-assisted research has been batch processing: submit a query, wait for the system to work through it, receive a result. The interaction is asynchronous, the feedback loop is slow, and the human is largely passive until the output arrives. A paper from early 2026 presents a
