
More Data Won’t Save Your LLM. Better Data Will.
There was a point, not long ago, when the dominant strategy for improving large language models was simple: feed them more. More tokens, more compute, more parameters. The scaling laws
Booth 21-25 | AI Data Management Zone | Tokyo Big Sight

There was a point, not long ago, when the dominant strategy for improving large language models was simple: feed them more. More tokens, more compute, more parameters. The scaling laws

For most of the last six years, the most capable LLMs were locked behind APIs. You could call them, you could build on them, but you couldn’t look inside, customize the weights, or run them on your own infrastructure. That was simply the reality of working with frontier models. In
