OpenAI’s o3 suggests AI models are scaling in new ways — but so are the costs

December 24, 2024

219

Last month, AI founders and investors told TechCrunch that we’re now in the “second era of scaling laws,” noting how established methods of improving AI models were showing diminishing returns. One promising new method they suggested could keep gains was “test-time scaling,” which seems to be what’s behind the performance of OpenAI’s o3 model – but it comes with drawbacks of its own.

Much of the AI world took the announcement of OpenAI’s o3 model as proof that AI scaling progress has not “hit a wall.” The o3 model does well on benchmarks, significantly outscoring all other models on a test of general ability called ARC-AGI, and scoring 25% on a difficult math test that no other AI model scored more than 2% on.

Of course, we at TechCrunch are taking all this with a grain of salt until we can test o3 for ourselves (very few have tried it so far). But even before o3’s release, the AI world is already convinced that something big has shifted.

The co-creator of OpenAI’s o-series of models, Noam Brown, noted on Friday that the startup is announcing o3’s impressive gains just three months after the startup announced o1 – a relatively short timeframe for such a jump in performance.

OpenAI’s o3 suggests AI models are scaling in new ways — but so are the costs

Blue Origin still doesn’t know why its New Glenn rocket blew up last month

Threads adds new features to Live Chats as it expands access

Acti puts AI agents directly into your smartphone keyboard

Most Popular

$quib: Erring Album Review | Pitchfork

How Lower Manhattan Shaped Retail and the Garment Industry in America

Blue Origin still doesn’t know why its New Glenn rocket blew up last month

Debit: Potpourri Album Review | Pitchfork

Recent Comments

ABOUT US

POPULAR POSTS

$quib: Erring Album Review | Pitchfork

How Lower Manhattan Shaped Retail and the Garment Industry in America

Blue Origin still doesn’t know why its New Glenn rocket blew up last month

POPULAR CATEGORY