@solarapparition 2025-11-05 ♥37 ↻3 original ↗
gpt-5 "feels small", so makes sense that it's still from a 4o base. i guess oai is all in on scaling purely via rl until that model size is wrung out, then up the model size somewhat, rinse and repeat

once again i'm further convinced that this is a slow-takeoff-y scenario. capabilities gain from rl within a model size just seems far narrower than back when pretraining scaling was still viable. and increasing model size would increase training time, which again has a natural slowing effect. really, it all ends up circling back to how much compute there is

author:solarapparition kind:tweet model:gpt-4o model:gpt-5 on:gpt-5 year:2025

cited on: gpt-5

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.