I imagine it comes down to economics.. there isn't much upside to fixing the last 20% of issues that the dumber faster models are missing.
The cost to serve, latency profile ,and internal demand for a maximal intelligence model would probably keep it pointed at harder and more valuable problems most of the time.
Like front-running Millennium prize solutions?