Would be very glad if anyone explained to me if and why this is revolutionary.
Here you go:
"Across algorithm engineering, mathematical optimization, and GPU kernel engineering, Dream-RSI achieves competitive or improved discovery quality while substantially reducing discovery cost in several settings."
Fairly certain all the labs are doing this (RSI) at this point. It's a question of how public their proclamations are about it and how they're positioning PR etc.
Even today's lighter weight models know how to write kernels and optimize them. I've had DeepSeek 4.1 Flash tune the crap out custom CUDA kernels on my own codebase and it was entirely competent at it. And cheap.
The innovation pieces will be in the harnesses to support this. Which I guess is partially what's going on here.
Probably isn't by virtue of it being publicly released