Sadly for them, Nvidia didn't stay still in the meantime and created the next generation of CUD...

fibonacci112358 • today at 6:25 AM • 2 replies • view on HN

Sadly for them, Nvidia didn't stay still in the meantime and created the next generation of CUDA, CuTile for Python and soon for C++, through CUDA Tile IR (using a similar compiler stack based on MLIR).

Event though it's not portable, it will likely have far greater usage than Mojo just by being heavely promoted by Nvidia, integrated in dev tools and working alongside existing CUDA code.

Tile IR was more likely a response to the threat of Triton rather than Mojo, at least from the pov of how easy is to write a decently performing LLM kernel.

Replies

melodyogonna • today at 7:26 AM

People keep mistaking Mojo as good syntax for writing GPU code, and so imagine Nvidia's Python frameworks already do that. But... would CuTile work on AMD GPUs and Apple Silicon? Whatever Nvidia does will still have vendor lock-in.

brcmthrowaway • today at 6:34 AM

Interesting, how big impact is CuTile?

alt Hacker News

Replies