logoalt Hacker News

bigyabaitoday at 4:54 PM0 repliesview on HN

Before the M5, there was no dedicated matrix multiplication hardware on the Apple Silicon GPU. Their solution was generally using the NPU and AMX coprocessors for tensor and matrix workloads.