As someone typing this on an M1 Mac with 32 GB of RAM who tried using 3.8 27B (Q4_K_M) yesterday in both LM Studio and llama.cpp, I wouldn't call it particularly usable in terms of token speed. (and that was with `--spec-type draft-mtp` for llama.cpp).
If you want to leave it running with the fans going crazy for 40 mins or overnight or something, fair enough, but otherwise it doesn't seem worth it to me. It's certainly not "interactive", even taking into account the over-thinking it does by default.
The 3.6 (maybe they'll release a 3.8?) MoE model is much more usable (but obviously not as good) on this machine spec.