> llama.cpp via Vulkan (AMD / Intel / NVIDIA) or CPU fallback
I got excited about someone paying attention to intel. Oh well.
From what I've seen, Vulkan adds a lot of overhead on Intel hardware.
I didn't see any benchmarks against vllm, sglang, exllama, etc
> llama.cpp via Vulkan (AMD / Intel / NVIDIA) or CPU fallback
I got excited about someone paying attention to intel. Oh well.