Mostly, I think, we don’t really understand your argument that Intel couldn’t easily replicate the parts needed only for inference.
back
1 comments
Yeah, for example llama.cpp runs on Intel GPUs via Vulkan or SYCL. The latter is actively being maintained by Intel developers.
Obviously that is only one piece of software, but its a certainly a useful one if you are using one of the many LLMs it supports.