Files
llama_cpp/ggml/src
Jeff BolzandGitHub 466300fe14 vulkan: optimize coopmat2 q4_k/q5_k dequant functions. (#11206)
Do masking on whole dwords, fetch all scales at once.
2025-01-16 22:23:49 +01:00
..
2025-01-07 08:37:02 +02:00