mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-09-03 10:37:38 +02:00
* vulkan: disable FA mask_opt on GCN to improve performance * reenable mask opt over attention head size 256
* vulkan: disable FA mask_opt on GCN to improve performance * reenable mask opt over attention head size 256