mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-07-29 14:11:13 -05:00
* vulkan: disable FA mask_opt on GCN to improve performance * reenable mask opt over attention head size 256
* vulkan: disable FA mask_opt on GCN to improve performance * reenable mask opt over attention head size 256