Files
llama.cpp/ggml
Jeff Bolz 788e07dc91 vulkan: Support Q2_0 (#25430)
* vulkan: Support Q2_0

The backend perf tests for mat-vec-mul weren't very good at first (worse than
q2_k), doubling the rows per workgroup made a big difference.

* reorder

* resolve merge conflict, adjust err threshold for f16->q2_0 set_rows
2026-07-17 08:42:59 +02:00
..
2026-07-17 08:42:59 +02:00
2024-07-13 18:12:39 +02:00
2026-07-10 10:28:39 +03:00