mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-08-02 16:10:48 -05:00
* Add q8_0 and q4_0 set_rows * Add fast(er) quantization set_rows path * formatting/naming * a little more naming * Remove unused constant * Don't override other override * Avoid bitcast * Narrow relaxation
396 KiB
396 KiB