Files
llama.cpp/conversion
Oliver Simons 60dcb35afe Fix conversion of https://huggingface.co/RedHatAI/Qwen3.8-27B-NVFP4
1. Per row FP8 wrongly fell to NVFP4 conversion, as NVFP4 checked on
   dims alone. Also check on dtype for FP8
2. Need to reshape FP8 QKV projections scales in the same way that
   weights are reshaped
2026-09-30 15:55:51 +02:00
..
2026-09-27 13:43:08 +02:00
2026-09-07 21:10:06 +02:00
2026-09-02 12:46:16 +02:00
2026-08-17 10:15:11 +02:00