Files
llama.cpp/tests
Oliver Simons 0211798e56 Make sharding HIP-compatible
1. Use ggml_cuda_get_physical_warp_size() to determine warp size flexibly
2. Add test with partial warp to test sum reduction on CUDA
2026-03-11 15:28:47 +01:00
..