This website requires JavaScript.
Explore
Help
Register
Sign In
upstream-archive
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-27 16:37:29 -05:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
b11108
llama.cpp
/
ggml
T
History
Bartowski
f95b0d9539
ggml : IQ1_M build prefix sums once per block (
#28706
)
2026-09-22 16:54:45 +03:00
..
cmake
CUDA: replace GGML_FA_ALL_QUANTS with GGML_FA_QUANTS, more control over what is compiled (
#28079
)
2026-09-09 12:50:08 +02:00
include
sycl : pinned memory use right device context instead of 0 (
#28895
)
2026-09-21 13:58:59 +03:00
src
ggml : IQ1_M build prefix sums once per block (
#28706
)
2026-09-22 16:54:45 +03:00
.gitignore
vulkan : cmake integration (
#8119
)
2024-07-13 18:12:39 +02:00
CMakeLists.txt
ggml : bump version to 0.24.0 (ggml/1627)
2026-09-14 16:45:33 +03:00