This website requires JavaScript.
Explore
Help
Register
Sign In
upstream-archive
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-21 13:37:29 -05:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
master
llama.cpp
/
ggml
T
Add File
New File
Upload File
Apply Patch
Copy Permalink
Download directory as ZIP
Download directory as TAR.GZ
History
Foad Abo Dahood
fb34fc262c
metal : fix mask bounds in flash attention block pre-pass (
#29220
)
2026-09-21 20:31:56 +03:00
..
cmake
CUDA: replace GGML_FA_ALL_QUANTS with GGML_FA_QUANTS, more control over what is compiled (
#28079
)
2026-09-09 12:50:08 +02:00
include
sycl : pinned memory use right device context instead of 0 (
#28895
)
2026-09-21 13:58:59 +03:00
src
metal : fix mask bounds in flash attention block pre-pass (
#29220
)
2026-09-21 20:31:56 +03:00
.gitignore
vulkan : cmake integration (
#8119
)
2024-07-13 18:12:39 +02:00
CMakeLists.txt
ggml : bump version to 0.24.0 (ggml/1627)
2026-09-14 16:45:33 +03:00