This website requires JavaScript.
Explore
Help
Register
Sign In
upstream-archive
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-30 09:57:38 -05:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
c0c7fa930d67c5ca63c22449b4790dcd15ae62dd
llama.cpp
/
include
T
History
Xuan Son Nguyen
c0c7fa930d
quantize: cap working memory size to avoid loading big tensors onto RAM
2026-08-27 13:48:24 +02:00
..
llama-cpp.h
llama : re-enable manual LoRA adapter free (
#19983
)
2026-03-18 12:03:26 +02:00
llama.h
quantize: cap working memory size to avoid loading big tensors onto RAM
2026-08-27 13:48:24 +02:00