This website requires JavaScript.
Explore
Help
Register
Sign In
upstream-archive
/
llama.cpp
Watch
1
Star
0
Fork
0
You've already forked llama.cpp
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-07-31 23:20:44 -05:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
ce5890b5f7d88fe3408398dfbbada00aec03d352
llama.cpp
/
ggml
/
src
/
ggml-webgpu
History
Chen Yuan
5306f4b3b5
fix(flash-attn): replace f32 with kv_type and q_type (
#23372
)
2026-05-21 07:58:49 -07:00
..
wgsl-shaders
fix(flash-attn): replace f32 with kv_type and q_type (
#23372
)
2026-05-21 07:58:49 -07:00
CMakeLists.txt
ggml webgpu: add support for emscripten builds (
#17184
)
2025-12-03 10:25:34 +01:00
ggml-webgpu-shader-lib.hpp
ggml-webgpu: makes the flash attn vec path subgroup-aware (
#23040
)
2026-05-14 09:31:36 -07:00
ggml-webgpu.cpp
ggml-webgpu : extend GDN for K>1 (
#23299
)
2026-05-19 09:45:41 +03:00
pre_wgsl.hpp
ggml webgpu: initial flashattention implementation (
#18610
)
2026-01-08 08:23:39 -08:00