This website requires JavaScript.
Explore
Help
Sign In
Public
/
llama.cpp
Watch
1
Star
0
Fork
0
You've already forked llama.cpp
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-09 14:29:06 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
14bb119b2eea8f1ae1ea7312f66fb4ffb2b89313
llama.cpp
/
ggml
/
src
/
ggml-webgpu
T
History
Masashi Yoshimura
f401bb1390
ggml-webgpu : refactor several wgsl files and simplify flash_attn wgsl. (
#26134
)
2026-08-10 09:29:41 +03:00
..
wgsl-shaders
ggml-webgpu : refactor several wgsl files and simplify flash_attn wgsl. (
#26134
)
2026-08-10 09:29:41 +03:00
CMakeLists.txt
ggml-webgpu: FlashAttention refactor + standardize quantization support (
#23834
)
2026-06-04 08:05:04 +03:00
ggml-webgpu-shader-lib.hpp
ggml-webgpu : refactor several wgsl files and simplify flash_attn wgsl. (
#26134
)
2026-08-10 09:29:41 +03:00
ggml-webgpu.cpp
ggml-webgpu : refactor several wgsl files and simplify flash_attn wgsl. (
#26134
)
2026-08-10 09:29:41 +03:00
pre_wgsl.hpp
ggml-webgpu: FlashAttention refactor + standardize quantization support (
#23834
)
2026-06-04 08:05:04 +03:00