This website requires JavaScript.
Explore
Help
Sign In
Public
/
llama.cpp
Watch
1
Star
0
Fork
0
You've already forked llama.cpp
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-08-28 16:11:21 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
ca3d5a3e10d53f7ea672cb9b6178faca3e2807bc
llama.cpp
/
include
T
History
Xuan-Son Nguyen
732707dff2
quantize: cap working memory size to avoid loading big tensors onto RAM (
#27795
)
2026-08-27 18:31:13 +02:00
..
llama-cpp.h
llama : re-enable manual LoRA adapter free (
#19983
)
2026-03-18 12:03:26 +02:00
llama.h
quantize: cap working memory size to avoid loading big tensors onto RAM (
#27795
)
2026-08-27 18:31:13 +02:00