This website requires JavaScript.
Explore
Help
Sign In
Public
/
llama.cpp
Watch
1
Star
0
Fork
0
You've already forked llama.cpp
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-02 19:11:05 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
6fdd0ac8907fd973a42b876357823ad2124cd8ed
llama.cpp
/
include
T
History
Xuan-Son Nguyen
732707dff2
quantize: cap working memory size to avoid loading big tensors onto RAM (
#27795
)
2026-08-27 18:31:13 +02:00
..
llama-cpp.h
llama : re-enable manual LoRA adapter free (
#19983
)
2026-03-18 12:03:26 +02:00
llama.h
quantize: cap working memory size to avoid loading big tensors onto RAM (
#27795
)
2026-08-27 18:31:13 +02:00