Files
llama.cpp/docs
Piotr Wilkin (ilintar) 5a4d0fecae CUDA: replace GGML_FA_ALL_QUANTS with GGML_FA_QUANTS, more control over what is compiled (#28079)
* CUDA: add configurable FA quant combinations

Assisted-by: Codex

* remove all flags but , add runtime fallback with warning for uncompiled combination

* Update docs/build.md

Co-authored-by: Johannes Gäßler <johannesg@5d6.de>

* apply code review comments

---------

Co-authored-by: Johannes Gäßler <johannesg@5d6.de>
2026-09-09 12:50:08 +02:00
..
2026-07-30 16:14:37 +03:00
2026-06-04 08:02:54 +03:00
2026-07-30 16:14:37 +03:00
2026-07-30 16:14:37 +03:00