mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-09-14 18:59:14 +02:00
78d2f52468
* cuda : concat implementation for quantized types * chore : apply am17an clever suggestion to shorten the code --------- Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com>