mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-08-29 00:21:21 +02:00
a035a88878
* * server: add spec-decode counters to /metrics endpoint * server: fixed review comments and now aligned param names exactly with vLLM.