mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-09-01 01:51:19 +02:00
1b2f992cd2
* test-backend-ops : use flops for some performance tests - parallelize tensor quantization - use a different set of cases for performance and correctness tests - run each test for at least one second