mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-09-17 08:19:38 +02:00
model : more uniform output id handling (#14275)
* model : more uniform output id handling ggml-ci * cont : revert n_outputs < n_tokens optimization ggml-ci * cont : fix out_ids initialization ggml-ci
This commit is contained in:
+432
-415
File diff suppressed because it is too large
Load Diff
Reference in New Issue
Block a user