llama : refactor tensor offloading as callback

This commit is contained in:
Georgi Gerganov
2023-10-29 12:35:07 +02:00
parent da936188d8
commit 15267192c0
+704 -726
View File
File diff suppressed because it is too large Load Diff