Files
llama.cpp/tools
Georgi Gerganov 3637576288 server : disable speculative decoding for SWA models (#13970)
* server : use swa-full fo draft context

ggml-ci

* server : disable speculative decoding for SWA models
2025-06-02 21:34:40 +03:00
..
2025-05-25 15:35:53 +03:00
2025-05-30 16:25:45 +03:00
2025-05-30 16:25:45 +03:00