main : add self-extend support (#4815)

* examples : add passkey test * passkey : better prints * passkey : select pass key pos from CLI * passkey : simplify n_past logic * llama : "self-extend"-like context extension * passkey : add comment * main : add Self-Extend support * llama : add comment about llama_kv_cache_seq_div
2025-08-25 09:38:35 -04:00 · 2024-01-08 11:18:32 +02:00
parent b0034d93ce
commit 52531fdff8
4 changed files with 87 additions and 24 deletions
--- a/llama.h
+++ b/llama.h
@@ -484,6 +484,10 @@ extern "C" {
                       llama_pos   p1,
                       llama_pos   delta);

+    // Integer division of the positions by factor of `d > 1`
+    // If the KV cache is RoPEd, the KV data is updated accordingly
+    // p0 < 0 : [0,  p1]
+    // p1 < 0 : [p0, inf)
    LLAMA_API void llama_kv_cache_seq_div(
            struct llama_context * ctx,
                    llama_seq_id   seq_id,