summaryrefslogtreecommitdiff
path: root/examples/batched.swift/Sources
diff options
context:
space:
mode:
authorGeorgi Gerganov <ggerganov@gmail.com>2023-11-23 19:07:56 +0200
committerGitHub <noreply@github.com>2023-11-23 19:07:56 +0200
commit6b0a7420d03b9d13cb0e9439a01ce8476d8bf093 (patch)
treef184d281cb47e357e4ead4a93a0d1fe504c74bbe /examples/batched.swift/Sources
parentd103d935c0e75769a6a597f7a64cab72c6cc3e79 (diff)
llama : KV cache view API + better KV cache management (#4170)
* llama : keep track of used KV cells + better KV cache management * llama : zero KV cache used upon clear ggml-ci * llama : allow exporting a view of the KV cache (#4180) * Allow exporting a view of the KV cache * Allow dumping the sequences per cell in common * Track max contiguous cells value and position as well * Fix max contiguous empty cells index calculation Make dump functions deal with lengths or sequences counts > 10 better * Fix off by one error in dump_kv_cache_view * Add doc comments for KV cache view functions Eliminate cell sequence struct; use llama_seq_id directly Minor cleanups * common : add -dkvc arg for enabling kv cache dumps --------- Co-authored-by: Kerfuffle <44031344+KerfuffleV2@users.noreply.github.com>
Diffstat (limited to 'examples/batched.swift/Sources')
0 files changed, 0 insertions, 0 deletions