* llama : only copy used KV cache in get / set state * switch to ggml for copying k, v * avoid designated initializers