Signals 4 · free daily AI digest

KVShareArena: KV-Cache Reuse Across Contexts and Model Checkpoints

Paper recorded by Signals 4 on 2026-09-09 in cs.CL. Abstract reproduced from arXiv; link to the original below.

Published 2026-09-09 on arXiv · recorded by Signals 4 on 2026-09-10

Category: cs.CL · 自然语言处理 · first seen 2026-09-10

Abstract

LLM serving systems already reuse KV caches, but only when the reused text sits at the very start of the prompt. Two growing workloads break this condition: a retrieval-augmented generation server assembles a different set of retrieved chunks for every query, and a multi-agent coordinator reads reports written by other agents. Reused inside a new prompt, a cache carries the wrong positions and nev

Read on arXiv →

#85 most recent of 186 cs.CL papers we have recorded · ↑ newer: The Semantic Bottleneck: Leveraging Semantic Representations for Non-I · ↓ older: DiSCo: A Distribution-First Steering and Cultural Prior Evaluation Fra
Cite this page: KVShareArena: KV-Cache Reuse Across Contexts and Model Checkpoints: the #85 most recent of 186 cs.CL papers we have recorded (as of 2026-09-09). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/kvsharearena-kv-cache-reuse-across-contexts-and-model-checkpoints.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable: papers.json
Related: More cs.CL papers · arXiv signals · All papers · Today in AI
Get 4 AI signals a day by email — free.
Subscribe free → See all plans →
Get 4 AI signals a day by email — free
All models · All repos · By company · Daily editions