Cache the Encoder Within: Compact, Reusable Memory across LLM Queries
Published in ArXiv preprint, 2026
EncBank reuses lower-layer document representations across queries and stores them at reduced precision for an adapted upper-layer reader.
Recommended citation: Liu et al., 2026.
Download Paper