TY - RPRT TI - LLMCache: Layer-Wise Caching Strategies for Accelerated Reuse in Transformer Inference AU - Harsh Vardhan Bansal PY - 2025 DO - 10.1109/ised67359.2025.11405274 UR - https://arxiv.org/abs/2512.16843 ID - 2512.16843 ER -