TF-Engram: A Train-Free Engram with SSD-Backed Memory for Large Language Models
The introduction of TF-Engram marks a significant advancement in the field of large language models (LLMs), providing a train-free Engram system that enhances memory storage and retrieval efficiency. This system utilizes SSD-backed memory and Early-Exit Guided Predictive Prefetching to improve performance during autoregressive decoding, resulting in a notable increase in downstream task scores for the Qwen3-0.6B model.