Post #963 200 Apr 15, 2026, 12:36 UTC i think this can be upgraded all the way to the infinite-context modelshttps://microsoft.github.io/memento/blogpost/ microsoft.github.io Memento: Teaching LLMs to Manage Their Own Context Memento teaches language models to manage their own context by segmenting reasoning into blocks, compressing each into a dense memento, and reasoning forward with sharply lower KV cache usage. ❤ 1