[ICLR 2026] Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs
-
Updated
Aug 31, 2026 - Python
[ICLR 2026] Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs
Un-LOCC: Universal Lossy Optical Context Compression for Vision-Based Language Models Achieve nearly 3x token compression at over 93% retrieval accuracy using existing Vision-Language Models.
Silence of the RAM: constant-memory safety probing for long-context LLMs. A streaming hard-max probe with a measured Theta(min(C,N)) memory law (flat 19.1 MiB to N=131,072 vs 652.1 MiB for softmax pooling), the subgradient obstruction it creates, and the fragmentation attack that defeats it.
A next-gen architecture for LLM memory handling.
🌟 Compress text using images to achieve nearly 3x reduction in tokens with over 93% retrieval accuracy for Vision-Language Models.
To associate your repository with the long-context-llm topic, visit your repo's landing page and select "manage topics."