Glossary · term
Semantic Compression
A technique that competes with long context: instead of keeping millions of tokens in the KV Cache, the model compresses data into dense semantic vectors that preserve logical relationships, which radically lowers GPU memory usage and makes it possible to handle large documents. A direction explored by, among others, Meta AI.
Training2025/26Wave 2 · 2024Maturity: 1/5
Maturity rationale
single R2 source, 2025-26 neologism