SALT: Salience-Aware Lexical Trie for Long-Context Compression
cs.PF, cs.AI, cs.LG
Submitted: 2026-07-20
Updated: 2026-08-31
Code: https://github.com/oteomamo/SALT
Terminology
Sources
- LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
- The Llama 3 Herd of Models
- RULER: What's the Real Context Size of Your Long-Context Language Models?
- FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration
- Towards General Text Embeddings with Multi-stage Contrastive Learning
- Compressive Transformers for Long-Range Sequence Modelling
- C-Pack: Packed Resources For General Chinese Embeddings
- Sentinel: Decoding Context Utilization via Attention Probing for Efficient LLM Context Compression