Glossary · term102 of 202
LLMLingua
A prompt-compression method that reports roughly 20x token reduction at about a 1.5-point accuracy cost, measured on short-answer reasoning benchmarks rather than agent tasks.
Related terms
At a glance
- cited by
- 1 post
- categories
- 1
- first used
- jul 2026