Still: Amortized KV Cache Compaction in a Single Forward Pass
You can shrink a language model's KV cache by 200×, in a single forward pass, and it still answers correctly.
@article{iverson2026still,
title = {Still: Amortized KV Cache Compaction in a Single Forward Pass},
author = {O'Neill, C. and Sandomirsky, A. and Partridge, H. and Jayasekara, M. and Kirkby, M.},
journal = {Base Labs},
year = {2026},
number = {001},
url = {https://baselabs-5feq68zvh-blueprintbyb10.vercel.app/001},
}



