Papers/2610.10552
🧪 Test?View on arXiv

Freeze the Decoder, Heal the Encoder: Parameter-Efficient Adaptation for SVD-Based KV-Cache Compression

Not provided in the abstract

fine-tuningKV-cache compressionparameter efficiencymodel adaptation
2610.10552
Builder Relevance
80%
2h ago

Abstract

The paper discusses the flaws in using a shared learning rate for parameter-efficient fine-tuning and proposes a method for low-rank KV-cache compression that achieves significant memory savings.

Reality Card

Core Claim

Freezing the decoder and healing only the encoder achieves parity with other methods while using 3x fewer trainable parameters and 3x less optimizer-state memory.

Method / Result

Achieved a real measured saving of 3x fewer trainable parameters and 3x less optimizer-state memory.

Limitations

The results are based on a specific compression ratio and may not generalize across different configurations or models.

Paper to code

Verified implementation resources so builders can test the paper’s claims instead of stopping at the abstract.

No verified implementation link has been attached yet. AIBuzzHub will keep this panel separate from unverified search results.
← Back to all papers