Skip to content

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Open access Jul 2026

HiKV: Hierarchical Importance-Aware KV Cache With Hardware Acceleration for LLM Decoding

HiKV is a novel algorithm-hardware co-design that exploits KV cache redundancy through hierarchical importance awareness and outperforms state-of-the-art importance-based methods by achieving an additional 1.87x reduction in external memory accesses.

Chao Fang, Jun Yin, Man Shi et al. · 0 citations