Skip to content

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

The Key Handoff: Retrieval in Hybrid Language Models

A two-hop question makes a language model retrieve twice: once to produce a bridge entity, and once to retrieve with it. Transformers resolve that entity in their early layers. Hybrid models replace most of the attention with a recurrent state, so where the key becomes usable, and where it is spent on an answer, is not...

Kaan Kale, Oğuzhan Başer, Sriram Vishwanath · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.