Give a model 20 documents and put the answer in the middle: performance craters. The information is IN the context — the model just can't find it there.
Everyone measured how long the input could be; almost nobody measured where the model actually looks.
The finding inverted assumptions: with the answer document fixed and only its position varied, accuracy is highest at the very beginning and end of the context — and dips sharply in between. The model behaves less like a careful reader and more like a skimmer whose attention collapses toward the anchors of the sequence. Performance is often highest when relevant information occurs at the beginning or end of the input — and middles are where facts go to hide.
The gap the paper exposed: fitting a context is not the same as using a context.
The model reads like someone skimming a morning paper under deadline: the headline (start) and the back page (recency) get real attention; page 14 of the sports section could contain the winning lottery numbers and they'd still miss it. The information was delivered. The reading wasn't.
Deliberately minimal tasks so position — not reasoning — is the only variable.
The mechanisms the paper and its successors pointed at.
The paper documents the phenomenon and explores causes: models attend preferentially to primacy (early tokens anchor the representation of everything after) and recency (late tokens are one step from the output position). The middle gets squeezed from both sides — an attention-allocation artifact baked into training on naturally edge-weighted data (documents begin with topic statements; answers often follow questions). The paper also shows a closed-book twist: models given no documents sometimes outperform models given the answer in the middle — the middle document actively hurts.
One shape summarized a generation of long-context failure.
Needle-in-a-haystack plots and context-ordering discipline are this paper's descendants.
Check your understanding of the key concepts from Lost in the Middle.
Everything you need to remember about this paper.