HOST: If I review long contracts, why should I care how a language model handles memory? EXPERT: It could help you ask whether an assistant can still consult an early clause when you question it much later. This paper explains the design choices; it does not test contract review. HOST: What choice does the model make about that early clause? EXPERT: One design keeps earlier pieces of text separately. Another looks at only selected pieces. Another folds earlier information into a running internal state, so the original pieces are no longer separately selectable. HOST: So if it selects fewer pieces, do you think it might skip the one I need? EXPERT: Yes, that is the trade-off the authors describe. Selecting fewer pieces can reduce work, but the needed evidence has to make the selection. HOST: So what did the authors actually find across models? EXPERT: They cataloged 59 documented release records. In a separate snapshot of 11 selected high-performing open-weight model endpoints, every architecture still had a way to retrieve individual tokens, pieces of text, even though their overall designs varied. HOST: So does that tell me which model will actually find a cancellation deadline? EXPERT: No, the authors say their selection is curated, and the comparison doesn't show that a memory design causes better answers. The practical takeaway is to ask what a system retains and retrieves, then evaluate the actual task separately.