# Preprint finds LLM memory scores depend heavily on how history is rendered

_Wednesday, August 26, 2026 at 12:00 AM EDT · Science · Latest · Tier 2 — Notable_

A new arXiv preprint introduces RENDER, a benchmark control that holds a conversation fixed while changing how its history is presented to an answering model. Across 500 LongMemEval questions and nine models, the authors report that matched-budget resolved packets outperformed recency-truncated raw dialogue by 42.4 to 72.6 points. The study also found mixed model-specific significance after judge rescoring. The results are a preprint evaluation, not evidence of an operational memory-system deployment.

## Sources

- [cs.AI updates on arXiv.org](https://arxiv.org/abs/2608.23568)

---
Canonical: https://techandbusiness.org/newswire/8gYRqYRhDaPrs96xPPUrc3
Retrieved: 2026-08-26T13:28:49.479Z
Publisher: Tech & Business (techandbusiness.org)
