Lost in the Haystack: Smaller Needles are More Difficult for LLMs to Find.

Owen Bianchi1,2, Mathew J Koretsky1,2, Maya Willey1,2

  • 1Center for Alzheimer's Disease and Related Dementias, NIA, NIH.

Arxiv
|October 1, 2025
PubMed
Summary

Smaller gold contexts degrade large language model (LLM) performance on needle-in-a-haystack tasks. This amplifies positional sensitivity, challenging agentic systems integrating varied information.