No AI summary available for this article.
Why It Matters
Language-model (LM) harnesses enable LMs to operate effectively over long contexts using additional compute.
Provenance
Discovered via ArXiv and published by ArXiv.
Key Claims
Original description
Language-model (LM) harnesses enable LMs to operate effectively over long contexts using additional compute. However, existing long-context evaluations are insufficient for distinguishing modern harnesses, reflected by saturated accuracy across harnesses and largely similar evaluation costs. In this paper, we introduce a benchmark for evaluating both the effectiveness and efficiency of long-context harnesses. Our tasks require diverse retrieval strategies, including lexical search and semantic matching, together with strategic and adaptive reasoning over global and local context. Much of the c...
Discovered via ArXiv
Research papers and preprints from arXiv.
Publisher: arxiv.org
ID: http://arxiv.org/abs/2609.38137v1 · Indexed about 1 hour ago