Papers
arxiv:2609.39394

Can Computation from Earlier Problems Help LLMs Solve New Ones?

Published on Sep 30
· Submitted by
Devinhe
on Oct 5
Authors:
,
,
,
,
,

Abstract

Large language models often solve independent problems in the same conversation. Can computation from earlier problems help them solve new ones? To answer this question, we first conduct preliminary experiments showing that retained history can raise or lower later-turn accuracy, even within the same domain. To understand these effects, we use controlled replay to isolate internal state changes specific to each problem-history pairing. Across different histories, these changes preserve similar relationships among current problems. To improve reasoning under retained history, we introduce STAIR (Stale-Token Attention for Inter-query Reuse). STAIR captures keys and values from earlier response generation in a fixed bank. It learns to redirect current queries when they read this bank during prompt processing. The base model remains frozen; only 12,288 parameters are trained. Across three Qwen models and four benchmarks, STAIR improves average later-turn accuracy by up to 11.67 percentage points over the unmodified model with history.

Community

Paper author Paper submitter

Can computation from earlier problems help LLMs solve new ones? We show that retained history can both help and hurt reasoning after task switches. We introduce STAIR, which learns to re-address historical K/V states with only 12,288 trainable parameters while keeping the LLM frozen, improving later-turn accuracy by up to 11.67 percentage points.

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2609.39394
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 0

No model linking this paper

Cite arxiv.org/abs/2609.39394 in a model README.md to link it from this page.

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/2609.39394 in a dataset README.md to link it from this page.

Spaces citing this paper 0

No Space linking this paper

Cite arxiv.org/abs/2609.39394 in a Space README.md to link it from this page.

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.