LGAIJul 3

No Time Like the Present: Agentic Test-Time Training for LLM Agents

arXiv:2607.034419.4
Predicted impact top 31% in LG · last 90 daysOriginality Incremental advance
AI Analysis

For LLM agent deployment in long-horizon tasks, aTTT mitigates performance degradation without requiring new training data.

LLM agents degrade over long episodes due to drift. Agentic Test-Time Training (aTTT) uses token-level reweighting to downweight repeated n-grams, improving success by up to 5.0 points on ALFWorld and 4.9 points on SWE-bench Lite with only 1.9x overhead.

LLM agents often degrade over long episodes: as trajectories grow, they revisit explored states, repeat failed actions, and lose strategies that previously worked. Test-time training (TTT) offers a way to adapt model weights to the evolving task state, but existing LLM TTT methods largely adapt once to a fixed input. We study continuous TTT in multi-turn agent episodes, where each update changes the policy that generates later training text. This creates a self-training loop that helps when new trajectory information appears, but can amplify drift when the agent gets stuck and repeatedly trains on similar text. We find that update-text repetition distinguishes these regimes and introduce Agentic Test-Time Training (aTTT), a token-level reweighting method that downweights the loss on tokens appearing in repeated $n$-grams from prior updates while leaving novel tokens fully weighted. To run such updates inside live episodes, we build a concurrent serving system using vLLM's runtime LoRA API, limiting overhead to 1.9$\times$ the no-TTT cost. aTTT improves success by up to 5.0 points on ALFWorld and 4.9 points on SWE-bench Lite. The gains concentrate where models already have task competence but drift over long trajectories, suggesting that aTTT mainly preserves existing competence rather than teaching new abilities.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes