arXiv cs.LGOctober 2, 2026
Decision Titan: Test-Time Training for Long-Term Memory in Offline Reinforcement Learning
Excerpt
arXiv:2610.01513v1 Announce Type: cross Abstract: Long-term dependencies remain a major challenge for sequential decision-making in the field of AI: RNNs suffer from vanishing gradients and the limited expressivity of vector-based hidden states, whilst Transformer-based models are limited by the quadratic scaling of attention. Recent work has proposed tackling this problem with the Test-Time Training (TTT) framework, which stores episodic memories in the parameters of a neural network through gr