arXiv cs.LGOctober 1, 2026
T-Router: Learning Thalamic Routing for Reasoning with Parameter-Efficient Reinforcement Learning
Excerpt
arXiv:2609.39109v1 Announce Type: new Abstract: Parameter-efficient reinforcement learning aims to improve reasoning with a compact trainable interface to a pretrained model. We introduce the Thalamic Router (T-Router), which concentrates adaptation on the reuse of completed computations. A compressed, addressable bank preserves block changes; a depth-recurrent controller conditions their selection and relative-scale writeback. This coupling gives thalamic context-dependent routing a concrete co