arXiv cs.AIOctober 7, 2026
LexiHorizon: Stabilizing Reinforcement Learning for Long-Horizon Deep Search
Excerpt
arXiv:2610.05119v1 Announce Type: new Abstract: Deep search agents tackle complex knowledge tasks through iterative retrieval, multi-hop reasoning, and evidence synthesis across multiple sources. Existing approaches typically assume relatively stable retrieval systems and operate over short-horizon tool interaction. However, when retrieval is sensitive to query formulation, even a semantically appropriate query may fail to surface critical evidence because of mismatched entity names, aliases, or