arXiv cs.AIOctober 7, 2026
Dr-CiK: A Benchmark testing Deep Research for Context-Aided Forecasting
Excerpt
arXiv:2605.27904v2 Announce Type: replace Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively discovered from noisy, heterogeneous information sources. Yet existing context-aided forecasting benchmarks typically assume that the supporting context is already provided, leaving open whether agents can identify it on their own. Therefore, we introduce Dr-CiK, a benchmark for evaluating whether