arXiv cs.LGOctober 2, 2026
Iterative Policy Refinement through Semantic Rollout Analysis
Excerpt
arXiv:2610.01652v1 Announce Type: new Abstract: Structured policies improve efficiency, robustness, and interpretability in imitation learning by introducing task-specific inductive bias, but existing structure generation methods rely either on extensive human input or on static domain knowledge encoded in LLMs, which may be inconsistent with the expert demonstrations. We propose a closed-loop framework that iteratively refines structured policies using LLM-guided analysis of policy rollouts. By