← Back to all articles
arXiv cs.LGOctober 1, 2026

Self-Evolving Harness on Multiple Tasks with the Agent as Its Own Optimizer

Excerpt

arXiv:2609.38372v1 Announce Type: cross Abstract: A harness is the code around a language-model agent that organizes prompts, calls tools, manages context, and controls execution. As models grow stronger, recent work has begun to let agents improve their own harnesses, a line of work known as self-evolving harnesses. In most existing methods, a separate proposer running on a human-designed harness modifies the solver's harness, and a separate harness is evolved for each benchmark. Real-world tas