← Back to all articles
arXiv cs.LGOctober 2, 2026

Flowing Faster to Coordinate: One-Step Online Multi-Agent Flow Policies

Excerpt

arXiv:2610.01882v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) provides a powerful framework for learning coordinated behaviors through interactions with the environment. Developing MARL policies requires balancing expressive modeling of complex and multimodal action distributions with efficient training and execution. Generative policies, particularly diffusionbased policies, can faithfully capture complex and multimodal behaviors, but costly iterative sampling hinder