← Back to all articles
arXiv cs.LGOctober 1, 2026

Which Models Work Well Together? Measuring Heterogeneity for LLM Team Selection

Excerpt

arXiv:2609.38274v1 Announce Type: cross Abstract: The performance ceiling of an LLM team is constrained not only by individual model capabilities, but also by inter-member error resonance and predictive differences. Although heterogeneous teaming is often observed to be effective in practice, existing approaches lack complementarity metrics that are computable, interpretable, and optimizable, leaving team composition to rely on heuristics. We propose a heterogeneity-driven team selection framewo