arXiv cs.LGOctober 1, 2026
Which Models Work Well Together? Measuring Heterogeneity for LLM Team Selection
Excerpt
arXiv:2609.38274v1 Announce Type: cross Abstract: The performance ceiling of an LLM team is constrained not only by individual model capabilities, but also by inter-member error resonance and predictive differences. Although heterogeneous teaming is often observed to be effective in practice, existing approaches lack complementarity metrics that are computable, interpretable, and optimizable, leaving team composition to rely on heuristics. We propose a heterogeneity-driven team selection framewo