arXiv cs.AIOctober 7, 2026
Vision Transformer Ensembles for Panoramic Street Segmentation
Excerpt
arXiv:2610.06063v1 Announce Type: cross Abstract: Semantic segmentation of street panoramas can support detailed descriptions of urban environments, yet small datasets and unequal training costs make model selection difficult. This paper presents the system used for a first place submission to the PalmCity challenge in the leaderboard snapshot dated 5 October 2026. Nine pretrained segmentation systems are compared using approximately equal computation budgets. The candidates include DeepLabV3+,