← Back to all articles
arXiv cs.CLSeptember 24, 2026

Recognized but Not Produced: A Generation Benchmark for Culturally Specific Kinship Terms

Excerpt

arXiv:2609.26942v1 Announce Type: new Abstract: Current literature evaluates large language models (LLMs) on multilingual kinship understanding using multiple choice benchmarks, treating it as a recognition problem. We instead prompt five open weight LLMs to generate kinship terms in three non Western languages (Hindi, Tamil, and Korean) across two communicative tasks and pair this with a matched option-supported selection baseline. On identical relation language cells, GPT OSS120B selects the c