← Back to all articles
arXiv cs.AIOctober 7, 2026

Playing social deduction games with reinforcement fine-tuned large language models

Excerpt

arXiv:2610.04261v1 Announce Type: cross Abstract: Reinforcement fine-tuning (RFT) is increasingly used in applications where large language models (LLMs) interact with humans and other agents. Here we use social deduction games to study how RFT changes LLMs' social behaviour. We let fine-tuned and base LLM agents play hidden-role games that require hidden-state inference, social reading and vote steering. Our results show that LLM agents do not reliably acquire social-deduction ability by direct