Playing social deduction games with reinforcement fine-tuned large language models
arXiv:2610.04261v1 Announce Type: new Abstract: Reinforcement fine-tuning (RFT) is increasingly used in applications where large language models (LLMs) interact with humans and other agents. Here we use social deduction games to study how RFT changes LLMs' social behaviour. We let fine-tuned and ba…
Read the full story at arXiv cs.CL ↗
Timeline · 1 report
- 2026-10-06 04:00 · arXiv cs.CL
Playing social deduction games with reinforcement fine-tuned large language models