Playing social deduction games with reinforcement-fine-tuned large language models
Read the original at arxiv.org→arXiv:2610.04261v1 Announce Type: new Abstract: Reinforcement fine-tuning (RFT) is increasingly used in applications where large language models (LLMs) interact with humans and other agents. Here we use social...
Original headline: "Playing social deduction games with reinforcement fine-tuned large language models"
Coverage timeline
- Oct 6, 04:00 UTC arXiv cs.CL lead source Playing social deduction games with reinforcement fine-tuned large language models