SOTOPIA-π: Interactive Learning of Socially Intelligent Language Agents

Ruiyi Wang; Haofei Yu; Wenxin Zhang; Zhengyang Qi; Maarten Sap; Yonatan Bisk; Graham Neubig; Hao Zhu

doi:10.18653/v1/2024.acl-long.698

SOTOPIA-π: Interactive Learning of Socially Intelligent Language Agents

Ruiyi Wang, Haofei Yu, Wenxin Zhang, Zhengyang Qi, Maarten Sap, Yonatan Bisk, Graham Neubig, Hao Zhu

Abstract

Humans learn social skills through both imitation and social interaction. This social learning process is largely understudied by existing research on building language agents. Motivated by this gap, we propose an interactive learning method, SOTOPIA-π, that improves the social intelligence of language agents. This method leverages behavior cloning and self-reinforcement based training on filtered social interaction data according to large language model (LLM) rating. We show that our training method allows a 7B LLM to reach the social goal completion ability of an expert model (GPT-4-based agent) without the loss of more generic abilities, such as the ability to answer knowledge-based questions. We also demonstrate that this training paradigm uncovers some weaknesses in standard evaluation and safety training paradigms that (1) LLM-based evaluation of social intelligence overestimates the abilities of the language agents trained specifically for social interaction, and that (2) despite not training for better safety or question answering (QA) ability, our methods improve the safety of language agents and maintain general QA ability on the MMLU benchmark.

Anthology ID:: 2024.acl-long.698
Volume:: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Month:: August
Year:: 2024
Address:: Bangkok, Thailand
Editors:: Lun-Wei Ku, Andre Martins, Vivek Srikumar
Venue:: ACL
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 12912–12940
Language:
URL:: https://aclanthology.org/2024.acl-long.698
DOI:: 10.18653/v1/2024.acl-long.698
Bibkey:
Cite (ACL):: Ruiyi Wang, Haofei Yu, Wenxin Zhang, Zhengyang Qi, Maarten Sap, Yonatan Bisk, Graham Neubig, and Hao Zhu. 2024. SOTOPIA-π: Interactive Learning of Socially Intelligent Language Agents. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 12912–12940, Bangkok, Thailand. Association for Computational Linguistics.
Cite (Informal):: SOTOPIA-π: Interactive Learning of Socially Intelligent Language Agents (Wang et al., ACL 2024)
Copy Citation:
PDF:: https://aclanthology.org/2024.acl-long.698.pdf

PDF Cite Search