Random forests, sound symbolism and Pokémon evolution

Alexander James Kilpatrick; Aleksandra Ćwiek; Shigeto Kawahara

doi:10.1371/journal.pone.0279350

Random forests, sound symbolism and Pokémon evolution

PLoS One. 2023 Jan 4;18(1):e0279350. doi: 10.1371/journal.pone.0279350. eCollection 2023.

Authors

Alexander James Kilpatrick¹, Aleksandra Ćwiek², Shigeto Kawahara³

Affiliations

¹ International Communication, Nagoya University of Commerce and Business, Nagoya, Aichi, Japan.
² Department Leibniz-Zentrum Allgemeine Sprachwissenschaft, Berlin, Germany.
³ Institute of Cultural and Linguistic Studies, Keio University, Tokyo, Japan.

Abstract

This study constructs machine learning algorithms that are trained to classify samples using sound symbolism, and then it reports on an experiment designed to measure their understanding against human participants. Random forests are trained using the names of Pokémon, which are fictional video game characters, and their evolutionary status. Pokémon undergo evolution when certain in-game conditions are met. Evolution changes the appearance, abilities, and names of Pokémon. In the first experiment, we train three random forests using the sounds that make up the names of Japanese, Chinese, and Korean Pokémon to classify Pokémon into pre-evolution and post-evolution categories. We then train a fourth random forest using the results of an elicitation experiment whereby Japanese participants named previously unseen Pokémon. In Experiment 2, we reproduce those random forests with name length as a feature and compare the performance of the random forests against humans in a classification experiment whereby Japanese participants classified the names elicited in Experiment 1 into pre-and post-evolution categories. Experiment 2 reveals an issue pertaining to overfitting in Experiment 1 which we resolve using a novel cross-validation method. The results show that the random forests are efficient learners of systematic sound-meaning correspondence patterns and can classify samples with greater accuracy than the human participants.

Copyright: © 2023 Kilpatrick et al. This is an open access article distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited.

Publication types

Research Support, Non-U.S. Gov't

MeSH terms

Algorithms
Humans
Random Forest*
Sound
Symbolism
Video Games*

Grants and funding

AK - Grant obtained from Japan Society for the Promotion of Science (Tokyo, JP) GRANT_NUMBER: 20K13055 https://www.jsps.go.jp/english/index.html The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript.