How poor is the stimulus? Evaluating hierarchical generalization in neural networks trained on child-directed speech (2301.11462v2)

Published 26 Jan 2023 in cs.CL

Abstract: When acquiring syntax, children consistently choose hierarchical rules over competing non-hierarchical possibilities. Is this preference due to a learning bias for hierarchical structure, or due to more general biases that interact with hierarchical cues in children's linguistic input? We explore these possibilities by training LSTMs and Transformers - two types of neural networks without a hierarchical bias - on data similar in quantity and content to children's linguistic input: text from the CHILDES corpus. We then evaluate what these models have learned about English yes/no questions, a phenomenon for which hierarchical structure is crucial. We find that, though they perform well at capturing the surface statistics of child-directed speech (as measured by perplexity), both model types generalize in a way more consistent with an incorrect linear rule than the correct hierarchical rule. These results suggest that human-like generalization from text alone requires stronger biases than the general sequence-processing biases of standard neural network architectures.

Authors (4)

Aditya Yedetore (1 paper)
Tal Linzen (73 papers)
Robert Frank (23 papers)
R. Thomas McCoy (33 papers)

Citations (13)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Tweets

https://twitter.com/AdityaYedetore/status/1843474694193983722

How poor is the stimulus? Evaluating hierarchical generalization in neural networks trained on child-directed speech (2301.11462v2)

Summary

Related Papers

Tweets