The words that children hear play a crucial role in the development of the vocabulary they can understand and use in communication. In this work, we aim to study the size of the lexicon to which Italian preschool children are exposed. We constructed a corpus of 224 diverse sources (children’s books, songs, cartoons, and parentchild interactions, comprising about 450,000 occurrences) and applied non-parametric frequency-of-frequencies (FoF) estimators of the latent lexicon size. Our results indicate that the Italian child-directed lexicon comprises at least 20,000 lemmas.
- Size estimation of lexicon directed to Italian preschool children via a frequencies-of-frequencies approach
- S CostantiniP PasqualettiP RinaldiD ChiarellaM FavillaM MajoranoLorenzo SpreaficoL Tardella
- Statistical Science: From Theory to Applied Research III, pp.62-68
- Martella F, Arima S, Marino M, Mollica C
- 9783032308801
- 9783032308818
- 3059-2135
- 3059-2143
- Italian Statistical Society Series on Advances in Statistics
- Springer
- 7
- 978-3-032-30880-1
(UNIBZ)98429169
991007387695601241 - n.a.
- Faculty of Education
- English
- Book chapter
- Costantini S, Pasqualetti P, Rinaldi P, Chiarella D, Favilla M, Majorano M, Spreafico L, Tardella L
- Editors/Supervisors: Martella F, Arima S, Marino M, Mollica C