<p>This paper describes a dataset consisting of manually annotated nouns from a corpus of longitudinal day-long audio and hour-long video recordings collected monthly from 44 babies from age 6 months to age 17 months. This dataset was created as part of a larger project, called SEEDLingS, that examines the development of infants’ language comprehension before and after their first birthday, from earliest comprehension to the early days of word production. This paper provides an overview of the corpus, describes how and why the nouns from the corpus were annotated, and discusses considerations for the reuse of this dataset for future work. The described annotations and relevant metadata are publicly available alongside this manuscript.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

A year of nouns from English-learning infants’ daily lives: The SEEDLingS-Nouns dataset

  • Evgenii Kalenkovich,
  • Sharath Koorathota,
  • Shaelise Tor,
  • Andrei Amatuni,
  • Shannon Egan-Dailey,
  • Charlotte Moore,
  • Catherine Laing,
  • Hallie Garrison,
  • Gladys Baudet,
  • Federica Bulgarelli,
  • Sarp Uner,
  • Lillianna Righter,
  • Elika Bergelson

摘要

This paper describes a dataset consisting of manually annotated nouns from a corpus of longitudinal day-long audio and hour-long video recordings collected monthly from 44 babies from age 6 months to age 17 months. This dataset was created as part of a larger project, called SEEDLingS, that examines the development of infants’ language comprehension before and after their first birthday, from earliest comprehension to the early days of word production. This paper provides an overview of the corpus, describes how and why the nouns from the corpus were annotated, and discusses considerations for the reuse of this dataset for future work. The described annotations and relevant metadata are publicly available alongside this manuscript.