Aller au contenu principal
Accès ouvert déclaré 2025 preprint

Characterizing young children’s everyday activities using video question-answering models

0Citations signalées — pas une note de qualité
1Institutions déclarées
1Pays d’affiliation déclarés

Résumé fourni par la source

Children are remarkably efficient learners compared to our most advanced computational models of learning. One key difference is that children seem to leverage regularities in the activities (e.g., eating) in which they participate to learn about words or objects (e.g., “pomegranate”), even under skewed, long-tailed distributions. While everyday activities have long been theorized to be important as supports for children’s learning, our understanding of the types, frequencies, and rhythms of these activities has been out of reach due to both a lack of naturalistic video datasets and the necessity for manual annotations. Here, we use the recent release of a large, egocentric dataset of children’s everyday experience (BabyView) (N=31 children, N=868 hours) and capitalize on innovations in video question-answering (VideoQA) models to quantify the what and where of children’s everyday experiences. Using these models, we classify both the activities (e.g., eating, dancing, exploring) and physical locations (e.g., living room, garage) in the infant view and generate natural-language descriptions for contiguous 10-second videos across the entire dataset. Notably, we find that (a) some activities and locations occur much more frequently than others, yet (b) there is wide variation across children. Moreover, (c) activities and locations exhibit structured transition probabilities (e.g., cooking often precedes eating), and (d) may decompose into distinct sub-clusters (e.g., different subtypes of reading). Compared with prior work analyzing static image content, our work highlights the advances possible by using VideoQA models to analyze the dynamic nature of children’s experiences. Our results provide a better understanding of children’s learning input in everyday contexts, informing developmentally-inspired models of early learning and cognitive development.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Characterizing young children’s everyday activities using video question-answering models
Date Crossref
10/10/2025
Éditeur
Center for Open Science
Type
posted-content

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude et ne compte pas comme une seconde source scientifique indépendante.

Institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Sujets associés

Innovative Teaching and Learning MethodsSpeech and dialogue systems

BNTIC News n’est pas le producteur de ces données. Recherche à la demande dans Crossref et Europe PMC, sans clé ; OpenAlex reste optionnel. Aucun service payant requis, aucune réponse conservée. Sources et limites.