Aller au contenu principal
Accès ouvert déclaré 2016 preprint

Temporal Convolutional Networks: A Unified Approach to Action\n Segmentation

14Citations signalées, ce qui n’est pas une note de qualité
0Institutions déclarées
0Pays d’affiliation déclarés

Le résumé fourni par la source

The dominant paradigm for video-based action segmentation is composed of two\nsteps: first, for each frame, compute low-level features using Dense\nTrajectories or a Convolutional Neural Network that encode spatiotemporal\ninformation locally, and second, input these features into a classifier that\ncaptures high-level temporal relationships, such as a Recurrent Neural Network\n(RNN). While often effective, this decoupling requires specifying two separate\nmodels, each with their own complexities, and prevents capturing more nuanced\nlong-range spatiotemporal relationships. We propose a unified approach, as\ndemonstrated by our Temporal Convolutional Network (TCN), that hierarchically\ncaptures relationships at low-, intermediate-, and high-level time-scales. Our\nmodel achieves superior or competitive performance using video or sensor data\non three public action segmentation datasets and can be trained in a fraction\nof the time it takes to train an RNN.\n

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

La source scientifique ouverte est momentanément indisponible.

Les sujets associés

Human Pose and Action RecognitionAdvanced Neural Network ApplicationsVideo Surveillance and Tracking Methods

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.