SSP-Net: Scalable sequential pyramid networks for real-Time 3D human pose regression - ETIS, équipe MIDI
Article Dans Une Revue Pattern Recognition Année : 2023

SSP-Net: Scalable sequential pyramid networks for real-Time 3D human pose regression

Résumé

In this paper we propose a highly scalable convolutional neural networks, end-to-end trainable, for real-time 3D human pose regression from still RGB images. We call this approach Scalable Sequential Pyramid Networks (SSP-Net) as it is trained with refined supervision at multiple scales in a sequential manner. Our network requires a single training procedure and is capable of producing its best predictions at 120 frames per second (FPS), or acceptable predictions at more than 200 FPS when cut at test time. We show that the proposed regression approach is invariant to the size of feature maps, allowing our method to perform multi-resolution intermediate supervisions and reaching results comparable to the state-of-the-art with very low resolution feature maps. We demonstrate the accuracy and the effectiveness of our method by providing extensive experiments on two of the most important publicly available datasets for 3D pose estimation, Human3.6M and MPI-INF-3DHP. Additionally, we provide relevant insights about our decisions on the network architecture and show its flexibility to meet the best precision-speed compromise.
Fichier principal
Vignette du fichier
2009.01998.pdf (1.44 Mo) Télécharger le fichier
Origine Fichiers produits par l'(les) auteur(s)

Dates et versions

hal-04124371 , version 1 (05-03-2024)

Identifiants

Citer

Diogo Carbonera Luvizon, Hedi Tabia, David Picard. SSP-Net: Scalable sequential pyramid networks for real-Time 3D human pose regression. Pattern Recognition, 2023, 142, pp.109714. ⟨10.1016/j.patcog.2023.109714⟩. ⟨hal-04124371⟩
204 Consultations
42 Téléchargements

Altmetric

Partager

More