MuTr: Multi-Stage Transformer for Hand Pose Estimation from Full-Scene Depth Image.

Sensors (Basel)

Department of Cybernetics and New Technologies for the Information Society, University of West Bohemia Technická 8, 301 00 Pilsen, Czech Republic.

Published: June 2023

AI Article Synopsis

Article Abstract

This work presents a novel transformer-based method for hand pose estimation-DePOTR. We test the DePOTR method on four benchmark datasets, where DePOTR outperforms other transformer-based methods while achieving results on par with other state-of-the-art methods. To further demonstrate the strength of DePOTR, we propose a novel multi-stage approach from full-scene depth image-MuTr. MuTr removes the necessity of having two different models in the hand pose estimation pipeline-one for hand localization and one for pose estimation-while maintaining promising results. To the best of our knowledge, this is the first successful attempt to use the same model architecture in standard and simultaneously in full-scene image setup while achieving competitive results in both of them. On the NYU dataset, DePOTR and MuTr reach precision equal to 7.85 mm and 8.71 mm, respectively.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC10305187PMC
http://dx.doi.org/10.3390/s23125509DOI Listing

Publication Analysis

Top Keywords

hand pose
12
pose estimation
8
full-scene depth
8
mutr multi-stage
4
multi-stage transformer
4
hand
4
transformer hand
4
pose
4
estimation full-scene
4
depth image
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!