HyperAIHyperAI

Command Palette

Search for a command to run...

3 months ago

Fusing Posture and Position Representations for Point Cloud-Based Hand Gesture Recognition

{Mattias P Heinrich Alexander Bigalke}

Abstract

Hand gesture recognition can benefit from directly processing 3D point cloud sequences, which carry rich geometric information and enable the learning of expressive spatio-temporal features. However, currently employed single-stream models cannot sufficiently capture multi-scale features that include both fine-grained local posture variations and global hand movements. We therefore propose a novel dual-stream model, which decouples the learning of local and global features. These are eventually fused in an LSTM for temporal modelling. To induce the global and local stream to capture complementary position and posture features, we propose the use of different 3D learning architectures in both streams. Specifically, state-of-the-art point cloud networks excel at capturing fine posture variations from raw point clouds in the local stream. To track hand movements in the global stream, we combine an encoding with residual basis point sets and a fully-connected DenseNet. We evaluate the method on the Shrec'17 and DHG dataset and report state-of-the-art results at a reduced computational cost. Source code is available at https://github.com/multimodallearning/hand-gesture-posture-position.

Benchmarks

BenchmarkMethodologyMetrics
hand-gesture-recognition-on-dhg-14FPPR-PCD
Accuracy: 92.0
hand-gesture-recognition-on-dhg-28FPPR-PCD
Accuracy: 91.7
hand-gesture-recognition-on-shrec-2017FPPR-PCD
14 Gestures Accuracy: 96.1
28 Gestures Accuracy: 95.2

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp
Fusing Posture and Position Representations for Point Cloud-Based Hand Gesture Recognition | Papers | HyperAI