HyperAIHyperAI

Command Palette

Search for a command to run...

3 months ago

Visual Alignment Constraint for Continuous Sign Language Recognition

Yuecong Min Aiming Hao Xiujuan Chai Xilin Chen

Visual Alignment Constraint for Continuous Sign Language Recognition

Abstract

Vision-based Continuous Sign Language Recognition (CSLR) aims to recognize unsegmented signs from image streams. Overfitting is one of the most critical problems in CSLR training, and previous works show that the iterative training scheme can partially solve this problem while also costing more training time. In this study, we revisit the iterative training scheme in recent CSLR works and realize that sufficient training of the feature extractor is critical to solving the overfitting problem. Therefore, we propose a Visual Alignment Constraint (VAC) to enhance the feature extractor with alignment supervision. Specifically, the proposed VAC comprises two auxiliary losses: one focuses on visual features only, and the other enforces prediction alignment between the feature extractor and the alignment module. Moreover, we propose two metrics to reflect overfitting by measuring the prediction inconsistency between the feature extractor and the alignment module. Experimental results on two challenging CSLR datasets show that the proposed VAC makes CSLR networks end-to-end trainable and achieves competitive performance.

Code Repositories

hulianyuyy/Temporal-Lift-Pooling
pytorch
Mentioned in GitHub
ycmin95/VAC_CSLR
Official
pytorch
Mentioned in GitHub

Benchmarks

BenchmarkMethodologyMetrics
sign-language-recognition-on-rwth-phoenixVAC
Word Error Rate (WER): 22.1

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp
Visual Alignment Constraint for Continuous Sign Language Recognition | Papers | HyperAI