British Sign Language Recognition via Late Fusion of Computer Vision and Leap Motion with Transfer Learning to American Sign Language

Jordan J Bird; Anikó Ekárt; Diego R Faria

doi:10.3390/s20185151

British Sign Language Recognition via Late Fusion of Computer Vision and Leap Motion with Transfer Learning to American Sign Language

Sensors (Basel). 2020 Sep 9;20(18):5151. doi: 10.3390/s20185151.

Authors

Jordan J Bird¹, Anikó Ekárt², Diego R Faria¹

Affiliations

¹ ARVIS Lab-Aston Robotics Vision and Intelligent Systems, Aston University, Birmingham B4 7ET, UK.
² School of Engineering and Applied Science, Aston University, Birmingham B4 7ET, UK.

Abstract

In this work, we show that a late fusion approach to multimodality in sign language recognition improves the overall ability of the model in comparison to the singular approaches of image classification (88.14%) and Leap Motion data classification (72.73%). With a large synchronous dataset of 18 BSL gestures collected from multiple subjects, two deep neural networks are benchmarked and compared to derive a best topology for each. The Vision model is implemented by a Convolutional Neural Network and optimised Artificial Neural Network, and the Leap Motion model is implemented by an evolutionary search of Artificial Neural Network topology. Next, the two best networks are fused for synchronised processing, which results in a better overall result (94.44%) as complementary features are learnt in addition to the original task. The hypothesis is further supported by application of the three models to a set of completely unseen data where a multimodality approach achieves the best results relative to the single sensor method. When transfer learning with the weights trained via British Sign Language, all three models outperform standard random weight distribution when classifying American Sign Language (ASL), and the best model overall for ASL classification was the transfer learning multimodality approach, which scored 82.55% accuracy.

Keywords: late fusion; multimodality; sign language recognition.

MeSH terms

Computers
Humans
Machine Learning*
Movement
Neural Networks, Computer*
Sign Language*
United Kingdom
United States