Logo image
Open Research University homepage
Surrey researchers Sign in
Changing the Representation: Examining Language Representation for Neural Sign Language Production
Preprint

Changing the Representation: Examining Language Representation for Neural Sign Language Production

Harry Walsh, Ben Saunders and Richard Bowden
arXiv.org
Cornell University Library, arXiv.org
16/09/2022

Abstract

Gloss Language translation Natural language processing Performance enhancement Representations Sentences Translating Sign Language
Neural Sign Language Production (SLP) aims to automatically translate from spoken language sentences to sign language videos. Historically the SLP task has been broken into two steps; Firstly, translating from a spoken language sentence to a gloss sequence and secondly, producing a sign language video given a sequence of glosses. In this paper we apply Natural Language Processing techniques to the first step of the SLP pipeline. We use language models such as BERT and Word2Vec to create better sentence level embeddings, and apply several tokenization techniques, demonstrating how these improve performance on the low resource translation task of Text to Gloss. We introduce Text to HamNoSys (T2H) translation, and show the advantages of using a phonetic representation for sign language translation rather than a sign level gloss representation. Furthermore, we use HamNoSys to extract the hand shape of a sign and use this as additional supervision during training, further increasing the performance on T2H. Assembling best practise, we achieve a BLEU-4 score of 26.99 on the MineDGS dataset and 25.09 on PHOENIX14T, two new state-of-the-art baselines.

Metrics

Details

Logo image

Usage Policy