Logo image
SIGNET: Motion-Level Knowledge Transfer for Cross-Language Sign Language Translation
Conference proceeding   Peer reviewed

SIGNET: Motion-Level Knowledge Transfer for Cross-Language Sign Language Translation

Computer Vision – ECCV 2026, Vol.In Press(In Press)
The 19th European Conference on Computer Vision (ECCV 2026) (Malmö, Sweden, 08/09/2026–12/09/2026)
2026

Abstract

Sign Language Translation Cross-lingual Transfer
Sign language translation (SLT) remains challenging due to its high spatio-temporal complexity, long sequences, and the need to model multiple articulators without relying on gloss annotations. Existing approaches are typically tailored to individual datasets or languages and struggle to scale, while overlooking the relationships between sign languages that could inform more effective cross-lingual transfer. We present SIGNET, a framework that enables motion-level knowledge transfer for cross-language sign language translation. Our key insight is that, although sign languages differ in grammar and lexicon, pretrained models capture motion-level visual patterns that can be reused across datasets and languages. SIGNET integrates multiple pretrained sign language backbones through an attention-based, hand-prior aggregation mechanism that guides a gated fusion network in dynamically selecting the most relevant experts. Comprehensive experiments on four benchmarks (How2Sign, Phoenix14T, CSL-Daily, and MeineDGS) demonstrate state-of-the-art translation performance, and SIGNET also surpasses prior methods on WLASL for sign language recognition.
pdf
signet9.31 MBDownloadView
Author's Accepted Manuscript Embargo until publication date CC BY V4.0
url
https://eccv.ecva.net/View
Event Website Conference website

Metrics

1 File views/ downloads
1 Record Views

Details

Logo image

Usage Policy