TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK	REMOVE
Sign Language Translation	CSL-Daily	MMTLB	BLEU-4	23.92	# 2
Sign Language Recognition	RWTH-PHOENIX-Weather 2014 T	MMTLB	Word Error Rate (WER)	22.45	# 6

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/a-simple-multi-modality-transfer-learning/sign-language-translation-on-csl-daily)](https://paperswithcode.com/sota/sign-language-translation-on-csl-daily?p=a-simple-multi-modality-transfer-learning)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/a-simple-multi-modality-transfer-learning/sign-language-recognition-on-rwth-phoenix-1)](https://paperswithcode.com/sota/sign-language-recognition-on-rwth-phoenix-1?p=a-simple-multi-modality-transfer-learning)`

A Simple Multi-Modality Transfer Learning Baseline for Sign Language Translation

CVPR 2022 · Yutong Chen, Fangyun Wei, Xiao Sun, Zhirong Wu, Stephen Lin ·

This paper proposes a simple transfer learning baseline for sign language translation. Existing sign language datasets (e.g. PHOENIX-2014T, CSL-Daily) contain only about 10K-20K pairs of sign videos, gloss annotations and texts, which are an order of magnitude smaller than typical parallel data for training spoken language translation models. Data is thus a bottleneck for training effective sign language translation models. To mitigate this problem, we propose to progressively pretrain the model from general-domain datasets that include a large amount of external supervision to within-domain datasets. Concretely, we pretrain the sign-to-gloss visual network on the general domain of human actions and the within-domain of a sign-to-gloss dataset, and pretrain the gloss-to-text translation network on the general domain of a multilingual corpus and the within-domain of a gloss-to-text corpus. The joint model is fine-tuned with an additional module named the visual-language mapper that connects the two networks. This simple baseline surpasses the previous state-of-the-art results on two sign language translation benchmarks, demonstrating the effectiveness of transfer learning. With its simplicity and strong performance, this approach can serve as a solid baseline for future research. Code and models are available at: https://github.com/FangyunWei/SLRT.

PDF Abstract CVPR 2022 PDF CVPR 2022 Abstract

Code

Add Remove Mark official

FangyunWei/SLRT official

205

FangyunWei/SLRT

205

rzhao-zhsq/cv-slt

edwardguil/MMTL

Tasks

Add Remove

Sign Language Recognition

Sign Language Translation

Transfer Learning

Translation

Datasets

Kinetics

Kinetics 400 WLASL RWTH-PHOENIX-Weather 2014 T CSL-Daily

Results from the Paper

Edit

Ranked #2 on Sign Language Translation on CSL-Daily

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Result	Benchmark
Sign Language Translation	CSL-Daily	MMTLB	BLEU-4	23.92	# 2		Compare
Sign Language Recognition	RWTH-PHOENIX-Weather 2014 T	MMTLB	Word Error Rate (WER)	22.45	# 6		Compare

Methods

Add Remove

No methods listed for this paper. Add relevant methods here

Edit Social Preview

A Simple Multi-Modality Transfer Learning Baseline for Sign Language Translation

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove