REVIEW 8 cited by
BBC-Oxford British Sign Language Dataset
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this work, we introduce the BBC-Oxford British Sign Language (BOBSL) dataset, a large-scale video collection of British Sign Language (BSL). BOBSL is an extended and publicly released dataset based on the BSL-1K dataset introduced in previous work. We describe the motivation for the dataset, together with statistics and available annotations. We conduct experiments to provide baselines for the tasks of sign recognition, sign language alignment, and sign language translation. Finally, we describe several strengths and limitations of the data from the perspectives of machine learning and linguistics, note sources of bias present in the dataset, and discuss potential applications of BOBSL in the context of sign language technology. The dataset is available at https://www.robots.ox.ac.uk/~vgg/data/bobsl/.
Forward citations
Cited by 8 Pith papers
-
Isharah: A Large-Scale Multi-Scene Dataset for Continuous Sign Language Recognition
Isharah is a new 30,000-clip, multi-scene Saudi Sign Language dataset with gloss and translation annotations, plus signer-independent and unseen-sentence benchmarks for continuous sign language recognition and translation.
-
Semantic Hardness Is Not Visual Hardness: Sign-Aware Hard Negative Mining for Sign Language Retrieval
Hard negatives selected by visual confusability in sign embeddings, not linguistic similarity, substantially raise fine-grained sign-language retrieval accuracy without collapsing coarse performance.
-
SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning
Sparse keyframe-conditioned Conditional Flow Matching produces fluid, articulate 3D sign language motion across four languages while enabling precise Keyframe-to-Pose editing.
-
Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation
A dual visual encoder with contrastive visual-text pretraining achieves the best reported BLEU-4 score among gloss-free sign language translation methods on Phoenix-2014T.
-
iLSU-T: an Open Dataset for Uruguayan Sign Language Translation
iLSU-T is a 187-hour Uruguayan Sign Language video dataset with Spanish text, 18 interpreters, and first baseline translation results.
-
Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation
LLM-generated pseudo glosses, reordered via weak video supervision, enable sign language translation that rivals gloss-supervised models while needing only 30 gloss examples.
-
Sign Spotting Disambiguation using Large Language Models
LLM-based beam search disambiguation improves dictionary sign spotting WER from 47.2% to 44.4% on an internal BSL dataset.
-
Using Sign Language Production as Data Augmentation to enhance Sign Language Translation
Adding synthetic sign-language data produced by stitching, a GAN, or Gaussian splatting to the training set improves sign-language translation, with the largest gains for skeleton-pose models.
Discussion (0). Continue with ORCID to comment.