REVIEW 12 cited by
BBC-Oxford British Sign Language Dataset
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this work, we introduce the BBC-Oxford British Sign Language (BOBSL) dataset, a large-scale video collection of British Sign Language (BSL). BOBSL is an extended and publicly released dataset based on the BSL-1K dataset introduced in previous work. We describe the motivation for the dataset, together with statistics and available annotations. We conduct experiments to provide baselines for the tasks of sign recognition, sign language alignment, and sign language translation. Finally, we describe several strengths and limitations of the data from the perspectives of machine learning and linguistics, note sources of bias present in the dataset, and discuss potential applications of BOBSL in the context of sign language technology. The dataset is available at https://www.robots.ox.ac.uk/~vgg/data/bobsl/.
Forward citations
Cited by 12 Pith papers
-
Isharah: A Large-Scale Multi-Scene Dataset for Continuous Sign Language Recognition
Isharah is a new 30,000-clip, multi-scene Saudi Sign Language dataset with gloss and translation annotations, plus signer-independent and unseen-sentence benchmarks for continuous sign language recognition and translation.
-
Semantic Hardness Is Not Visual Hardness: Sign-Aware Hard Negative Mining for Sign Language Retrieval
Hard negatives selected by visual confusability in sign embeddings, not linguistic similarity, substantially raise fine-grained sign-language retrieval accuracy without collapsing coarse performance.
-
SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning
Sparse keyframe-conditioned Conditional Flow Matching produces fluid, articulate 3D sign language motion across four languages while enabling precise Keyframe-to-Pose editing.
-
Contrastive Pretraining with Dual Visual Encoders for Gloss-Free Sign Language Translation
A dual visual encoder with contrastive visual-text pretraining achieves the best reported BLEU-4 score among gloss-free sign language translation methods on Phoenix-2014T.
-
iLSU-T: an Open Dataset for Uruguayan Sign Language Translation
iLSU-T is a 187-hour Uruguayan Sign Language video dataset with Spanish text, 18 interpreters, and first baseline translation results.
-
Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation
LLM-generated pseudo glosses, reordered via weak video supervision, enable sign language translation that rivals gloss-supervised models while needing only 30 gloss examples.
-
2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset
2M-BELEBELE is a new multilingual speech and ASL comprehension benchmark built from BELEBELE and FLEURS, with human recordings for 74 spoken languages and ASL video with glosses.
-
SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
A masked cluster-prediction transformer over four sign-language streams sets state-of-the-art results on multiple ASL translation and recognition benchmarks using only public pre-training data.
-
ODE-Based Transformer Decoders for Iterative Sign Language Translation
Replacing residual decoder updates with RK-2 and RK-4 numerical integration steps improves BLEU-4 on two sign language benchmarks against a matched iterative refinement baseline without adding decoder parameters.
-
Sign Spotting Disambiguation using Large Language Models
LLM-based beam search disambiguation improves dictionary sign spotting WER from 47.2% to 44.4% on an internal BSL dataset.
-
Using Sign Language Production as Data Augmentation to enhance Sign Language Translation
Adding synthetic sign-language data produced by stitching, a GAN, or Gaussian splatting to the training set improves sign-language translation, with the largest gains for skeleton-pose models.
-
Deaf in AI: AI language technologies and the erosion of linguistic rights
The paper argues that current AI sign language technologies, trained on interpreter-mediated data and framed as substitutes for interpreters, threaten deaf people's linguistic rights and require deaf-led design and go...
Discussion (0). Continue with ORCID to comment.