Pith. sign in

REVIEW 1 cited by

Word-level Persian Lipreading Dataset

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.04068 v1 pith:DUYIIDB6 submitted 2023-04-08 cs.CV

classification cs.CV
keywords datasetword-leveladvanceslip-readinglipreadingmethodpersianused
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Lip-reading has made impressive progress in recent years, driven by advances in deep learning. Nonetheless, the prerequisite such advances is a suitable dataset. This paper provides a new in-the-wild dataset for Persian word-level lipreading containing 244,000 videos from approximately 1,800 speakers. We evaluated the state-of-the-art method in this field and used a novel approach for word-level lip-reading. In this method, we used the AV-HuBERT model for feature extraction and obtained significantly better performance on our dataset.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Integrating Persian Lip Reading in Surena-V Humanoid Robot for Human-Robot Interaction

    cs.CV 2025-01 conditional novelty 5.0 of 10

    A custom 7-word Persian lip-reading dataset is used to train an LSTM that reports 89% accuracy and is deployed on the Surena-V humanoid robot for real-time command recognition.

Pith tools