REVIEW 1 cited by
Word-level Persian Lipreading Dataset
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Lip-reading has made impressive progress in recent years, driven by advances in deep learning. Nonetheless, the prerequisite such advances is a suitable dataset. This paper provides a new in-the-wild dataset for Persian word-level lipreading containing 244,000 videos from approximately 1,800 speakers. We evaluated the state-of-the-art method in this field and used a novel approach for word-level lip-reading. In this method, we used the AV-HuBERT model for feature extraction and obtained significantly better performance on our dataset.
Forward citations
Cited by 1 Pith paper
-
Integrating Persian Lip Reading in Surena-V Humanoid Robot for Human-Robot Interaction
A custom 7-word Persian lip-reading dataset is used to train an LSTM that reports 89% accuracy and is deployed on the Surena-V humanoid robot for real-time command recognition.
Discussion (0). Continue with ORCID to comment.