A new deep architecture learns character prototypes from line transcriptions to produce scalable paleographic measurements on historical scripts.
Rethinking Text Line Recognition Models, April 2021
3 Pith papers cite this work. Polarity classification is still indexing.
representative citing papers
Nougat applies a visual transformer to convert academic PDFs into markup language while accurately handling mathematical content on a new scientific document dataset.
A new dataset and open-source OCR pipeline transcribes medieval English legal manuscripts at up to 88% word accuracy using CNN+LSTM and language model correction.
citing papers explorer
-
Leveraging Morphology for Historical Script Metrological Analysis
A new deep architecture learns character prototypes from line transcriptions to produce scalable paleographic measurements on historical scripts.
-
Nougat: Neural Optical Understanding for Academic Documents
Nougat applies a visual transformer to convert academic PDFs into markup language while accurately handling mathematical content on a new scientific document dataset.
-
Democratizing the medieval English legal tradition
A new dataset and open-source OCR pipeline transcribes medieval English legal manuscripts at up to 88% word accuracy using CNN+LSTM and language model correction.