NaviDC-OCR is a 1.2B parameter document parser that combines deformation-aware training, adaptive layout sampling, and content-structure decoupled learning to reach state-of-the-art scores on OmniDocBench v1.6, Wild-OmniDocBench, and PureDocBench.
Forcennet: Foreground-centric network for document image rectification
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
NaviDC-OCR is a 1.2B parameter document parser that combines deformation-aware training, adaptive layout sampling, and content-structure decoupled learning to reach state-of-the-art scores on OmniDocBench v1.6, Wild-OmniDocBench, and PureDocBench.