Pith. sign in

REVIEW 1 cited by

Orientation-Independent Chinese Text Recognition in Scene Images

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.01081 v1 pith:VK4ZXMTG submitted 2023-09-03 cs.CV

classification cs.CV
keywords textrecognitionchineseimagescontentinformationmethodorientation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Scene text recognition (STR) has attracted much attention due to its broad applications. The previous works pay more attention to dealing with the recognition of Latin text images with complex backgrounds by introducing language models or other auxiliary networks. Different from Latin texts, many vertical Chinese texts exist in natural scenes, which brings difficulties to current state-of-the-art STR methods. In this paper, we take the first attempt to extract orientation-independent visual features by disentangling content and orientation information of text images, thus recognizing both horizontal and vertical texts robustly in natural scenes. Specifically, we introduce a Character Image Reconstruction Network (CIRN) to recover corresponding printed character images with disentangled content and orientation information. We conduct experiments on a scene dataset for benchmarking Chinese text recognition, and the results demonstrate that the proposed method can indeed improve performance through disentangling content and orientation information. To further validate the effectiveness of our method, we additionally collect a Vertical Chinese Text Recognition (VCTR) dataset. The experimental results show that the proposed method achieves 45.63% improvement on VCTR when introducing CIRN to the baseline model.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Zero-Shot Chinese Character Recognition with Hierarchical Multi-Granularity Image-Text Aligning

    cs.CV 2025-05 conditional novelty 6.0 of 10

    A multi-granularity contrastive framework, Hi-GITA, aligns Chinese character images with stroke, radical, and structure sequences and improves zero-shot recognition accuracy by large margins on several benchmarks.

Pith tools