Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T11:24:16.101773Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2412.15523.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T11:24:16.101773Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9ed2e4b1-4105-4225-b75c-1c21a627f890 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting Context Perception Parallel Decoder for Scene Text Recognition
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e2161a-9d07-4967-993d-74c303dda212 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting InstructDiffusion: A Generalist Modeling Interface for Vision Tasks
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cbb866a-e573-4159-aeec-6e46a2ad7c9a · outbound
InstructOCR: Instruction Boosting Scene Text Spotting OCR-free Document Understanding Transformer
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c1003f8-8de2-47c5-8774-3b8ac79bff0f · outbound
InstructOCR: Instruction Boosting Scene Text Spotting Segment Anything
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55638be9-ac1e-4fa9-b141-dde60e63ca80 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcfc6fd0-f36c-4591-a414-ec1fbebe0002 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting SPTS v2: Single-Point Scene Text Spotting
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ee1b0aa8-656e-4374-b20a-32bffae4b191 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting KOSMOS-2.5: A Multimodal Literate Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08b6e3a0-caf1-43b4-82ba-31959c3f891f · outbound
InstructOCR: Instruction Boosting Scene Text Spotting In 2017 14th IAPR international con- ference on document analysis and recognition (ICDAR), vol- ume 1, 1454–1459
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4dac2de8-9d73-4393-ae63-54579e5874c6 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting UPOCR: Towards Unified Pixel-Level OCR Interface
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c6d0c6cb-54eb-499f-adce-c11763938a7a · outbound
InstructOCR: Instruction Boosting Scene Text Spotting UReader: Universal OCR-free Visually-situated Language Understanding with Multimodal Large Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f113a607-6fe4-4838-bc18-46fc8ff6dbc9 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting In 12th international conference on document analysis and recognition, 1484–1493
Reference 2013
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8f093086-639d-4626-b5c1-3082f28b7b1b · outbound
InstructOCR: Instruction Boosting Scene Text Spotting In 13th international conference on document analysis and recognition, 1156–1160
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 02c21729-2244-4df4-b7dc-09dfb6d184a5 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting In 2017 14th IAPR international conference on document anal- ysis and recognition (ICDAR), volume 1, 935–942
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d915012a-b1e2-44be-92fa-3a29e0daee91 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5d838e3-9ea4-449b-9aac-ab771b998daa · outbound
InstructOCR: Instruction Boosting Scene Text Spotting An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dca4bb9-7760-4cb3-8a55-75c53a62262a · outbound
InstructOCR: Instruction Boosting Scene Text Spotting Pix2seq: A Language Modeling Framework for Object Detection
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7507f045-ef53-4b51-9d07-75b3bad64f5a · outbound
InstructOCR: Instruction Boosting Scene Text Spotting GIT: A Generative Image-to-text Transformer for Vision and Language
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c522d627-393f-42a0-9000-e8c0a3b9b3d5 · outbound
InstructOCR: Instruction Boosting Scene Text Spotting Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c07e379-aa0b-41b6-927f-a3aafe304c2c · outbound
InstructOCR: Instruction Boosting Scene Text Spotting OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.