Pith. sign in

Paper Citation Record · LEDGER

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction

As of 22 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 2 inbound Pith citation observations for arXiv:2411.17835.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17835 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:33:45.080377Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:22:11.188805Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T11:23:45.109420Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4837e81d-6a60-413f-8fe6-24139b5f0700 · outbound

This paper cites Nougat: Neural Optical Understanding for Academic Documents.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction Nougat: Neural Optical Understanding for Academic Documents

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:44.936180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:44.936180Z digest=sha256:c2b24e2e4d86b20912c7013797b6406115a8990ff7dd184b2c5505783f20ce44

Observation ffcd77cc-cc39-43a9-bc5d-c278c9ab6d88 · outbound

This paper cites Multilingual Denoising Pre-training for Neural Machine Translation.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction Multilingual Denoising Pre-training for Neural Machine Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:44.942126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:44.942126Z digest=sha256:9826087051fc4ef870a67082570dcf88d3c70470fb6922c38aafd4f3cfb11c81

Observation 4a69a187-f985-4460-bb32-b18ced067217 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:44.947796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:44.947796Z digest=sha256:2e4ed40c8e6a297699970eb66f43d2250485a98882531a6074fde74fefc15aa6

Observation 2d0240c6-d716-49d5-9230-e82ba1663c57 · outbound

This paper cites LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.549158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.953568Z digest=sha256:266a406ff9b52ae626a618f16d103e605e385672a211eb42cd1f252c787eabf8

Observation d1492924-1050-43bb-9927-97c4bf5b8561 · outbound

This paper cites Donut: Document Understand- ing Transformer without OCR,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction Donut: Document Understand- ing Transformer without OCR,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.531499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.959243Z digest=sha256:1a94ff133ba4698bfbfdde71c278dc0ea9a9271c659751287926c8712d9816b4

Observation d50c3a25-08c4-4cba-a60f-a874ffcd8159 · outbound

This paper cites DS-YOLOv5: Deformable Single Shot YOLO for Document Parsing,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction DS-YOLOv5: Deformable Single Shot YOLO for Document Parsing,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.515172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.965533Z digest=sha256:f37ef9f318d58dc08ab57d31996e45de480ffe1997da6b661436f52d4873ed44

Observation a15524d1-9ee4-4cc5-99b0-ea0718fbeaa9 · outbound

This paper cites https://www.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction https://www

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.498479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.971711Z digest=sha256:32201101780d5406660052086c9edb8b07b19ba0ed3a1fd28867862dfa611bd1

Observation 2aaf8da5-9c7d-4ca8-9955-24b753052384 · outbound

This paper cites Khatt: An Open Arabic Hand- written Text Database,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction Khatt: An Open Arabic Hand- written Text Database,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.481255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.978031Z digest=sha256:8c8197a36c2326c03896ecee52f3ea5715441652313e5f94e7682893e3266407

Observation e1f9116e-b8ed-4cde-ba59-916ea9f69792 · outbound

This paper cites PubLayNet: Largest Dataset Ever for Document Layout Analysis,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction PubLayNet: Largest Dataset Ever for Document Layout Analysis,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.464742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.983514Z digest=sha256:f0f66c187f208b6b2757cb2e157e54fdc1763c56785d9a44b438eac0df2997f7

Observation 7949b6c3-66a3-48d7-adfd-25d7e54244cd · outbound

This paper cites VisionLAN: Visual Alignment Network for Scene Text Recognition,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction VisionLAN: Visual Alignment Network for Scene Text Recognition,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.447120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.988835Z digest=sha256:7307e0ffa9566a5e8c71e5ca07eae229307efe0b0c6d21eb6dd2b5303f717a56

Observation 08ea73c4-027d-46c0-91cf-6eae061111b0 · outbound

This paper cites TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction TrOCR: Transformer-based Optical Character Recognition with Pre-trained Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:44.993280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:44.993280Z digest=sha256:e21d7774add0857c6ef68235393c1c574cfc5072630d13c3816b6a8431b92074

Observation d1cd6c35-943e-40ec-9a70-70c31cec0432 · outbound

This paper cites riotu-lab/Aranizer-PBE-86k · Hug- ging Face,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction riotu-lab/Aranizer-PBE-86k · Hug- ging Face,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.428607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:44.998200Z digest=sha256:5b09ce00f331c2acd5704cc1a46dc2237d242d3114fba2d468c81c2aa71be062

Observation a89dca2e-939f-40e8-9106-0a630a558da8 · outbound

This paper cites MohamedRashad/arabic- img2md · Hugging Face,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction MohamedRashad/arabic- img2md · Hugging Face,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.411904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:45.003126Z digest=sha256:ad30acaa3c753481087c77f6ab4c5228bab35a3dac9a1c7c3424b382232372c5

Observation 939a491e-f688-44aa-905d-691e69b057e3 · outbound

This paper cites MohamedRashad/arabic-books · Hugging Face,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction MohamedRashad/arabic-books · Hugging Face,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.394303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:45.008716Z digest=sha256:7188b2f76b3c68d1da57c7c47c62eb861b3a1572d20429343a3bb6c093c78c30

Observation dd57a647-5892-484b-ac17-4d54589000e6 · outbound

This paper cites LayoutLM: Pre-training of Text and Layout for Document Image Understanding,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction LayoutLM: Pre-training of Text and Layout for Document Image Understanding,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.377256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:45.024606Z digest=sha256:792f3cff7002c1808f6fbe28611f8492bbb079d5b07d469299ae8b77fec39812

Observation 1d723a6b-530a-4d01-bb1c-f61f04c7225a · outbound

This paper cites BERTgrid: Contextualized Embedding for 2D Document Representation and Understanding.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction BERTgrid: Contextualized Embedding for 2D Document Representation and Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:45.029959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:45.029959Z digest=sha256:093fc9172305c1a6d70f6d086cd9852cb51b631aa4783e8407a9cd8f90cd6191

Observation 07845baa-8c01-4b49-b3d1-87da34a8aad4 · outbound

This paper cites Mathematical Formula Detection in Document Im- ages: A New Dataset and a New Approach,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction Mathematical Formula Detection in Document Im- ages: A New Dataset and a New Approach,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.359733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:45.035445Z digest=sha256:9d078c352e54dd0a8c92d884dafcd0e553bb7b95048d4ae741c1f265f80cfc83

Observation aef4e67c-eaf9-4782-9954-8ccbdc4f15d6 · outbound

This paper cites OmniParser: A Unified Framework for Text Spotting, Key Information Ex- traction and Table Recognition,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction OmniParser: A Unified Framework for Text Spotting, Key Information Ex- traction and Table Recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.340090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:45.041287Z digest=sha256:6a975a6d477f9c7898dac9c7f2cb74f05c5c97b81bcffe035f36f046f0a31e25

Observation 8c19a771-7244-4420-8605-7120f2b21645 · outbound

This paper cites General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:45.046203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:45.046203Z digest=sha256:916e9c01a54a341f41d245f14a1ac0f15172b1539262c75aa931226b942d70dc

Observation 8648eb88-2452-4ad1-bacc-481badefaa23 · outbound

This paper cites Focus Anywhere for Fine-grained Multi-page Document Understanding.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction Focus Anywhere for Fine-grained Multi-page Document Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:45.051137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:45.051137Z digest=sha256:e1395b210e3464f7d116916ea1842f0e472e593d3d8067fa9ba6916cf219286e

Observation 66bab495-d4ab-4ccd-b68f-a59783f6139f · outbound

This paper cites UReader: Universal OCR-free Visually-situated Language Understanding with Multimodal Large Language Model.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction UReader: Universal OCR-free Visually-situated Language Understanding with Multimodal Large Language Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:45.057217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:45.057217Z digest=sha256:a0d28b3c3d0b696c06e94ecc357a9ef81424b8f2c34c45773b86ee49a5d3a28d

Observation 1111587f-4e5a-4b54-a3ee-6f00151088b8 · outbound

This paper cites mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:45.064176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:45.064176Z digest=sha256:0d642bd80b86620fd391282b136dcb4508d2fdc123791a7377337e9e294f062f

Observation 441bfc6a-4ae6-47e0-bbbd-f84e4d373c22 · outbound

This paper cites MPLUG-PaperOwl: Scientific Dia- gram Analysis with the Multimodal Large Language Model,.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction MPLUG-PaperOwl: Scientific Dia- gram Analysis with the Multimodal Large Language Model,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:33:45.320579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T17:33:45.069611Z digest=sha256:453656fed287f677a210661c05b2045c557758a3f7bebedf3ae6d25a09288c9f

Observation 64294183-90b7-46e2-ba21-f46190dd377a · outbound

This paper cites mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding.

Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T17:33:45.080377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:33:45.080377Z digest=sha256:c3f9050a7f2a40bfe2726187327b32a5839ac9dbc1d3db375115f0b633e9dda8

Pith citing papers

Observation 2d4041ed-b40c-4e20-bcf7-c06581eec7b2 · inbound

SARD: A Large-Scale Synthetic Arabic OCR Dataset for Book-Style Text Recognition cites this paper.

SARD: A Large-Scale Synthetic Arabic OCR Dataset for Book-Style Text Recognition Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:11.188805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:11.188805Z digest=sha256:f78a103133c254e608304b9b87256ca35c3c7291b84a9691fd7980f3d11a0863

Observation ae8a7989-cf4c-4d87-a82f-68e64818835e · inbound

A-SEA3L-QA: A Fully Automated Self-Evolving, Adversarial Workflow for Arabic Long-Context Question-Answer Generation cites this paper.

A-SEA3L-QA: A Fully Automated Self-Evolving, Adversarial Workflow for Arabic Long-Context Question-Answer Generation Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T11:23:45.317541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T11:23:44.439355Z digest=sha256:8b3e8668a0e92bb352ab763769e278ad74a27f51278adec9c6262aa81fb0afc9