Pith. sign in

Paper Citation Record · LEDGER

TextSquare: Scaling up Text-Centric Visual Instruction Tuning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2404.12803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.12803 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:46:40.923386Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:28:33.630729Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 822ce715-504b-4053-8584-7d2b2a31b041 · inbound

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data cites this paper.

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T10:43:50.171896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:43:50.171896Z digest=sha256:bada8a0654c309f16a0627017a05d3f1e5c7fa72a7f534c53e1d677230b5cc0a

Observation 01e56a5f-a498-4082-a7f6-1e59a7314bce · inbound

Knowledge-Embedded and Hypernetwork-Guided Few-Shot Substation Meter Defect Image Generation Method cites this paper.

Knowledge-Embedded and Hypernetwork-Guided Few-Shot Substation Meter Defect Image Generation Method TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T10:43:42.108402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:43:42.108402Z digest=sha256:172d27099229b017c549330f41f3f8a367a401551f519be4e6f01c03719809e4

Observation bd92c6a1-d0bf-4604-ae03-a3cafb9d2fb9 · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 208

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:32.105894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:32.105894Z digest=sha256:5dd1cf8f5ad0ed5d7e9496328625e92351c28c62c99c7d43e2a8dadb5f5780c2

Observation a5ca0822-c50f-46b7-a980-f88c23f889ed · inbound

Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction cites this paper.

Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:03:12.253874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T20:01:57.029143Z digest=sha256:49e8b3f553909f7b1cc08516c3b38a10be9812938e2922f451d2bc8e8dbecec2

Observation ac61a334-dc66-49ad-9551-bf85733bfbcd · inbound

Adaptive Slicing-Assisted Hyper Inference for Enhanced Small Object Detection in High-Resolution Imagery cites this paper.

Adaptive Slicing-Assisted Hyper Inference for Enhanced Small Object Detection in High-Resolution Imagery TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:04.359252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T02:15:57.389751Z digest=sha256:f5d55d85fa12f75a353994b731a35c9f5bb6370d48b70d32d74977eda851a0cc

Observation ba04cc73-5bf7-4da4-a5ea-f83bd0c69520 · inbound

Feature Perturbation Pool-based Fusion Network for Unified Multi-Class Industrial Defect Detection cites this paper.

Feature Perturbation Pool-based Fusion Network for Unified Multi-Class Industrial Defect Detection TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:29:21.636628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T03:28:11.747826Z digest=sha256:98de389b5a7498055b560df63c26a88a5d806cb61b8eec9c2f5acf24bffac9d0

Observation 2a1a8163-c6b6-490e-9b68-4382978557ce · inbound

Lightweight Real-Time Rendering Parameter Optimization via XGBoost-Driven Lookup Tables cites this paper.

Lightweight Real-Time Rendering Parameter Optimization via XGBoost-Driven Lookup Tables TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:21:14.833499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T17:25:18.561710Z digest=sha256:35ee3d2ab669d237f501d013a04378ae4c78cfdc47dcda5ac74c200b25e12c54

Observation 740e9900-d310-4cf1-a8fc-bfd9796e6b38 · inbound

Image Classification via Random Dilated Convolution with Multi-Branch Feature Extraction and Context Excitation cites this paper.

Image Classification via Random Dilated Convolution with Multi-Branch Feature Extraction and Context Excitation TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:22:02.532748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T17:07:55.177114Z digest=sha256:409db9bb83174246add62cfbfdb1a9c1d873d6066d120165d06cd9433d18fe42

Observation 0179feda-077d-4f05-ab5b-c4aafccb2388 · inbound

Gait Recognition via Deep Residual Networks and Multi-Branch Feature Fusion cites this paper.

Gait Recognition via Deep Residual Networks and Multi-Branch Feature Fusion TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:27.643408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T08:33:09.327681Z digest=sha256:2459938d8e08d0c3ccfbe681fab3d66d5444473f1890df835fb32bf317da6061

Observation eae58fe0-2409-416e-bfaa-b0718705b91f · inbound

Multi-Branch Non-Homogeneous Image Dehazing via Concentration Partitioning and Image Fusion cites this paper.

Multi-Branch Non-Homogeneous Image Dehazing via Concentration Partitioning and Image Fusion TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:26:08.654646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T20:03:36.551943Z digest=sha256:abc195df2f8670a45b91e814b2dc824f0e9171d27241293e27981e7a7b7dc91f

Observation 7553f3ca-bec9-4de8-b43b-0d39a4bcbc72 · inbound

Local Spatiotemporal Convolutional Network for Robust Gait Recognition cites this paper.

Local Spatiotemporal Convolutional Network for Robust Gait Recognition TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:08:29.328878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T02:07:25.917495Z digest=sha256:16268330e311f517923e368e62a623a763d5869f2bd521dbce894252ef3aaa68

Observation d567303d-d6d9-4289-8df1-e5879cf84718 · inbound

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting cites this paper.

Do You Need Text Rectification? Soft Attention Mask Embedding for Rectification-Free Scene Text Spotting TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:33:14.581568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T11:28:51.711892Z digest=sha256:30b1169e8b6acce14534f3a1d0532b856c912a0e26ca744524570fff32922a62

Observation 234cd75a-2aed-40cf-a176-fde70cb44795 · inbound

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation cites this paper.

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:14:18.736535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T06:13:11.013395Z digest=sha256:53375b686571e13b261ec9330377d696dc3fef2135fc96bfd5c70c94d2a083f2

Observation 116d7829-bd22-4b15-bb9e-f3a093a389c9 · inbound

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation cites this paper.

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:05:28.684332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T07:01:26.911110Z digest=sha256:d7c4eed005a9f49b42d4d2dc594c8aa31fdfdea96eb6bfe8e0498b441ebee00e

Observation eb0e8ae8-ccf5-477f-aade-773b29521671 · inbound

DAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal Reasoning cites this paper.

DAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal Reasoning TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:04:21.356682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T06:01:20.803078Z digest=sha256:204959ec08aaf855868fbad6cefe551f7b747a1abc578e283f68f8563ac4c2f8

Observation 2aef371b-8e21-4678-a89b-5b1558d8e906 · inbound

ProWAFT: A ROMA-LPD Instance for Workload-Aware and Dynamic Fault Tolerance in FPGA-Based CNN Accelerators cites this paper.

ProWAFT: A ROMA-LPD Instance for Workload-Aware and Dynamic Fault Tolerance in FPGA-Based CNN Accelerators TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:28:33.632467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T15:23:48.748933Z digest=sha256:4fcced9f0c0b75a9f96ce56ae3ff88c18e6f37d7a18a25494245f8945fca4cbe

Observation 013a1907-7e2a-43a6-b44a-61f42d20e6a2 · inbound

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth cites this paper.

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 160

Resolution
unresolved
no resolver link, observed 2026-08-02T14:55:37.403124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:55:37.403124Z digest=sha256:9420fb4af38a800e2a8bc2b87abef344bc72250058e742b3befa91e93a61c4ab

Observation 01c50331-21cb-40f6-a6ec-3fa45d58515e · inbound

DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards cites this paper.

DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards TextSquare: Scaling up Text-Centric Visual Instruction Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T00:46:40.923386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:46:40.923386Z digest=sha256:2e0a8a637db773817fa15c1adafc60df2838d26e6bc7727b1beabaecd3c46cd9