Pith. sign in

Paper Citation Record · LEDGER

Vision Transformer with Quadrangle Attention

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2303.15105.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.15105 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:33:54.289229Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T06:02:37.541296Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 846b657f-d763-4204-ad3c-939633a8cbf2 · inbound

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding cites this paper.

LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding Vision Transformer with Quadrangle Attention

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:02:37.545249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-23T06:01:00.775721Z digest=sha256:f8c657ba2b6f99634bcb80a69308f5ed531f9d0c7f50af85ff301873acf07e39

Observation 1ffd1b68-9ee3-4d03-8634-e20268023f95 · inbound

Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness cites this paper.

Facial Dynamics in Video: Instruction Tuning for Improved Facial Expression Perception and Contextual Awareness Vision Transformer with Quadrangle Attention

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:54.289229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:54.289229Z digest=sha256:b737b644cee9d9f88e8e83f40cf8e55796f8039b7fa4e22febe012e4d6433fa5

Observation 087b8d9f-8df3-488a-8814-c9a5f9fc6b88 · inbound

LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs cites this paper.

LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs Vision Transformer with Quadrangle Attention

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T22:24:35.992190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:24:35.992190Z digest=sha256:5a84f071bac99a5ceb385d58320379d1dec25bf35647fab14f52cb8994110b51

Observation fd7ae621-bd2f-40f0-8604-32794fccae95 · inbound

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams cites this paper.

ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams Vision Transformer with Quadrangle Attention

Reference 3

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T09:03:25.057325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T09:01:28.605086Z digest=sha256:e6acb691874544b60c13fa6fc70ba859911735fd2efccd5fdc453ec27a1bff7d

Observation 50cae3d4-20d9-4b12-aaec-50482c25540f · inbound

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment cites this paper.

X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment Vision Transformer with Quadrangle Attention

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-01T07:12:17.721095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:12:17.721095Z digest=sha256:50268d7aaaa46055186ed7b68ca4f1b413e402c71cd3dd7a6f21fc6030370e3d