Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2004.00849.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-13T20:29:13.300410Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T15:17:07.110860Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 546fb144-4606-4b62-9fba-c2f0036bb45a · inbound
GIT: A Generative Image-to-text Transformer for Vision and Language Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f8337825-ee2a-41f7-b9bf-622fe1072f80 · inbound
The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision) Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41625290-b7e6-4102-898a-aaf729961f51 · inbound
LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 197
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 17b0b189-121f-48ae-9ba4-4a09a08e92c7 · inbound
Agent AI: Surveying the Horizons of Multimodal Interaction Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 287
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b271ee9a-b7ac-4699-8be2-4a2b025c1f69 · inbound
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e62cc5bd-370a-4985-a8b9-d5f8d555a58b · inbound
Kinky vortons in the 2HDM Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05e19b0a-5d10-4397-bf3b-759b7091ba53 · inbound
HyFL-CLIP: Hyperbolic Fine-Tuning of CLIP for Robust Long-Context Understanding Pixel-BERT: Aligning Image Pixels with Text by Deep Multi-Modal Transformers
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.