Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:1909.11740.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:12:08.524191Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T17:44:57.758784Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 629aeb46-409d-49bc-9c0a-18157c1721af · inbound
Unicoder-VL: A Universal Encoder for Vision and Language by Cross-modal Pre-training UNITER: UNiversal Image-TExt Representation Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5468228b-b5b5-4972-891c-b5da5a2b711a · inbound
Language Models are Few-Shot Learners UNITER: UNiversal Image-TExt Representation Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0903af5b-c090-436b-9f7f-fb006c8b7a09 · inbound
DetailCLIP: Injecting Image Details into CLIP's Feature Space UNITER: UNiversal Image-TExt Representation Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation be956813-70c8-4907-803b-68d82d645b5f · inbound
Everything is a Video: Unifying Modalities through Next-Frame Prediction UNITER: UNiversal Image-TExt Representation Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d2363b8-29c3-4e49-8ad5-570f7f323a1a · inbound
Approximate Fiber Product: A Preliminary Algebraic-Geometric Perspective on Multimodal Embedding Alignment UNITER: UNiversal Image-TExt Representation Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19852c09-dec0-4569-89f8-f635ef29bc6e · inbound
Exploring Large Vision-Language Models for Robust and Efficient Industrial Anomaly Detection UNITER: UNiversal Image-TExt Representation Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671a0c85-5eee-4c68-bcdb-5b3249cb232b · inbound
Survey on Question Answering over Visually Rich Documents: Methods, Challenges, and Trends UNITER: UNiversal Image-TExt Representation Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f8190f-6cb2-4dec-818f-fcf540a2c64d · inbound
LA-RCS: LLM-Agent-Based Robot Control System UNITER: UNiversal Image-TExt Representation Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cb7be6-a06c-463a-8359-e2eb0820f02c · inbound
Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation UNITER: UNiversal Image-TExt Representation Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fde2c972-902f-483c-8a35-38d25d02358d · inbound
AME: Aligned Manifold Entropy for Robust Vision-Language Distillation UNITER: UNiversal Image-TExt Representation Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e18712-028d-4062-86a6-8106ca9f484a · inbound
SEPS: Semantic-enhanced Patch Slimming Framework for fine-grained cross-modal alignment UNITER: UNiversal Image-TExt Representation Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45c0e4d9-8c5e-4266-a334-b1f1403cb4a7 · inbound
Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments UNITER: UNiversal Image-TExt Representation Learning
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cbc61fdb-edcc-469f-b999-c395dde7d123 · inbound
MASCOT: Model-Aware Submodular Coverage for Composite-Attribute Text-to-Image Retrieval UNITER: UNiversal Image-TExt Representation Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.