Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2203.07190.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:27.682010Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T09:50:00.732703Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1dc487ff-e6c1-4cc3-99e6-44a31205782b · inbound
Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 57ded672-e299-492f-b903-43903541fc5f · inbound
Advancing Myopia To Holism: Fully Contrastive Language-Image Pre-training CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15b8295a-9ff7-47af-bd7f-be6e5e19c222 · inbound
Human Action CLIPs: Detecting AI-generated Human Motion CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea5dd575-ea5b-44c1-a9c8-32ed09ae7ccc · inbound
Nearly Solved? Robust Deepfake Detection Requires More than Visual Forensics CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be67acef-79f1-4123-8589-88189b65c44e · inbound
DiffCLIP: Few-shot Language-driven Multimodal Classifier CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33f3847e-a73c-4d95-b38b-ffb1e24975a7 · inbound
How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adad59bb-9f9f-406c-9c95-8386147827c9 · inbound
Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cacea6f-70fa-4dcc-b2fe-fbd51df36e67 · inbound
Multi-Branch Collaborative Learning Network for Video Quality Assessment in Industrial Video Search CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbcb3662-ca8e-4610-946d-517c1f28888d · inbound
Logits DeConfusion with CLIP for Few-Shot Learning CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4bf2ebf-9f7f-49f2-8a34-66c2ce042822 · inbound
(Almost) Free Modality Stitching of Foundation Models CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b20cc5d5-8999-419f-b50d-de5067b26b04 · inbound
Sparse and Dense Retrievers Learn Better Together: Joint Sparse-Dense Optimization for Text-Image Retrieval CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56c330f5-d58b-42c0-a4f4-b76187d16700 · inbound
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 849ce55d-f81e-411d-a613-d97ed986faaf · inbound
WRF4CIR: Weight-Regularized Fine-Tuning Network for Composed Image Retrieval CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.