Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:17:42.908883Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2505.15576.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:17:42.908883Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ec091c7e-84e8-4160-8c8d-34a155b859f5 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Distill- ing knowledge from text-to-image generative models im- proves visio-linguistic reasoning in clip
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 92b560c6-34bd-476c-b3a9-d2147cb97ed8 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Cross-modal common representation learning by hybrid transfer network
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 02561694-c93f-40cf-8a5b-ae8bedb68762 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Structure-clip: Towards scene graph knowledge to en- hance multi-modal structured representations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e666f391-3fb3-411a-9ee2-73ff1fef7a39 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Adaptive prompt-based semantic embedding with inspire potential of implicit knowledge for cross-modal retrieval
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 85db73d7-d937-4086-91b1-bef36901bcea · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b3abd114-552d-46df-bc4c-e6de579628bf · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Microsoft coco: Com- mon objects in context
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02da6426-1133-42ef-abc0-1f88d3ffe75b · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Image segmentation using text and image prompts
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 72b8af50-9fce-4487-bdc8-3f60a2142d9e · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models AutoCLIP: Auto-tuning Zero-Shot Classifiers for Vision-Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6b11b91a-b0be-477e-896a-6acd0a4a5ea5 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Textattack: A frame- work for adversarial attacks, data augmentation, and ad- versarial training in nlp
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5d4cd703-11a3-4ee7-9db8-1a5e6f0ca69d · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Valse: A task-independent benchmark for vision and language models centered on linguistic phe- nomena
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0cbf1f8a-a4bb-4e26-8d84-6aa6fd31065c · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Learning transferable visual models from nat- ural language supervision
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cb4a62ac-5a42-473d-9441-1154ee5e00a6 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 84b12e22-3366-4d4b-8043-9a845ee50fa2 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Image as a foreign language: Beit pretrain- ing for vision and vision-language tasks
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a201851e-5b78-4fa4-8326-dfef90c5a8e3 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Groupvit: Semantic segmentation emerges from text supervision
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f15431b8-7660-43e9-abf1-13e05e3da31f · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models CREPE: open-domain ques- tion answering with false presuppositions
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dff164eb-6c6a-4acf-8cb4-bee651993ea6 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models When and why vision-language models behave like bags-of-words, and what to do about it?
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e203c0d3-7549-434a-9fba-2910c9f836f2 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 330b64f6-7ef8-41e1-ab72-c3da8578474a · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2954376-bc86-4cf7-a5db-f3cedd1d4238 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Sugar- crepe: fixing hackable benchmarks for vision-language compositionality
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8a557498-957f-4da6-bc18-93cd963a7d59 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Smith, Yejin Choi, and Hannaneh Ha- jishirzi
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7ccaf003-a3a7-45dc-8058-ea0b90f595cf · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Chils: Zero- shot image classification with hierarchical label sets
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a189848c-4dbe-48c3-b195-b6e52c4a40f6 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Winoground: Probing vision and language models for visio-linguistic compositionality
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c418a5b3-7238-4f24-87f9-f94242e88f99 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models spacy 2: Natural lan- guage understanding with bloom embeddings, convolu- tional neural networks and incremental parsing.To appear,
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6ef37cb2-d20d-4616-82b2-d37e257a9bfd · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models CyCLIP: Cyclic contrastive language-image pretraining
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8e556a4b-9df3-43d6-a5e9-a246ffaee2c1 · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Going beyond nouns with vision & language models using synthetic data
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 00bc3172-cf4b-41a1-a036-4a5a0eb713bf · outbound
Visual Perturbation and Adaptive Hard Negative Contrastive Learning for Compositional Reasoning in Vision-Language Models Blip: Bootstrapping language-image pre- training for unified vision-language understanding and generation
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.