Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T22:26:00.495364Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2602.16918.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T22:26:00.495364Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4b508cec-537d-41dd-8a86-794fba215770 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Symbolic Discovery of Optimization Algorithms
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 861bfd8e-0120-4558-bb20-82a55f20ac88 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0fc1917-471d-412f-96ca-d39244f146d2 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Deconstructing Denoising Diffusion Models for Self-Supervised Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cc71124-0f00-41db-97c1-36ec5518baa5 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Meta CLIP 2: A Worldwide Scaling Recipe
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1b7200c-2e27-405f-97c9-d69808dc5286 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Vision Transformers Need Registers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7e408fc-18e5-4f02-b520-14d1d29dcb93 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Improving CLIP Training with Language Rewrites
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b594624c-d56b-4114-b2bf-95085c3e3573 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data SimCSE: Simple Contrastive Learning of Sentence Embeddings
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5676fdd-2261-48a0-a90d-caa06f22b53e · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data The Llama 3 Herd of Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699fa30a-6d82-4273-ade3-3ced50b5e6ed · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data LoRA: Low-Rank Adaptation of Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 504cca27-18fb-4424-a97a-6feb7e8edbbc · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data The Kinetics Human Action Video Dataset
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01c48d91-fe66-4049-8a96-885d6737b0ab · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95340127-4d20-444d-9118-4f8ce75c5177 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Exploring the Limits of Weakly Supervised Pretraining
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6fa7d81-3589-4fe0-a5cf-6eeadcac4002 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data DINOv2: Learning Robust Visual Features without Supervision
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 696c96a6-75cc-40bf-b530-93bb54c9f38a · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data LAION-5B: An open large-scale dataset for training next generation image-text models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dbfe0ee-661d-4a8a-9b15-41a9782eab60 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Adafactor: Adaptive Learning Rates with Sublinear Memory Cost
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4cc25f9-992f-4ba6-9599-89d16fb03e88 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data DINOv3
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b55b2978-a860-4702-a0b5-39b37c7870e3 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e0582b5-d228-4703-97d3-dc1607de51c5 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Pooling And Attention: What Are Effective Designs For LLM-Based Embedding Models?
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbc80d26-e53f-4a8a-aa16-25f31f6abe37 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data LLaMA: Open and Efficient Foundation Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db195f75-9445-4edb-8d16-bd3c9152f423 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a03cbcfc-fb3c-4eed-b62c-572431ca23dc · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Representation Learning with Contrastive Predictive Coding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0966b5a3-9fd4-4390-aa95-823c9abd1975 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Demystifying CLIP Data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a317e63-dab9-481c-8136-3d3b0d18d5e5 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Msr-vtt: A large video description dataset for bridging video and language.2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pages 5288–5296,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b902530-fcc4-4c89-8711-732db56544e2 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Deep Residual Learning for Image Recognition
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4a64b7c-4877-4d2d-9b1b-2d934359eb4f · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5522b7c-7839-4be4-8dce-a74a5544fd79 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Hong-You Chen, Zhengfeng Lai, Haotian Zhang, Xinze Wang, Marcin Eichner, Keen You, Meng Cao, Bowen Zhang, Yinfei Yang, and Zhe Gan
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e7d7fe6-1d9e-4aab-9d44-46a5177bd1ac · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data LocCa: Visual Pretraining with Location-aware Captioners
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77d81c88-65f5-4b27-a17a-116f7dfc034f · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Language Models are Few-Shot Learners
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eebaee77-cf9e-4e9c-b509-705411c184e5 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Quo vadis, action recognition? a new model and the kinetics dataset
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc6b390b-0fa7-4974-ba07-ac700ebf8f6d · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Improving fine-grained understanding in image-text pre-training
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e131ad-0a4e-46a9-9c23-526340582a32 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Poggio, and Thomas Serre
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a5f23d3-fca3-48fc-bbbe-f97a1817f0e2 · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data Token Merging: Your ViT But Faster
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ca2d4af-bf68-4a7a-b8bf-f51fae9b828e · outbound
Xray-Visual Models: Scaling Vision models on Industry Scale Data V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.